Jobs Companies Specter Platform Site Reliability Engineer

Über diese Platform Site Reliability Engineer Stelle bei Specter

Specter · Vor Ort · San Francisco

Company Background

Specter's mission is to help automate the physical world.

Today, we build video sensors with state-of-the-art AI agents that answer any question, anywhere in their environments. Our systems can automatically detect and reason about any physical activity captured on camera, from security incidents (e.g. perimeter intrusion, theft, LPR), to safety monitoring (e.g. PPE detection, injured people), to operational efficiency (e.g. material tracking, congestion monitoring). We offer both long range wireless (1km range) and wired sensor variants to suit any deployment.

Soon, we will build robots, trained on top of the data we collect, to take action in these environments as well.

Our co-founders Xerxes and Philip are passionate about empowering our partners in the fast approaching world of physical AI and robotics. We are a small, fast growing team who hail from Anduril, Tesla, Uber, and the U.S. Special Forces.


The Role

We’re hiring a Platform Site Reliability Engineer to own the operational health, reliability, and scalability of the cloud platform behind our connected sensor fleet.

This is a high-ownership role at the intersection of site reliability and platform engineering. You’ll operate and improve our Kubernetes-based infrastructure, manage cloud resources through Terraform, strengthen observability and incident response, and build the systems that allow our engineering teams to deploy safely and move quickly.

You’ll work primarily across our AWS infrastructure and Kubernetes environments while partnering with application, AI, embedded systems, and fleet teams. You’ll help resolve production issues when they occur—and then improve the platform so they are less likely to happen again.


Responsibilities

Reactive — Triage & Recovery

  • Debug production issues across Kubernetes clusters, Linux systems, AWS infrastructure, networking, and application workloads.

  • Lead incidents from detection through recovery, coordinating across teams when failures span multiple parts of the system.

  • Participate in an on-call rotation and follow incidents through to durable fixes.

Systems Builder — Close the Loop

  • Build, operate, and improve our Kubernetes platform and the AWS infrastructure supporting it.

  • Manage production infrastructure with Terraform, including reusable modules, automated validation, and safe change workflows.

  • Reduce operational toil through automation while improving deployment tooling, CI/CD, and developer workflows.

Observability Owner — Platform Visibility

  • Design and improve observability into system health, functionality, and performance through logging, metrics, tracing, dashboards, and alerting across Kubernetes workloads and AWS infrastructure.

  • Define meaningful service-level indicators and objectives, and close telemetry gaps before they become incidents.

  • Develop runbooks, incident-response procedures, post-incident reviews, and operational readiness standards.


Qualifications

  • Strong Linux systems knowledge and experience diagnosing production systems.

  • Hands-on experience operating Kubernetes in production, including networking, storage, resource management, upgrades, and troubleshooting.

  • Strong experience using Terraform to manage production cloud infrastructure.

  • Experience with AWS, including IAM, networking, compute, storage, and EKS.

  • Solid networking fundamentals, including DNS, load balancing, firewalls, VPNs, subnets, and routing.

  • Experience building operational tooling and automation using Python, Go, Bash, or a similar language.

  • Strong ownership during incidents and the ability to turn ambiguous failures into lasting improvements.


Nice to Have

  • Experience supporting connected devices, edge computing, or on-premises infrastructure alongside cloud systems.

  • Experience with CI/CD, GitOps, or Kubernetes multi-cluster environments.

  • Familiarity with cloud and Kubernetes security practices.

  • Experience reading firmware logs or low-level Rust or C code when debugging across the edge-to-cloud boundary.

  • Experience with NixOS and managing Nix-based development infrastructure.

Bereit, sich bei Specter zu bewerben?
Bei Specter bewerben

Ähnliche Jobs

Anthropic
Staff+ Site Reliability Engineer, Safeguards ML Infra
Anthropic
⚡ Früh bewerben Remote-Friendly (Travel-Requir... · standortgebunden $405,000–$485,000
● Neu 👁 Gesehen ✓ Beworben vor 4 Std.
Anthropic
Staff Software Engineer, AI Reliability
Anthropic
⚡ Früh bewerben San Francisco, CA | New York C... Vor Ort $325,000–$485,000
● Neu 👁 Gesehen ✓ Beworben vor 4 Std.
Waabi
Vehicle Reliability Engineer
Waabi
⚡ Früh bewerben Dallas, TX Hybrid
● Neu 👁 Gesehen ✓ Beworben vor 18 Std.
Pinterest
Sr. Site Reliability Engineer, tvScientific
Pinterest
⚡ Früh bewerben San Francisco, CA, US; Remote,... · standortgebunden $139,764–$287,749
● Neu 👁 Gesehen ✓ Beworben vor 19 Std.
Pinterest
Site Reliability Engineer II, tvScientific
Pinterest
⚡ Früh bewerben San Francisco, CA, US; Remote,... · standortgebunden $114,297–$235,319
● Neu 👁 Gesehen ✓ Beworben vor 19 Std.
Ōura
Hardware Reliability Engineer
Ōura
⚡ Früh bewerben Hybrid - San Francisco, Califo... Hybrid
● Neu 👁 Gesehen ✓ Beworben vor 23 Std.
Sierra
Software Engineer, Site Reliability (SRE)
Sierra
⚡ Früh bewerben San Francisco, CA Vor Ort $230,000–$390,000
● Neu 👁 Gesehen ✓ Beworben vor 1 Tg.
Inferact
Member of Technical Staff, Site Reliability Engineer
Inferact
⚡ Früh bewerben San Francisco Vor Ort $200,000–$400,000
● Neu 👁 Gesehen ✓ Beworben vor 1 Tg.
Okta
Manager, Site Reliability Engineering
Okta
⚡ Früh bewerben Bellevue, Washington; Chicago,... Vor Ort $204,000–$306,000
● Neu 👁 Gesehen ✓ Beworben vor 1 Tg.

Registrieren für Vorschläge, die auf die von Ihnen geöffneten Jobs und gespeicherten Suchen zugeschnitten sind.

Mehr Jobs bei Specter

Alle Jobs bei Specter ansehen →

Jetzt bewerben
🤖

Moment — langsam

JobsRadar wurde für echte Menschen gebaut, die eine schwere Zeit bei der Jobsuche haben — nicht für automatisierte Anfragen. Sie klicken viel zu schnell und sind jetzt vorübergehend blockiert.

Kommen Sie später wieder. Wenn Sie wirklich auf Jobsuche sind, stehen wir hinter Ihnen — verhalten Sie sich einfach wie ein Mensch.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Verschaffe dir einen Vorsprung bei der Jobsuche.

Tritt unserem Telegram-Kanal bei für das, was dir hilft, die Stelle zu bekommen — Gehaltsbenchmarks, den wöchentlichen Marktpuls und neue Feature-Drops. Kein Spam, nur Signal.

Dem Kanal beitreten — kostenlos