Jobs Companies Vannevar Site Reliability Engineer

À propos de ce poste Site Reliability Engineer chez Vannevar

Vannevar · Sur site · San Diego, California

Vannevar is a defense technology company building AI to deter our adversaries. In the 21st century, conflict moves at algorithmic speed and foresight equals firepower. Our agentic AI is purpose-built to compete with China—from cross-Strait conflict to gray zone coercion. Trained on the most mission-relevant datasets in defense, our technology models adversary behavior, simulates campaigns, and recommends the best course of action to decision makers. Our AI systems are some of the most trusted in the industry and actively used on the front lines of the Indo-Pacific to keep the peace and save lives.

Exceptional technology starts with exceptional people. Vannevar is a small agile team combining world-class engineers with veteran strategists who bring deep expertise in defense and tradecraft. We’re building a company defined by mission impact, user empathy, and disciplined growth. In just three years, we grew from $3M to $80M in ARR, achieved early profitability, and reached unicorn status—proving that disruption doesn’t require an ego, and staying power doesn’t mean standing still.

About the role

We are looking for an Site Reliability Engineer to own the reliability, health, and deployment automation of the platform at Vannevar Labs. In this role you'll be the person watching the system's pulse — monitoring dashboards, catching health issues before they become incidents, and owning the debugging process from first alert to resolution. Your decisions today will have a large impact on the company's future.
We believe that simple systems are easier to understand, maintain, and scale. You will be making trade-offs as you work to ensure that our systems are prepared to operate reliably in high-side environments at scale. A strong sense of judgment matters here: knowing when to dig deeper into a problem yourself and when to pull in the right people to escalate. Clear, calm communication — during an incident and in day-to-day work — is a must.

What you'll do

  • Monitor dashboards and system telemetry to detect health issues, performance degradation, and reliability risks — often before anyone else notices them.
  • Own the debugging and incident response process end to end, exercising good judgment about when to investigate more deeply and when to escalate.
  • Build logging, monitoring, and observability tooling to visualize the state of the platform and continuously mature our SRE practices.
  • Develop, maintain, and be responsible for overall platform health, scaling, and capacity planning.
  • Understand and help improve the deployment process, and automate build & deployment pipelines.
  • Identify bottlenecks in engineering workflows and drive improvements that make the whole team faster and more reliable.
  • Develop self-service tools and automation to improve engineering efficiency.
  • Play a critical part in implementing a secure, robust, high-availability delivery pipeline.
  • Communicate system status, trade-offs, and post-incident learnings clearly with teammates and stakeholders.

Qualifications

  • 5+ years of experience in SRE, DevOps, or software engineering.
  • Hands-on experience monitoring production systems and responding to incidents — comfortable owning a debugging process and making the call on when to dig in versus escalate.
  • Excellent communication skills, especially the ability to stay clear and organized while troubleshooting live issues.
  • Experience with the PLG stack, Datadog, or other enterprise monitoring/observability tools.
  • Experience participating in an on-call rotation and running or contributing to post-mortems.
  • Knowledge of AWS cloud technologies.
  • Familiarity with infrastructure-as-code technologies such as Terraform and Pulumi.
  • Experience with Python, Bash, or other scripting languages.
  • Experience working in an agile scrum environment, with the ability to work independently.
  • Able to quickly learn new and existing technologies.
  • Strong attention to detail and analytical capabilities.
  • Willingness and ability to work on-site in San Diego, CA.
  • U.S. Citizenship status is required, as this position requires the ability to access U.S.-only data systems and export-controlled data.
  • TS/SCI Clearance required.

Additional Qualifications (Nice to haves)

  • Experience defining and tracking SLOs/SLIs and error budgets.
  • Experience crafting CI/CD processes and automation.
  • Proficient with containerization technologies like Docker.
  • Experience working in AWS GovCloud.
  • Experience with modern web services architectures.
  • Experience with relational database systems, including SQL and relational design.
  • Experience working with Elasticsearch/OpenSearch.
  • Strong collaboration and negotiation skills, with the ability to work on cross-functional projects with internal partner engineering teams.

What we offer

We’re proud to offer competitive benefits that support our employees. Some key highlights of our benefits package include:
  • Health, dental, and vision insurance
  • 100% remote first culture. You can work from anywhere in the US and all full time employees have WeWork access
  • Unlimited PTO including competitive vacation and holiday schedules
  • Lifestyle stipends - Monthly mental health, wellness & fitness stipend, in-home office setup stipend and family planning assistance
  • Salary top-up during military reserve duty
  • Fully paid parental leave
  • Child and pet care reimbursement during travel

 

Vannevar is an equal opportunity employer, and qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender perception or identity, national origin, age, marital status, protected veteran status, or disability status.
 
We encourage candidates from all backgrounds to apply, even if you don't feel like you're a perfect fit. If you're passionate about contributing to our mission, we'd love to hear from you!
 
IMPORTANT NOTICE
We are committed to protecting the privacy of all applicants. Official emails from the company will come from an @vannevarlabs.com domain. Under no circumstances will a legitimate representative from our company contact you to request passwords, financial information, or other sensitive personal data. Please be vigilant of potential scams.
Prêt à postuler chez Vannevar ?
Postuler chez Vannevar

À propos de Vannevar

Vannevar builds next generation defense software for the public servants keeping our country safe.  As a team, we exist because we believe in public service, and we think that our democracy and government improve only if we put serious, collective effort into improving them, including the technology our government uses.   
 
IMPORTANT NOTICE
We are committed to protecting the privacy of all applicants. Official emails from the company will come from an @vannevarlabs.com domain. Under no circumstances will a legitimate representative from our company contact you to request passwords, financial information, or other sensitive personal data. Please be vigilant of potential scams.

 

Voir tous les emplois chez Vannevar →

Emplois similaires

Shieldai
Senior Staff Lead Site Reliability Engineer (R5803)
Shieldai
⚡ Postuler tôt San Diego, California Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 1 sem.
Skydio
Senior Hardware Test and Reliability Engineer
Skydio
⚡ Postuler tôt San Mateo, California, United... Sur site $170,000–$230,000
● Nouveau 👁 Vu ✓ Postulé il y a 1 h
Rain
Site Reliability Engineer
Rain
⚡ Postuler tôt Monde entier $170,000–$225,000
● Nouveau 👁 Vu ✓ Postulé il y a 1 h
Ironclad
Staff Site Reliability Engineer
Ironclad
⚡ Postuler tôt San Francisco Hybride $220,000–$235,000
● Nouveau 👁 Vu ✓ Postulé il y a 2 h
Antares
Reliability Engineer II-Senior
Antares
⚡ Postuler tôt Los Angeles Sur site $112,000–$185,000
● Nouveau 👁 Vu ✓ Postulé il y a 2 h
Shieldai
Senior Reliability and Maintainability Engineer (R5301)
Shieldai
⚡ Postuler tôt Dallas, Texas Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 2 h
EarnIn
Staff Site Reliability Engineer (SRE)
EarnIn
⚡ Postuler tôt Mountain View, US Hybride
● Nouveau 👁 Vu ✓ Postulé il y a 4 h
Anduril Industries
Technical Site Reliability Engineer
Anduril Industries
⚡ Postuler tôt Abu Dhabi, United Arab Emirate... Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 4 h
Roku
Senior Software Engineer,  SRE
Roku
⚡ Postuler tôt Bengaluru, India Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 4 h

Inscrivez-vous pour des suggestions adaptées aux emplois que vous ouvrez et aux recherches que vous enregistrez.

Plus d’emplois chez Vannevar

Voir tous les emplois chez Vannevar →

Postuler maintenant
🤖

Doucement — un instant

JobsRadar a été conçu pour de vraies personnes qui traversent une période difficile dans leur recherche d’emploi — pas pour des requêtes automatisées. Vous cliquez beaucoup trop vite et vous êtes maintenant temporairement bloqué.

Revenez plus tard. Si vous cherchez réellement un emploi, nous sommes de votre côté — agissez simplement comme un être humain.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Prenez une longueur d’avance dans votre recherche d’emploi.

Rejoignez notre canal Telegram pour ce qui vous aide à décrocher le poste — références salariales, le pouls hebdomadaire du marché et les annonces de nouveautés. Pas de spam, que du signal.

Rejoindre le canal — c’est gratuit