Jobs Companies Navan Senior Site Reliability Engineer

Sobre este puesto de Senior Site Reliability Engineer en Navan

Navan · Presencial · London, UK

At Navan, “It’s all about the user. All of them.” We’re passionate about providing a seamless one-stop experience for business travelers, no matter how they travel, where they stay, or where they’re going.

We are constantly striving to make the most reliable and scalable systems possible to ensure that our services are available to our travelers when they need it most. With our exponential growth, we have many exciting challenges ahead and we’re looking for a passionate Senior Site Reliability Engineer to join our team in London. As a Senior SRE you will design and develop tooling, automation and infrastructure services that power the Navan services, used by thousands of travelers on a daily basis.  You will work closely with development teams, release and productivity teams and security teams to identify customer needs and build innovative solutions to solve them. 

You will work across a vast array of systems and technologies, aiming to build an autonomous, monitored, fault-tolerant infrastructure that is optimized for both simplicity and uptime. You will collaborate with the backend and frontend engineering teams to ensure that product solutions are scalable, efficient, and reliable. You will design infrastructure to support our massive growth and work with the team to maintain the highest level of service.

 

What You'll Do:

  • Building a fast moving, high growth service. Navan is revolutionizing travel and expense services for the enterprise, and the product is evolving quickly. You are comfortable in a startup environment, enjoy seeing the product take shape, and have strong ownership of the success of your services.
  • Designing, implementing and operating cloud infrastructure. You’re a fit for us if you think in terms of infrastructure as code, deployment pipelines, and building the guardrails to make going fast also going safely.
  • Identifying reliability anti-patterns and solving them systemically. You dive deep into the data to evaluate the health of your systems, and you use it to improve visibility and reliability across the fleet of services.
  • Finding and automating the toil out of our processes. You’d prefer to automate it entirely, or build a tool to empower your users rather than be the gatekeeper to the tool.
  • Leveraging AI tools and platforms in your daily work to achieve autonomous operations, reduce toil, and improve system observability.
  • Defining and driving the adoption of system reliability standards, including formalizing SLO/SLI frameworks, observability standards, and blameless post-mortem practices across multiple engineering teams.
  • Driving the adoption of AI-assisted developer tools and platforms to increase engineering productivity, enforce code quality standards, and enable real-time architectural validation.

 

What We’re Looking For:

  • 5+ years of progressive experience as a Senior SRE or DevOps Lead (or equivalent role)
  • 2+ years of experience in working on a production, 24x7 product environment
  • Passionate about solving problems and learning new tools and technologies
  • Excellent communication skills working with stakeholders and domain experts across the company to design solutions to user problems
  • Thrive in a fast-paced environment 
  • Demonstrated experience mentoring and leading junior and mid-level engineers, and acting as a technical owner for cross-functional infrastructure projects.
  • Operate with a strong sense of ownership demonstrated through shipping production-quality code and infrastructure equipped with testing, monitoring and documentation
  • Hands-on operational experience with Java based applications and services including JVM profiling and performance tuning (python, Node.js and Go are a plus)
  • Hands-on experience building and operating distributed systems in a public cloud environment (preferably AWS), using CI/CD to deploy, manage and operate production systems, focusing on tooling and automation using tools such as maven and Jenkins.
  • Hands-on experience with microservice architecture and related reliability and resiliency patterns such as throttling, queueing, and retries
  • Hands-on experience with writing Infrastructure as Code in Terraform or Cloudformation or similar tools
  • A passion for automating away everything, using scripting languages such as python, bash groovy (we prefer lazy engineers)
  • Built, using, and automating monitoring systems such as NewRelic, DataDog, SignalFX, Kibana,
  • Hands-on experience deploying, operating, and monitoring production-grade AI/ML microservices (e.g., RAG pipelines, agentic systems) on cloud platforms like AWS Fargate/ECS.
  • Experience leveraging AI/LLM platforms (e.g., Gemini, Braintrust) and managing their secrets and infrastructure using Infrastructure as Code (Terraform) and AWS SSM.
  • Demonstrated ability to integrate AI-specific telemetry and advanced observability practices to enable predictive insights and systemic root-cause analysis.
¿Listo para postularte en Navan?
Postúlate en Navan

Sobre Navan

ABOUT TRIPACTIONS

TripActions is the fastest-growing corporate travel platform disrupting a $1.5T industry and shaping the future of business travel.

TripActions is a story of inspiration born of frustration. Road warriors and co-founders Ariel Cohen and Ilan Twig believed that companies deserved a travel solution that takes the pain out of work trips –– so that their travelers can focus on being productive and meeting in-person, not wasting valuable time booking travel. So in 2015, they created TripActions. Since then, we’ve been a mission to power the face-to-face, in-person connections that move people, ideas and businesses forward.

TripActions’ platform offers a vast selection of inventory that travelers can choose from, a personalized, intuitive user interface driven by machine learning, and 24/7 proactive real human, customer support. Companies enjoy complete travel program visibility, over 30% cost savings on average and seamless integrations with their HR and expense systems.

Globally, TripActions has grown to over 600 employees across 7 offices in 4 countries. We support over 1,500 customers, with innovative brands like Lyft, Dropbox, Sara Lee Frozen Bakery, Allbirds, Robinhood and the ACLU relying on TripActions for their business travel needs. As one of Silicon Valley’s newest “unicorns”, TripActions has a valuation north of $1B and a total of $232M in funding. We’ve recently received $154M in our Series C funding round –– led by new investor Andreessen Horowitz, with participation from repeat investors Lightspeed Venture Partners, Zeev Ventures and SGVC.

TripActions was recently recognized as one of Fast Company’s Most Innovative Companies for 2019, #12 in LinkedIn’s Top Startups 2018 and #3 in the U.S. for Happiest Employees by Comparably.

We’re redefining what it means to travel for work. Come help us build the future of business travel.

Ver todos los empleos en Navan →

Empleos similares

Rightmove Careers
Engineering Manager (Site Reliability)
Rightmove Careers
⚡ Postúlate pronto London, UK Presencial
● Nuevo 👁 Visto ✓ Postulado hace 1d
Ping Identity
Staff Site Reliability Engineer
Ping Identity
⚡ Postúlate pronto UK - Remote · restringido por ubicación
● Nuevo 👁 Visto ✓ Postulado hace 2d
Anthropic
Staff Software Engineer, AI Reliability Engineering
Anthropic
⚡ Postúlate pronto London, UK Presencial £325,000–£390,000
● Nuevo 👁 Visto ✓ Postulado hace 1sem
Graphcore
Senior Systems Engineer – Performance & Reliability
Graphcore
⚡ Postúlate pronto London, UK Presencial
● Nuevo 👁 Visto ✓ Postulado hace 1sem
Graphcore
Senior Systems Engineer – Performance & Reliability
Graphcore
⚡ Postúlate pronto Bristol, UK Presencial
● Nuevo 👁 Visto ✓ Postulado hace 1sem
Graphcore
Senior Systems Engineer – Performance & Reliability (Analysis)
Graphcore
⚡ Postúlate pronto Bristol, UK Presencial
● Nuevo 👁 Visto ✓ Postulado hace 2sem
Miro
Senior Network Site Reliability Engineer
Miro
⚡ Postúlate pronto Copenhagen, DK; London, UK; Re... · restringido por ubicación
● Nuevo 👁 Visto ✓ Postulado hace 2sem
GoCardless
Site Reliability Engineer
GoCardless
⚡ Postúlate pronto London, UK Presencial $270,400–$270,400
● Nuevo 👁 Visto ✓ Postulado hace 2sem
ComplyAdvantage
Site Reliability Engineering Manager (Data Infra)
ComplyAdvantage
⚡ Postúlate pronto Lisbon, Portugal Presencial
● Nuevo 👁 Visto ✓ Postulado hace 4sem

Regístrate para recibir sugerencias adaptadas a los empleos que abres y las búsquedas que guardas.

Más empleos en Navan

Ver todos los empleos en Navan →

Postúlate ahora
🤖

Un momento — para

JobsRadar se creó para personas reales que están pasando un mal momento en su búsqueda de empleo — no para solicitudes automatizadas. Estás haciendo clic demasiado rápido y ahora estás bloqueado temporalmente.

Vuelve más tarde. Si de verdad estás buscando empleo, cuentas con nosotros — solo compórtate como una persona.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Toma ventaja en tu búsqueda de empleo.

Únete a nuestro canal de Telegram para lo que te ayuda a conseguir el puesto — referencias salariales, el pulso semanal del mercado y avisos de nuevas funciones. Sin spam, solo señal.

Únete al canal — es gratis