Jobs Companies Forward Networks Site Reliability Engineer

Sobre este puesto de Site Reliability Engineer en Forward Networks

Forward Networks · Presencial · Santa Clara, CA

Forward was founded in 2013 by four Stanford Ph.D.s, building the industry's first network digital twin: a mathematically accurate model of the production network. It's the foundation for autonomous networking, giving engineers and AI agents the ability to know the impact of every change before it touches production. That founding instinct still defines how we work. We're accurate and evidence-driven, relentless about clarity, and we'd rather be certain than comfortable, building a groundbreaking platform that transforms how teams run and secure networks across every major cloud and vendor environment.

Global leaders like Goldman Sachs, PayPal, S&P Global, IBM, and Dell trust Forward, alongside fast-growing enterprises and government agencies, realizing an average of $14.2 million in annual benefits, according to IDC. Backed by top-tier investors, including A. Capital, Andreessen Horowitz, Goldman Sachs, MSD Partners, Omega Venture Partners, Section 32, and Threshold Ventures, and headquartered in Santa Clara, we're most proud of our team: curious people who'd rather build what doesn't exist than accept how things have always been done.

Forward is looking for a Site Reliability Engineer

About the Role This is not a "keep the lights on" SRE role. As our first or early SRE hire you will be building the reliability engineering function at Forward — defining how we think about availability, observability, incident response, and operational excellence across a complex, distributed SaaS platform. You will work closely with engineering, infrastructure, and product to ensure our platform meets the reliability bar our enterprise customers demand.

If you thrive in environments where you're handed a problem rather than a playbook this role is for you.

What You'll Own

  • Define and drive SRE practices from the ground up — SLOs, SLIs, error budgets, and the frameworks the engineering org will actually use
  • Drive the reliability and operational excellence of the Forward SaaS platform
  • Build and maintain observability infrastructure — logging, metrics, tracing, and alerting — so the team always knows what's happening before customers do
  • Lead incident response: on-call rotations, runbooks, post-mortems, and the follow-through to make sure the same incident doesn't happen twice
  • Partner with engineering teams to embed reliability thinking into the SDLC — capacity planning, load testing, chaos engineering, and production readiness reviews
  • Help define and build the SRE team as the company scales — this is a foundational hire with a path to leadership

What We're Looking For

  • 6+ years of experience in site reliability engineering, DevOps, or infrastructure engineering in a SaaS or cloud environment
  • Proven experience building or significantly maturing an SRE function — not just operating within one someone else built
  • Strong fundamentals in networking — TCP/IP, DNS, routing, switching, firewalls, and load balancing. Experience with network management or observability platforms is a significant plus
  • Hands-on experience with Kubernetes and container orchestration in production environments
  • Deep proficiency with observability tooling — Prometheus, Grafana, Datadog, Splunk, or similar
  • Strong scripting and automation skills in Python, Bash, or similar
  • Experience with cloud platforms — AWS, GCP, or Azure — including infrastructure as code (Terraform, Ansible, or equivalent)
  • Track record of owning and improving incident response processes including blameless post-mortems and SLO-driven reliability improvements
  • Ability to communicate clearly with both engineering teams and non-technical stakeholders — you can explain an outage to a customer-facing team without jargon and explain an SLO to an executive without losing them

Nice to Have

  • Experience supporting enterprise or federal government customers with high availability requirements
  • Experience in a foundational or early SRE hire capacity at a growth stage company

What This Role Is Not

  • A pure ops or NOC role — you are building and engineering, not just monitoring
  • A siloed function — you will be deeply embedded with product and engineering teams
  • A ticket-taker — you will be proactively identifying and solving reliability problems before they become incidents

Why Forward

  • You'll be building something from scratch at a company with real enterprise traction and world-class investors behind it
  • Our customers include some of the most complex network environments on the planet — the reliability bar is high and the work is genuinely interesting
  • People-centric culture built by Stanford Ph.D.s who care deeply about doing things the right way
  • Competitive compensation, equity, and the opportunity to grow into a leadership role as the SRE function scales

The base pay range for this role is between $230,000 and $250,000. Base pay will depend on your skills, qualifications, experience, and location

¿Listo para postularte en Forward Networks?
Postúlate en Forward Networks

Cómo se compara este salario de SRE

Este puesto paga $240,000/yrpor encima de el rango típico para los puestos de SRE.

$127,350 la mediana de $136,475 $182,950

Rango típico $131,250–$158,500/yr, a partir de 8 ofertas comparables de SRE en JobsRadar (salario anualizado en USD). Ver datos salariales de SRE →

Empleos similares

NVIDIA
Senior Site Reliability Engineer - HPC
NVIDIA
⚡ Postúlate pronto US, CA, Santa Clara Híbrido
● Nuevo 👁 Visto ✓ Postulado hace 2d
NVIDIA
Senior Site Reliability Engineer - Storage
NVIDIA
⚡ Postúlate pronto US, CA, Santa Clara Presencial
● Nuevo 👁 Visto ✓ Postulado hace 5d
NVIDIA
Principal Software Engineer, At-Scale Reliability and Fleet Intelligence — CSP Engagements
NVIDIA
⚡ Postúlate pronto US, CA, Santa Clara Presencial
● Nuevo 👁 Visto ✓ Postulado hace 1sem
Applied Materials
Product Quality & Reliability Engineer IV
Applied Materials
⚡ Postúlate pronto Santa Clara,CA Presencial $133,500–$183,500
● Nuevo 👁 Visto ✓ Postulado hace 1sem
Applied Materials
DfSafety and Reliability Engineer
Applied Materials
⚡ Postúlate pronto Santa Clara,CA Presencial $100,000–$136,500
● Nuevo 👁 Visto ✓ Postulado hace 2sem
Applied Materials
Product Quality & Reliability Engineer IV (E4)
Applied Materials
⚡ Postúlate pronto Austin,TX Presencial $116,000–$159,500
● Nuevo 👁 Visto ✓ Postulado hace 1 mes
Sustainable Talent
Site Reliability Engineer
Sustainable Talent
⚡ Postúlate pronto Santa Clara, CA Presencial $135,200–$135,200
● Nuevo 👁 Visto ✓ Postulado hace 2 meses
Gatik AI
Senior/Staff Site Reliability Engineer
Gatik AI
⚡ Postúlate pronto Santa Clara, CA Presencial $180,000–$180,000
● Nuevo 👁 Visto ✓ Postulado hace 2 meses
RELX
Site Reliability Engineering Lead
RELX
⚡ Postúlate pronto Florida Presencial $118,300–$219,800
● Nuevo 👁 Visto ✓ Postulado hace 1h

Regístrate para recibir sugerencias adaptadas a los empleos que abres y las búsquedas que guardas.

Más empleos en Forward Networks

Ver todos los empleos en Forward Networks →

Postúlate ahora
🤖

Un momento — para

JobsRadar se creó para personas reales que están pasando un mal momento en su búsqueda de empleo — no para solicitudes automatizadas. Estás haciendo clic demasiado rápido y ahora estás bloqueado temporalmente.

Vuelve más tarde. Si de verdad estás buscando empleo, cuentas con nosotros — solo compórtate como una persona.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Toma ventaja en tu búsqueda de empleo.

Únete a nuestro canal de Telegram para lo que te ayuda a conseguir el puesto — referencias salariales, el pulso semanal del mercado y avisos de nuevas funciones. Sin spam, solo señal.

Únete al canal — es gratis