Jobs Companies Forward Networks Site Reliability Engineer

À propos de ce poste Site Reliability Engineer chez Forward Networks

Forward Networks · Sur site · Santa Clara, CA

Forward is transforming how the world’s most complex networks are managed and secured. Founded in 2013 by four Stanford Ph.D.s, we built the industry’s first network digital twin — a mathematically precise model of the production network that gives IT teams unmatched visibility, verification, and agility across every major cloud and vendor environment.

Our customers include global leaders such as Goldman Sachs, PayPal, S&P Global, IBM, and Dell, as well as fast-growing enterprises and government agencies. According to IDC, Forward customers realize an average of $14.2 million in annual benefits through improved efficiency and security.

Backed by world-class investors including Andreessen Horowitz, Goldman Sachs, MSD Partners, and Threshold Ventures, Forward offers a people-centric, innovative culture where brilliant minds are shaping the future of network reliability, security, and AI-ready operations.

Forward is looking for a Site Reliability Engineer

About the Role This is not a "keep the lights on" SRE role. As our first or early SRE hire you will be building the reliability engineering function at Forward — defining how we think about availability, observability, incident response, and operational excellence across a complex, distributed SaaS platform. You will work closely with engineering, infrastructure, and product to ensure our platform meets the reliability bar our enterprise customers demand.

If you thrive in environments where you're handed a problem rather than a playbook this role is for you.

What You'll Own

  • Define and drive SRE practices from the ground up — SLOs, SLIs, error budgets, and the frameworks the engineering org will actually use
  • Drive the reliability and operational excellence of the Forward SaaS platform
  • Build and maintain observability infrastructure — logging, metrics, tracing, and alerting — so the team always knows what's happening before customers do
  • Lead incident response: on-call rotations, runbooks, post-mortems, and the follow-through to make sure the same incident doesn't happen twice
  • Partner with engineering teams to embed reliability thinking into the SDLC — capacity planning, load testing, chaos engineering, and production readiness reviews
  • Help define and build the SRE team as the company scales — this is a foundational hire with a path to leadership

What We're Looking For

  • 6+ years of experience in site reliability engineering, DevOps, or infrastructure engineering in a SaaS or cloud environment
  • Proven experience building or significantly maturing an SRE function — not just operating within one someone else built
  • Strong fundamentals in networking — TCP/IP, DNS, routing, switching, firewalls, and load balancing. Experience with network management or observability platforms is a significant plus
  • Hands-on experience with Kubernetes and container orchestration in production environments
  • Deep proficiency with observability tooling — Prometheus, Grafana, Datadog, Splunk, or similar
  • Strong scripting and automation skills in Python, Bash, or similar
  • Experience with cloud platforms — AWS, GCP, or Azure — including infrastructure as code (Terraform, Ansible, or equivalent)
  • Track record of owning and improving incident response processes including blameless post-mortems and SLO-driven reliability improvements
  • Ability to communicate clearly with both engineering teams and non-technical stakeholders — you can explain an outage to a customer-facing team without jargon and explain an SLO to an executive without losing them

Nice to Have

  • Experience supporting enterprise or federal government customers with high availability requirements
  • Experience in a foundational or early SRE hire capacity at a growth stage company

What This Role Is Not

  • A pure ops or NOC role — you are building and engineering, not just monitoring
  • A siloed function — you will be deeply embedded with product and engineering teams
  • A ticket-taker — you will be proactively identifying and solving reliability problems before they become incidents

    The base pay range for this role is between $230,000 and $250,000. Base pay will depend on your skills, qualifications, experience, and location
Prêt à postuler chez Forward Networks ?
Postuler chez Forward Networks

Comment se compare ce salaire pour SRE

Ce poste paie $240,000/yrau-dessus de la fourchette habituelle pour les postes SRE.

$87,000 la médiane $175,000 $246,220

Fourchette typique $134,225–$208,625/yr, à partir de 469 annonces SRE comparables sur JobsRadar (rémunération annualisée en USD). Voir les aperçus de salaire pour SRE →

Emplois similaires

PH
Site Reliability Engineering — AI Accelerator Infrastructure
Phizenix
⚡ Postuler tôt Santa Clara, CA (3 Days Onsite... Hybride $155,000–$235,000
● Nouveau 👁 Vu ✓ Postulé il y a 1 sem.
PH
Director, Site Reliability Engineering — AI Accelerator Infrastructure
Phizenix
⚡ Postuler tôt Santa Clara, CA (3 Days Onsite... Hybride $195,000–$285,000
● Nouveau 👁 Vu ✓ Postulé il y a 2 sem.
Sustainable Talent
Site Reliability Engineer
Sustainable Talent
⚡ Postuler tôt Santa Clara, CA Sur site $135,200–$135,200
● Nouveau 👁 Vu ✓ Postulé il y a 2 sem.
Gatik AI
Senior/Staff Site Reliability Engineer
Gatik AI
⚡ Postuler tôt Santa Clara, CA Sur site $180,000–$180,000
● Nouveau 👁 Vu ✓ Postulé il y a 2 sem.
Arkestro
Senior Site Reliability Engineer
Arkestro
⚡ Postuler tôt United States Sur site $160,000–$180,000
● Nouveau 👁 Vu ✓ Postulé il y a 37 min
Verisign
Site Reliability Engineer - IBM AIX
Verisign
⚡ Postuler tôt Reston,Virginia,United States Hybride $135,800–$183,800
● Nouveau 👁 Vu ✓ Postulé il y a 55 min
Verisign
SRE - Linux
Verisign
⚡ Postuler tôt Reston,Virginia,United States Hybride $135,800–$183,800
● Nouveau 👁 Vu ✓ Postulé il y a 55 min
Verisign
Site Reliability Engineer
Verisign
⚡ Postuler tôt Reston,Virginia,United States Hybride $135,800–$183,800
● Nouveau 👁 Vu ✓ Postulé il y a 55 min
Okta
Staff Site Reliability Engineer (FedRAMP)
Okta
⚡ Postuler tôt Bellevue, Washington; Chicago,... Sur site $194,000–$267,000
● Nouveau 👁 Vu ✓ Postulé il y a 1 h

Inscrivez-vous pour des suggestions adaptées aux emplois que vous ouvrez et aux recherches que vous enregistrez.

Plus d’emplois chez Forward Networks

Voir tous les emplois chez Forward Networks →

Postuler maintenant
🤖

Doucement — un instant

JobsRadar a été conçu pour de vraies personnes qui traversent une période difficile dans leur recherche d’emploi — pas pour des requêtes automatisées. Vous cliquez beaucoup trop vite et vous êtes maintenant temporairement bloqué.

Revenez plus tard. Si vous cherchez réellement un emploi, nous sommes de votre côté — agissez simplement comme un être humain.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Prenez une longueur d’avance dans votre recherche d’emploi.

Rejoignez notre canal Telegram pour ce qui vous aide à décrocher le poste — références salariales, le pouls hebdomadaire du marché et les annonces de nouveautés. Pas de spam, que du signal.

Rejoindre le canal — c’est gratuit