Jobs Companies Fluidstack Customer Reliability Engineer

Über diese Customer Reliability Engineer Stelle bei Fluidstack

Fluidstack · Vor Ort · San Francisco, CA

About Fluidstack

We exist to make humanity more free. For most of human history, you farmed or you starved. Technology gave people more time for the things they wanted to do, instead of things they had to do. Powerful AI will be the biggest lever for human choice we've ever built - but only if models are aligned with what humanity actually wants. There are groups building AI who don't share these goals. Whoever deploys frontier compute infrastructure fastest will decide whether AI expands human freedom or shrinks it.

We're singularly focused on delivering 10 to 100s of GWs of compute faster than anyone else, rethinking every layer of the stack. We acquire power, design and build data centers, and operate them - with teams spanning hardware and software. Speed and scale are our key differentiators. Come be a part of building civilization-scale infrastructure for AI.


We hire people who care deeply about this problem space. If that is you, please apply!

How We Operate

  • Extreme ownership. Full autonomy. Own things end to end often taking on scope outside your core role without being asked to get things done.

  • Velocity. We drive everything forward as fast as possible.

  • First principles. Challenge every assumption. Zero analogy thinking, no egos, the best idea wins.

  • Love of the game. The frontier of AI is the most interesting problem of our time. We put in long hours at high intensity to push the frontier forward.

The Data Center Operations Team

Examples of key problems the team is working on

  • Operate at the scale of a nation, not a building. The fleet you run will draw more power than some countries, on the way to 10s to 100s of GWs.

  • Fly the plane while it's being built. Sites come online in pieces, and you keep the live ones running flawlessly while construction continues around them.

  • Write the playbook, don't inherit it. No prior operations org has run at this speed and scale, so the standards you set become the standard.

Role Scope

  • Own reliability for named customer workloads: their clusters, their SLAs, their escalations.

  • Debug across the full stack, hardware to fabric to scheduler, when a training run degrades.

  • Run customer-facing incident communication with technical depth and no spin.

  • Turn recurring customer pain into engineering fixes with the production teams.

What We're Looking For

  • The below is a starting point. We always make space for exceptional people, so if you don't fit this role exactly, tell us where you would.

  • You've supported large-scale compute customers (HPC, cloud, or AI labs) at a technical level.

  • You debug distributed systems methodically across layers you don't own.

  • You've written incident updates customers trusted more after reading.

  • You push internal teams to fix causes, not symptoms, and follow up until they do.

  • Bonus: GPU training workloads. InfiniBand or RoCE. Slurm or Kubernetes. NCCL debugging.

We are committed to pay equity and transparency.

Fluidstack is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Fluidstack will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.

You will receive a confirmation email once your application has successfully been accepted. If there is an error with your submission and you did not receive a confirmation email, please email careers@fluidstack.io with your resume/CV, the role you've applied for, and the date you submitted your application-- someone from our recruiting team will be in touch.

Bereit, sich bei Fluidstack zu bewerben?
Bei Fluidstack bewerben

Wie sich dieses Gehalt für SRE vergleicht

Diese Stelle zahlt $238,500/yrüber der üblichen Spanne für SRE Stellen.

$171,540 dem Median $217,000 $286,095

Übliche Spanne $198,050–$236,075/yr, aus 18 vergleichbaren SRE Anzeigen auf JobsRadar (Vergütung auf USD hochgerechnet). Gehaltseinblicke für SRE ansehen →

Ähnliche Jobs

Fluidstack
Principal Operations Engineer, Reliability
Fluidstack
⚡ Früh bewerben Austin, TX Vor Ort $242,000–$278,000
● Neu 👁 Gesehen ✓ Beworben vor 4 Std.
Hadrian Automation
Site Reliability Engineer, Client Platform
Hadrian Automation
⚡ Früh bewerben Los Angeles, CA Vor Ort $164,000–$270,000
● Neu 👁 Gesehen ✓ Beworben vor 11 Std.
Hadrian Automation
Site Reliability Engineer, Robotics
Hadrian Automation
⚡ Früh bewerben Los Angeles, CA Vor Ort $164,000–$270,000
● Neu 👁 Gesehen ✓ Beworben vor 11 Std.
Fluidstack
Site Reliability Engineer, Compute
Fluidstack
⚡ Früh bewerben San Francisco, CA Vor Ort $208,000–$269,000
● Neu 👁 Gesehen ✓ Beworben vor 12 Std.
DoorDash USA
Senior Reliability Engineer
DoorDash USA
⚡ Früh bewerben San Francisco, CA; Oakland, CA Vor Ort $139,400–$205,000
● Neu 👁 Gesehen ✓ Beworben vor 16 Std.
Fluidstack
Reliability Engineer, R&D
Fluidstack
⚡ Früh bewerben Austin, TX Vor Ort $203,000–$232,000
● Neu 👁 Gesehen ✓ Beworben vor 3 Tg.
Pinterest
Site Reliability Engineer II, tvScientific
Pinterest
⚡ Früh bewerben San Francisco, CA, US; Remote,... · standortgebunden $114,297–$235,319
● Neu 👁 Gesehen ✓ Beworben vor 3 Tg.
Anthropic
Staff Software Engineer, AI Reliability
Anthropic
⚡ Früh bewerben San Francisco, CA | New York C... Vor Ort $325,000–$485,000
● Neu 👁 Gesehen ✓ Beworben vor 6 Tg.
DoorDash USA
Software Engineer, Reliability Platforms
DoorDash USA
⚡ Früh bewerben San Francisco, CA; Sunnyvale,... Vor Ort $159,800–$235,000
● Neu 👁 Gesehen ✓ Beworben vor 1 Wo.

Registrieren für Vorschläge, die auf die von Ihnen geöffneten Jobs und gespeicherten Suchen zugeschnitten sind.

Mehr Jobs bei Fluidstack

Alle Jobs bei Fluidstack ansehen →

Jetzt bewerben
🤖

Moment — langsam

JobsRadar wurde für echte Menschen gebaut, die eine schwere Zeit bei der Jobsuche haben — nicht für automatisierte Anfragen. Sie klicken viel zu schnell und sind jetzt vorübergehend blockiert.

Kommen Sie später wieder. Wenn Sie wirklich auf Jobsuche sind, stehen wir hinter Ihnen — verhalten Sie sich einfach wie ein Mensch.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Verschaffe dir einen Vorsprung bei der Jobsuche.

Tritt unserem Telegram-Kanal bei für das, was dir hilft, die Stelle zu bekommen — Gehaltsbenchmarks, den wöchentlichen Marktpuls und neue Feature-Drops. Kein Spam, nur Signal.

Dem Kanal beitreten — kostenlos