Jobs Companies Liquid AI Member of Technical Staff - GPU Infrastructure Engineer

Über diese Member of Technical Staff - GPU Infrastructure Engineer Stelle bei Liquid AI

Liquid AI · Hybrid · San Francisco

About Liquid AI

Spun out of MIT CSAIL, we build general-purpose AI systems that run efficiently across deployment targets, from data center accelerators to on-device hardware, ensuring low latency, minimal memory usage, privacy, and reliability. We partner with enterprises across consumer electronics, automotive, life sciences, and financial services. We are scaling rapidly and need exceptional people to help us get there.

The Opportunity

Our Cluster Infrastructure team owns the compute environments that power foundation model training and research at Liquid AI. We are looking for a hands-on software engineer to keep our GPU clusters reliable, improve resource efficiency, and build the tooling that allows researchers to focus on model development rather than infrastructure.

This role matters because infrastructure issues can delay training by days, while improvements in utilization, storage management, and automation can significantly increase research velocity and reduce compute costs. You will work closely with researchers and infrastructure engineers, owning problems from immediate operational response through long-term platform improvements.

What We’re Looking For

We need someone who:

  • Brings order to complex systems: You identify root causes and build durable fixes rather than repeatedly firefighting.

  • Is an engineer first: You can go deep across Linux, networking, storage, schedulers, and distributed systems.

  • Balances operations and engineering: You handle urgent issues while steadily replacing manual work with automation.

  • Owns outcomes: You communicate clearly, prioritize effectively, and drive problems to resolution across internal teams and external providers.

The Work

  • Own the reliability and operation of the GPU clusters used for training and research.

  • Debug issues across compute, storage, networking, schedulers, and distributed workloads.

  • Improve CPU, GPU, and storage utilization through better tooling and automation.

  • Onboard and migrate workloads across GPU providers and hardware platforms.

  • Build monitoring, validation, and platform abstractions that reduce operational work for researchers.

  • Contribute to the longer-term architecture of Liquid AI’s training infrastructure and GPU platform.

Desired Experience

Must-have

  • Strong software engineering experience, with the ability to build production-quality infrastructure tooling and automation.

  • Deep knowledge of distributed systems, Linux, networking, and storage.

  • Experience operating a shared compute cluster or distributed training platform.

  • A track record of supporting production users and turning recurring failures into durable solutions.

  • The technical depth to partner effectively with senior research and infrastructure engineers.

Nice-to-have

  • Experience with SLURM, Kubernetes, Ray, Hadoop, or another distributed compute platform.

  • Experience supporting GPU, HPC, or large-scale AI training infrastructure.

  • Experience with distributed storage, cluster schedulers, cloud providers, or infrastructure control planes.

What Success Looks Like (Year One)

  1. Researchers spend less time resolving infrastructure and resource-allocation issues.

  2. GPU, CPU, and storage resources are used more efficiently across the fleet.

  3. Recurring operational problems are replaced with automation, monitoring, and dependable platform tooling.

  4. Liquid AI has the beginnings of a durable internal platform that hides infrastructure complexity from researchers.

What We Offer

  • High-impact ownership: Own infrastructure that directly affects how quickly and efficiently we train foundation models.

  • Compensation: Competitive base salary with equity in a unicorn-stage company.

  • Health: We pay 100% of medical, dental, and vision premiums for employees and dependents.

  • Financial: 401(k) matching up to 4% of base pay.

  • Time Off: Unlimited PTO plus company-wide Refill Days throughout the year.

Bereit, sich bei Liquid AI zu bewerben?
Bei Liquid AI bewerben

Ähnliche Jobs

Harvey
Frontend Platform Engineer
Harvey
⚡ Früh bewerben New York Hybrid $193,400–$290,000
● Neu 👁 Gesehen ✓ Beworben vor 1 Std.
Front
Senior Software Engineer, Mobile Platform (React Native)
Front
⚡ Früh bewerben San Francisco, CA Hybrid $205,000–$235,000
● Neu 👁 Gesehen ✓ Beworben vor 1 Std.
Redwood Materials
Infrastructure Software Engineer, Energy Storage
Redwood Materials
⚡ Früh bewerben San Francisco, California, Uni... Vor Ort $180,000–$237,500
● Neu 👁 Gesehen ✓ Beworben vor 2 Std.
Stitch Fix
ML Platform Engineer
Stitch Fix
⚡ Früh bewerben Remote, USA · standortgebunden $136,000–$167,000
● Neu 👁 Gesehen ✓ Beworben vor 2 Std.
Drata
Staff Software Engineer, Monetization Platform
Drata
⚡ Früh bewerben Hybrid - San Francisco Hybrid $200,000–$271,500
● Neu 👁 Gesehen ✓ Beworben vor 7 Std.
Zipline
Senior Integration and Test Software Engineer - Long Range Platform
Zipline
⚡ Früh bewerben South San Francisco, Californi... Vor Ort 🛂 Visa-Sponsoring
● Neu 👁 Gesehen ✓ Beworben vor 8 Std.
Pinterest
Principal Engineer, AI Platform
Pinterest
⚡ Früh bewerben San Francisco, CA, US; Remote,... · standortgebunden $242,634–$499,541
● Neu 👁 Gesehen ✓ Beworben vor 8 Std.
Pinterest
Sr. Staff Software Engineer, Data Product Platform
Pinterest
⚡ Früh bewerben San Francisco, CA, US; Remote,... · standortgebunden $208,592–$429,454
● Neu 👁 Gesehen ✓ Beworben vor 8 Std.
Phonic
Platform Engineer
Phonic
⚡ Früh bewerben San Francisco Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 9 Std.

Registrieren für Vorschläge, die auf die von Ihnen geöffneten Jobs und gespeicherten Suchen zugeschnitten sind.

Mehr Jobs bei Liquid AI

Alle Jobs bei Liquid AI ansehen →

Jetzt bewerben
🤖

Moment — langsam

JobsRadar wurde für echte Menschen gebaut, die eine schwere Zeit bei der Jobsuche haben — nicht für automatisierte Anfragen. Sie klicken viel zu schnell und sind jetzt vorübergehend blockiert.

Kommen Sie später wieder. Wenn Sie wirklich auf Jobsuche sind, stehen wir hinter Ihnen — verhalten Sie sich einfach wie ein Mensch.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Verschaffe dir einen Vorsprung bei der Jobsuche.

Tritt unserem Telegram-Kanal bei für das, was dir hilft, die Stelle zu bekommen — Gehaltsbenchmarks, den wöchentlichen Marktpuls und neue Feature-Drops. Kein Spam, nur Signal.

Dem Kanal beitreten — kostenlos