Jobs Companies Liquid AI Member of Technical Staff - GPU Infrastructure Engineer

About this Member of Technical Staff - GPU Infrastructure Engineer role at Liquid AI

Liquid AI · Hybrid · San Francisco

About Liquid AI

Spun out of MIT CSAIL, we build general-purpose AI systems that run efficiently across deployment targets, from data center accelerators to on-device hardware, ensuring low latency, minimal memory usage, privacy, and reliability. We partner with enterprises across consumer electronics, automotive, life sciences, and financial services. We are scaling rapidly and need exceptional people to help us get there.

The Opportunity

Our Cluster Infrastructure team owns the compute environments that power foundation model training and research at Liquid AI. We are looking for a hands-on software engineer to keep our GPU clusters reliable, improve resource efficiency, and build the tooling that allows researchers to focus on model development rather than infrastructure.

This role matters because infrastructure issues can delay training by days, while improvements in utilization, storage management, and automation can significantly increase research velocity and reduce compute costs. You will work closely with researchers and infrastructure engineers, owning problems from immediate operational response through long-term platform improvements.

What We’re Looking For

We need someone who:

  • Brings order to complex systems: You identify root causes and build durable fixes rather than repeatedly firefighting.

  • Is an engineer first: You can go deep across Linux, networking, storage, schedulers, and distributed systems.

  • Balances operations and engineering: You handle urgent issues while steadily replacing manual work with automation.

  • Owns outcomes: You communicate clearly, prioritize effectively, and drive problems to resolution across internal teams and external providers.

The Work

  • Own the reliability and operation of the GPU clusters used for training and research.

  • Debug issues across compute, storage, networking, schedulers, and distributed workloads.

  • Improve CPU, GPU, and storage utilization through better tooling and automation.

  • Onboard and migrate workloads across GPU providers and hardware platforms.

  • Build monitoring, validation, and platform abstractions that reduce operational work for researchers.

  • Contribute to the longer-term architecture of Liquid AI’s training infrastructure and GPU platform.

Desired Experience

Must-have

  • Strong software engineering experience, with the ability to build production-quality infrastructure tooling and automation.

  • Deep knowledge of distributed systems, Linux, networking, and storage.

  • Experience operating a shared compute cluster or distributed training platform.

  • A track record of supporting production users and turning recurring failures into durable solutions.

  • The technical depth to partner effectively with senior research and infrastructure engineers.

Nice-to-have

  • Experience with SLURM, Kubernetes, Ray, Hadoop, or another distributed compute platform.

  • Experience supporting GPU, HPC, or large-scale AI training infrastructure.

  • Experience with distributed storage, cluster schedulers, cloud providers, or infrastructure control planes.

What Success Looks Like (Year One)

  1. Researchers spend less time resolving infrastructure and resource-allocation issues.

  2. GPU, CPU, and storage resources are used more efficiently across the fleet.

  3. Recurring operational problems are replaced with automation, monitoring, and dependable platform tooling.

  4. Liquid AI has the beginnings of a durable internal platform that hides infrastructure complexity from researchers.

What We Offer

  • High-impact ownership: Own infrastructure that directly affects how quickly and efficiently we train foundation models.

  • Compensation: Competitive base salary with equity in a unicorn-stage company.

  • Health: We pay 100% of medical, dental, and vision premiums for employees and dependents.

  • Financial: 401(k) matching up to 4% of base pay.

  • Time Off: Unlimited PTO plus company-wide Refill Days throughout the year.

Ready to apply to Liquid AI?
Apply to Liquid AI

Similar jobs

Redwood Materials
Infrastructure Software Engineer, Energy Storage
Redwood Materials
⚡ Apply early San Francisco, California, Uni... Onsite $180,000–$237,500
● New 👁 Seen ✓ Applied 11m ago
Stitch Fix
ML Platform Engineer
Stitch Fix
⚡ Apply early Remote, USA · location restricted $136,000–$167,000
● New 👁 Seen ✓ Applied 12m ago
Speechify
Software Engineer, Platform - San Francisco, CA, USA
Speechify
⚡ Apply early San Francisco, CA, USA Onsite
● New 👁 Seen ✓ Applied 4h ago
Harvey
Frontend Platform Engineer
Harvey
⚡ Apply early New York Hybrid $193,400–$290,000
● New 👁 Seen ✓ Applied 7h ago
Front
Senior Software Engineer, Mobile Platform (React Native)
Front
⚡ Apply early San Francisco, CA Hybrid $205,000–$235,000
● New 👁 Seen ✓ Applied 7h ago
Speechify
Software Engineer, Data Infrastructure & Acquisition - San Francisco, CA, USA
Speechify
⚡ Apply early San Francisco, CA, USA Onsite $140,000–$200,000
● New 👁 Seen ✓ Applied 10h ago
Drata
Staff Software Engineer, Monetization Platform
Drata
⚡ Apply early Hybrid - San Francisco Hybrid $200,000–$271,500
● New 👁 Seen ✓ Applied 13h ago
Zipline
Senior Integration and Test Software Engineer - Long Range Platform
Zipline
⚡ Apply early South San Francisco, Californi... Onsite 🛂 Visa sponsorship
● New 👁 Seen ✓ Applied 13h ago
Pinterest
Principal Engineer, AI Platform
Pinterest
⚡ Apply early San Francisco, CA, US; Remote,... · location restricted $242,634–$499,541
● New 👁 Seen ✓ Applied 14h ago

Sign up for suggestions tailored to the jobs you open and the searches you save.

More jobs at Liquid AI

See all jobs at Liquid AI →

Apply now
🤖

Whoa — hold up

JobsRadar was built for real people having a rough time in their job search — not for automated requests. You're clicking way too fast and you're now temporarily blocked.

Come back later. If you're genuinely job hunting, we've got your back — just act like a human.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Get an edge on your job hunt.

Join our Telegram channel for the stuff that helps you land the role — salary benchmarks, the weekly market pulse, and new-feature drops. No spam, just signal.

Join the channel — it's free