Jobs Companies Astera Labs Machine Learning Infrastructure Engineer

À propos de ce poste Machine Learning Infrastructure Engineer chez Astera Labs

Astera Labs · Sur site · San Jose, California, United States

Astera Labs (NASDAQ: ALAB) provides rack-scale AI infrastructure through purpose-built connectivity solutions. By collaborating with hyperscalers and ecosystem partners, Astera Labs enables organizations to unlock the full potential of modern AI. Astera Labs’ Intelligent Connectivity Platform integrates CXL®, Ethernet, NVLink, PCIe®, and UALink™ semiconductor-based technologies with the company’s COSMOS software suite to unify diverse components into cohesive, flexible systems that deliver end-to-end scale-up, and scale-out connectivity. The company’s custom connectivity solutions business complements its standards-based portfolio, enabling customers to deploy tailored architectures to meet their unique infrastructure requirements. Discover more at www.asteralabs.com.

 

Machine Learning Infrastructure Engineer

Location: San Jose, CA
Experience: 1–5 years
Team: Applied AI

The role

We’re hiring a Machine Learning Infrastructure Engineer to build the runtime, platform, and operational backbone for modern AI systems. This role is for someone who wants to work on the systems behind the systems: model access layers, routing, serving paths, telemetry, observability, evaluation infrastructure, and the controls needed to make fast-moving AI work reliable in practice.

 

This is a platform role, but not in the old sense. The work is tightly coupled to how modern AI systems are actually built and used: multiple model providers, agent runtimes, skill and tool layers, inference telemetry, cost-aware routing, AI spend visibility, and governance that is strong enough for real internal adoption.

 

What you’ll do

  • Build and improve internal AI infrastructure for LLM applications, agents, retrieval systems, and model-backed engineering workflows.
  • Own inference deployment paths across managed and self-serve environments, including access control, monitoring, and operational reliability.
  • Build platform layers such as model gateways, routing, runtime integrations, telemetry, and controls for safe execution at scale.
  • Develop AI Ops capabilities across evaluation, release readiness, observability, incident triage, regression detection, and cost monitoring.
  • Build dashboards, tracing, logging, and alerting for production AI systems, including spend and usage visibility across tools and teams.
  • Improve performance and unit economics through routing, caching, batching, failover, and latency/cost optimization.
  • Create reusable APIs, SDKs, and platform abstractions that make AI systems easier to deploy, evaluate, govern, and operate.

What we’re looking for

  • 1–5 years of experience in software engineering, ML infrastructure, MLOps, platform engineering, or related backend/infrastructure roles.
  • Strong Python plus strong systems instincts.
  • Experience with AWS or GCP and real production service ownership.
  • Familiarity with inference deployments, model APIs, gateways, serving systems, or runtime infrastructure for LLM/ML workloads.
  • Experience with observability, telemetry, reliability engineering, and incident response.
  • Understanding of eval systems, release workflows, retrieval-backed systems, and debugging non-deterministic AI behavior.
  • Ability to translate messy platform needs into scalable internal infrastructure.

What strong candidates often look like

They have built or operated systems where latency, routing, cost, telemetry, and reliability actually matter. They understand that modern AI infrastructure is not just about getting a model endpoint running. It is about building the runtime, visibility, controls, and developer experience that let an applied AI team move fast without losing quality or trust.

 

Why this role is interesting

The team is building AI-ready infrastructure in the most literal sense: observability, access control, AI spend tracking, secure managed platforms, skill/tool infrastructure, and telemetry that spans requests, tools, models, and outcomes. If you want to work on the platform layer that makes modern agentic systems possible — and do it in a setting where the downstream users are serious engineers with high expectations — this is that role.

 

The base pay compensation range for this role is between $140,000 - $165,000

We know that creativity and innovation happen more often when teams include diverse ideas, backgrounds, and experiences, and we actively encourage everyone with relevant experience to apply, including people of color, LGBTQ+ and non-binary people, veterans, parents, and individuals with disabilities.

Prêt à postuler chez Astera Labs ?
Postuler chez Astera Labs

Comment se compare ce salaire pour Platform Engineer

Ce poste paie $152,500/yrdans la fourchette habituelle pour les postes Platform Engineer.

$139,375 la médiane $176,875 $206,600

Fourchette typique $143,425–$204,213/yr, à partir de 6 annonces Platform Engineer comparables sur JobsRadar (rémunération annualisée en USD). Voir les aperçus de salaire pour Platform Engineer →

Emplois similaires

BILL
Principal Partner Solution Engineer - Partner Embed Platform
BILL
⚡ Postuler tôt Draper, Utah, United States; S... Sur site $20,800–$260,000
● Nouveau 👁 Vu ✓ Postulé il y a 5 j
Samsung Semiconductor
Staff Engineer, Workbench Platform
Samsung Semiconductor
⚡ Postuler tôt San Jose, California, United S... Sur site $163,000–$253,000
● Nouveau 👁 Vu ✓ Postulé il y a 1 sem.
Archer
Senior Staff Backend Engineer, Platform Core
Archer
⚡ Postuler tôt San Jose, California, United S... Sur site $182,500–$220,000
● Nouveau 👁 Vu ✓ Postulé il y a 2 sem.
Archer
Senior Backend Engineer, Platform
Archer
⚡ Postuler tôt San Jose, California, United S... Sur site $126,700–$150,000
● Nouveau 👁 Vu ✓ Postulé il y a 2 sem.
Archer
Sr Staff Engineer, Data Infrastructure
Archer
⚡ Postuler tôt San Jose, California, United S... $182,400–$228,000
● Nouveau 👁 Vu ✓ Postulé il y a 2 mois
Lambda
Staff Software Engineer - Infrastructure Storage
Lambda
⚡ Postuler tôt San Francisco Office (Fremont... Hybride $314,000–$465,000
● Nouveau 👁 Vu ✓ Postulé il y a 4 h
Horizon3 AI
Principal Engineer, Tech Lead - Platform Integration Team
Horizon3 AI
⚡ Postuler tôt US, Remote · lieu restreint $213,000–$273,000
● Nouveau 👁 Vu ✓ Postulé il y a 4 h
Farsight AI
Senior Platform Engineer
Farsight AI
⚡ Postuler tôt New York City Sur site $140,000–$180,000
● Nouveau 👁 Vu ✓ Postulé il y a 4 h
Applied Intuition
Software Engineer - Developer Infrastructure
Applied Intuition
⚡ Postuler tôt Sunnyvale Sur site $135,000–$265,000
● Nouveau 👁 Vu ✓ Postulé il y a 4 h

Inscrivez-vous pour des suggestions adaptées aux emplois que vous ouvrez et aux recherches que vous enregistrez.

Plus d’emplois chez Astera Labs

Voir tous les emplois chez Astera Labs →

Postuler maintenant
🤖

Doucement — un instant

JobsRadar a été conçu pour de vraies personnes qui traversent une période difficile dans leur recherche d’emploi — pas pour des requêtes automatisées. Vous cliquez beaucoup trop vite et vous êtes maintenant temporairement bloqué.

Revenez plus tard. Si vous cherchez réellement un emploi, nous sommes de votre côté — agissez simplement comme un être humain.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Prenez une longueur d’avance dans votre recherche d’emploi.

Rejoignez notre canal Telegram pour ce qui vous aide à décrocher le poste — références salariales, le pouls hebdomadaire du marché et les annonces de nouveautés. Pas de spam, que du signal.

Rejoindre le canal — c’est gratuit