Jobs Companies DiDi Labs Senior/Sr. Staff AI Infrastructure Engineer, Inference & Optimization

À propos de ce poste Senior/Sr. Staff AI Infrastructure Engineer, Inference & Optimization chez DiDi Labs

DiDi Labs · Sur site · San Jose, CA

About the Company

DiDi's autonomous driving unit was established in 2016 with the mission of developing Level 4 autonomous driving (AD) technology to make transportation safer and more efficient. In August 2019, the unit became an independent company, DiDi Autonomous Driving, dedicated to advanced AD R&D, product application, and business expansion. We believe integrating AD technology into a shared-mobility fleet will generate immense social value. By leveraging DiDi's specialized technology, operational expertise, and integrated ecosystem, we are positioned to build and operate a highly efficient, user-oriented autonomous fleet.

 

About The Role

We are seeking an experienced and mission-driven Senior/Sr. Staff AI Infrastructure Engineer, Inference & Optimization to lead the performance tuning, deployment, and resource scheduling of cutting-edge AI models across on-vehicle and cloud infrastructure. In this role, you will design high-efficiency inference pipelines, build system-level stability frameworks, and optimize hardware execution to ensure ultra-low latency and rock-solid operational reliability. You will act as a technical leader in AI infrastructure, accelerating model iteration and bridging the gap between frontier deep learning algorithms and real-time autonomous systems.

 

Responsibilities

  • Own the deployment, optimization, and resource scheduling of vehicle-side AI models, ensuring high efficiency, low latency, and robust execution within embedded constraints.

  • Lead vehicle-side system stability initiatives, conducting independent root-cause analysis and driving resolution for complex, system-level performance bottlenecks and runtime anomalies.

  • Architect and scale service-oriented deployment environments for Large Language Models (LLMs) and foundational models to support offline simulation, automated annotation, and rapid model validation.

  • Track and evaluate cutting-edge industry methodologies, continuously integrating advanced optimization toolchains, quantization techniques, and execution engines.

  • Establish system-level profiling and telemetry frameworks using CUDA tools to monitor, analyze, and maximize hardware utilization across target GPU architectures.

  • Collaborate cross-functionally with Autonomous Driving Perception/Prediction, Cloud Infrastructure, and Safety teams to enable rapid algorithm iteration and scalable vehicle deployment.

 

Qualifications

  • Master’s or higher degree in Computer Science, Software Engineering, Systems Engineering, or a closely related technical field.

  • 3-8+ years of industry experience in high-performance computing, AI infrastructure, model optimization, or embedded deployment.

  • Strong proficiency in C++ and Python, with solid expertise in parallel programming (CUDA, OpenMP) and low-level system profiling tools.

  • Deep familiarity with mainstream inference engines (e.g., TensorRT, ONNX Runtime) and specialized LLM inference/serving frameworks (e.g., vLLM, SGLang, TensorRT-LLM).

  • Practical understanding of modern GPU hardware architectures (e.g., NVIDIA Hopper, Thor) and memory bandwidth management.

  • Demonstrated ability to diagnose complex software-hardware integration issues and drive scalable, production-grade solutions.

Preferred Qualifications

  • Hands-on experience optimizing and deploying AI models on the NVIDIA Thor platform, including hardware resource scheduling and acceleration.

  • Proven track record of serving large foundation models (e.g., LLaMA, Qwen, GPT) in production or high-throughput cloud pipelines using frameworks like vLLM, SGLang, TGI, or LightLLM.

  • Background in deep learning training frameworks (PyTorch) and practical experience with model quantization (INT8/FP8/AWQ), kernel fusion, or graph compilation.

  • Experience deploying real-time, high-availability AI workloads in autonomous vehicles, robotics, or edge devices.

 

The base salary range for this full-time position is $169,783 - $351,000 annually in addition to bonus, equity and benefits. Our salary ranges are determined by role, level, and location. Within the range, individual pay is determined by work location and additional factors, including job-related skills, experience, and relevant education or training.

I acknowledge that prior to submitting this application, I have read and accepted the Privacy Notice for California Residents which is available on https://v.didi.cn/AQnxlBa

Prêt à postuler chez DiDi Labs ?
Postuler chez DiDi Labs

Comment se compare ce salaire pour Platform Engineer

Ce poste paie $260,392/yrau-dessus de la fourchette habituelle pour les postes Platform Engineer.

$170,088 la médiane $204,500 $251,275

Fourchette typique $183,500–$223,500/yr, à partir de 15 annonces Platform Engineer comparables sur JobsRadar (rémunération annualisée en USD). Voir les aperçus de salaire pour Platform Engineer →

Emplois similaires

Veeam Software
Senior Platform Engineer (Cloud Workloads)
Veeam Software
⚡ Postuler tôt San Jose, CA, USA Sur site $178,200–$297,000
● Nouveau 👁 Vu ✓ Postulé il y a 8 h
MS
Software Engineer, Simulation Infrastructure
Muon Space
⚡ Postuler tôt San Jose, CA Hybride $156,000–$186,000
● Nouveau 👁 Vu ✓ Postulé il y a 19 h
Weride
Software Engineer - Backend / Infrastructure
Weride
⚡ Postuler tôt San Jose, CA Sur site $130,000–$182,000
● Nouveau 👁 Vu ✓ Postulé il y a 2 j
Weride
Senior Software Engineer - Cloud Infrastructure
Weride
⚡ Postuler tôt San Jose, CA Sur site $143,000–$197,438
● Nouveau 👁 Vu ✓ Postulé il y a 2 j
Speechify
Software Engineer, Data Infrastructure & Acquisition - San Jose, CA, USA
Speechify
⚡ Postuler tôt San Jose, CA, USA Sur site $140,000–$200,000
● Nouveau 👁 Vu ✓ Postulé il y a 2 j
MS
Senior Software Engineer, Software Platform
Muon Space
⚡ Postuler tôt San Jose, CA Hybride $184,000–$208,000
● Nouveau 👁 Vu ✓ Postulé il y a 1 sem.
Speechify
Software Engineer, Platform - San Jose, CA, USA
Speechify
⚡ Postuler tôt San Jose, CA, USA Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 1 sem.
Figure
Staff Infrastructure Engineer
Figure
⚡ Postuler tôt San Jose, CA Sur site $175,000–$250,000
● Nouveau 👁 Vu ✓ Postulé il y a 1 sem.
Figure
Helix AI Engineer, Data Infrastructure
Figure
⚡ Postuler tôt San Jose, CA Sur site $150,000–$400,000
● Nouveau 👁 Vu ✓ Postulé il y a 1 sem.

Inscrivez-vous pour des suggestions adaptées aux emplois que vous ouvrez et aux recherches que vous enregistrez.

Plus d’emplois chez DiDi Labs

Voir tous les emplois chez DiDi Labs →

Postuler maintenant
🤖

Doucement — un instant

JobsRadar a été conçu pour de vraies personnes qui traversent une période difficile dans leur recherche d’emploi — pas pour des requêtes automatisées. Vous cliquez beaucoup trop vite et vous êtes maintenant temporairement bloqué.

Revenez plus tard. Si vous cherchez réellement un emploi, nous sommes de votre côté — agissez simplement comme un être humain.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Prenez une longueur d’avance dans votre recherche d’emploi.

Rejoignez notre canal Telegram pour ce qui vous aide à décrocher le poste — références salariales, le pouls hebdomadaire du marché et les annonces de nouveautés. Pas de spam, que du signal.

Rejoindre le canal — c’est gratuit