Jobs Companies NVIDIA Senior Site Reliability Engineer

À propos de ce poste Senior Site Reliability Engineer chez NVIDIA

NVIDIA · Sur site · Israel, Yokneam

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.

We are seeking a Staff DB SRE to build the runtime foundation for NVIDIA’s enterprise AI platforms — with a strong emphasis on database infrastructure at scale. This role blends large-scale database transformation with the building and development of GPU-accelerated platforms. You'll develop the software systems, automation frameworks, and high-performance database services that power NVIDIA’s AI workloads at scale.

What you'll be doing:

  • Design and operate highly available database clusters (MySQL, MSSQL, Oracle) with automated replication, failover, point-in-time recovery, and disaster-recovery strategies at enterprise scale.

  • Drive database performance engineering — own query optimization, indexing strategies, connection pooling, lock-contention analysis, and storage-engine tuning for production systems handling millions of transactions.

  • Build self-service database lifecycle automation — from one-click cluster provisioning and schema migrations to zero-downtime upgrades, blue-green deployments, and automated capacity scaling.

  • Bridge relational and AI-native data infrastructure — extend traditional database expertise into vector search, GPU-accelerated query engines, and hybrid data platforms that serve both classic workloads and AI applications.

  • Compose and build software platforms that transform legacy database systems into modern and scalable architectures.

  • Run vector & graph database services and query engines to handle AI/ML data workloads with ultra-low latency.

  • Build automation frameworks for provisioning, schema evolution, scaling, and failover integrated directly into CI/CD workflows.

  • Build developer-focused tooling for monitoring, profiling, and debugging database performance in real time.

  • Contribute to architecture, coding standards, and guidelines for long-term platform evolution.

  • Participate in on-call rotations to ensure flawless operation of critical database services.

What we need to see:

  • BS, MS, or PhD in Computer Science, Engineering, or a related field—or equivalent experience.

  • 8+ years of Database engineering experience with deep expertise in database systems or distributed data platforms

  • Deep hands-on expertise with one or more major relational database engines (like Oracle, MySQL, MSSQL) — including replication topologies, backup/restore strategies, and high-availability architecture.

  • Proven background in query optimization, data partitioning, and large-scale performance tuning.

  • Experience building or operating managed database services or internal Database-as-a-Service platforms — automating provisioning, monitoring, patching, and failover for database fleets at scale.

  • Strong programming skills in Python, Go, with a track record of building production-grade systems.

  • Demonstrable experience crafting high-performance, high-availability relational database services.

  • Experience with container orchestration (Kubernetes) and cloud-native database deployment patterns.

  • Hands-on experience establishing DevOps guidelines, e.g., CI/CD, monitoring, alerting, SLAs, capacity forecasting, etc.

Ways to stand out from the crowd:

  • Strong Kubernetes/Infrastructure as code and coding experience in addition to core DBA skills.

  • Expertise in hybrid/multi-region database replication strategies for low-latency AI workloads.

  • Demonstrable understanding of observability and performance profiling tools for complex data systems.

  • You’ve built an internal Database-as-a-Service offering — engineers request a cluster and get a production-ready, fully monitored, backed-up database without filing a ticket.

  • Hands-on experience with database migration tooling, schema evolution pipelines, and zero-downtime upgrade strategies across heterogeneous database engines

NVIDIA is committed to encouraging a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/ 

Prêt à postuler chez NVIDIA ?
Postuler chez NVIDIA

À propos de NVIDIA

NVIDIA pioneered accelerated computing. Today, our AI infrastructure powers global intelligence, transforming every industry. Learn more about NVIDIA .

Voir tous les emplois chez NVIDIA →

Emplois similaires

NVIDIA
Senior Site Reliability Engineer
NVIDIA
⚡ Postuler tôt Israel, Yokneam Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 1 sem.
NVIDIA
Senior Site Reliability Engineering - Storage
NVIDIA
⚡ Postuler tôt Israel, Yokneam Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 1 sem.
NVIDIA
Senior HPC Site Reliability Engineer
NVIDIA
⚡ Postuler tôt Israel, Yokneam Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 5 mois
RELX
Site Reliability Engineering Lead
RELX
⚡ Postuler tôt Florida Sur site $118,300–$219,800
● Nouveau 👁 Vu ✓ Postulé il y a 38 min
RELX
Site Reliability Engineer II
RELX
⚡ Postuler tôt Home based-Georgia · lieu restreint $71,600–$119,400
● Nouveau 👁 Vu ✓ Postulé il y a 38 min
RELX
Site Reliability Engineer III
RELX
⚡ Postuler tôt Alpharetta, GA Sur site $86,600–$144,400
● Nouveau 👁 Vu ✓ Postulé il y a 38 min
AES
Engineer, Reliability
AES
⚡ Postuler tôt US, Louisville, CO Sur site $94,000–$112,625
● Nouveau 👁 Vu ✓ Postulé il y a 1 h
AES
Senior Reliability Engineer
AES
⚡ Postuler tôt US, Louisville, CO Sur site $113,000–$141,525
● Nouveau 👁 Vu ✓ Postulé il y a 1 h
Anduril Industries
Senior Infrastructure Reliability Engineer
Anduril Industries
⚡ Postuler tôt Costa Mesa, California, United... Sur site $166,000–$220,000
● Nouveau 👁 Vu ✓ Postulé il y a 6 h

Inscrivez-vous pour des suggestions adaptées aux emplois que vous ouvrez et aux recherches que vous enregistrez.

Plus d’emplois chez NVIDIA

Voir tous les emplois chez NVIDIA →

Postuler maintenant
🤖

Doucement — un instant

JobsRadar a été conçu pour de vraies personnes qui traversent une période difficile dans leur recherche d’emploi — pas pour des requêtes automatisées. Vous cliquez beaucoup trop vite et vous êtes maintenant temporairement bloqué.

Revenez plus tard. Si vous cherchez réellement un emploi, nous sommes de votre côté — agissez simplement comme un être humain.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Prenez une longueur d’avance dans votre recherche d’emploi.

Rejoignez notre canal Telegram pour ce qui vous aide à décrocher le poste — références salariales, le pouls hebdomadaire du marché et les annonces de nouveautés. Pas de spam, que du signal.

Rejoindre le canal — c’est gratuit