Jobs Companies NVIDIA Senior Site Reliability Engineering, Storage

À propos de ce poste Senior Site Reliability Engineering, Storage chez NVIDIA

NVIDIA · Sur site · India, Bengaluru

NVIDIA has been redefining computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s an outstanding legacy of innovation that’s fueled by phenomenal technology – and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.

 

We are seeking a Senior Site Reliability Engineer – Storage, you will own the reliability, performance, and scalability of our global NAS, SAN, and Object Storage platforms that power critical internal and external services. You will combine deep storage expertise with strong automation and SRE practices to design, build, and operate highly available storage systems at scale.

 

What you will be doing:

  • Lead design, deployment, and operations of production NAS, SAN, and Object Storage platforms, ensuring reliability, performance, and security.

  • Capture requirements from partner teams, architect storage solutions, and drive end‑to‑end implementation for new and existing services.

  • Develop, maintain, and improve automation for provisioning, configuration, monitoring, incident response, and lifecycle management of storage infrastructure.

  • Participate in on‑call and incident response, lead troubleshooting of complex storage and performance issues, and drive root cause analysis and preventive actions.

  • Define and track SLOs/SLIs and error budgets for storage services, using observability and analytics to continuously improve reliability and efficiency.

  • Build and maintain runbooks, standard operating procedures, and comprehensive documentation for storage services and automation.

  • Analyze capacity and usage trends, perform forecasting, and recommend scaling or optimization strategies to support business growth.

  • Collaborate closely with SRE, infrastructure, networking, and application teams in a follow‑the‑sun model to deliver consistent, high‑quality service.

  • Mentor junior engineers, share best practices, and help drive adoption of SRE principles across the team.

 

What we need to see:

  • 12+ years of experience in Site Reliability, DevOps, or Infrastructure Engineering, with significant focus on storage systems.

  • Bachelor’s degree in Computer Science, Computer Engineering, or a related technical field or equivalent practical experience.

  • Strong hands‑on experience with design, deployment, and operations of enterprise‑grade NAS, SAN, and/or Object Storage platforms.

  • Solid understanding of SRE concepts (SLOs/SLIs, error budgets, incident management, observability, postmortems).

  • Proficiency with Infrastructure as Code and configuration management tools (e.g., Terraform, Ansible, Puppet, SaltStack) and source control systems.

  • Experience building and operating highly available, scalable infrastructure, including automation for provisioning, monitoring, and remediation.

  • Experience with container and virtualization platforms (e.g., Docker, Kubernetes, hypervisors) and modern CI/CD and version control tools.

  • Strong scripting or programming skills (e.g., Python, Go, Shell) to build tools, automate workflows, and integrate systems.

  • Excellent communication and collaboration skills, with the ability to work effectively across distributed and cross‑functional teams.

 

Ways to stand out from the crowd:

  • Experience with storage for high‑performance computing, AI/ML workloads, or large‑scale data analytics.

  • Proven ability to debug complex, distributed systems and storage performance issues.

  • History of driving reliability improvements through data‑driven analysis and automation.

  • Experience leading technical initiatives, mentoring engineers, or acting as a technical lead on critical projects.

Prêt à postuler chez NVIDIA ?
Postuler chez NVIDIA

À propos de NVIDIA

NVIDIA pioneered accelerated computing. Today, our AI infrastructure powers global intelligence, transforming every industry. Learn more about NVIDIA .

Voir tous les emplois chez NVIDIA →

Emplois similaires

NVIDIA
Senior System Software Engineer – Linux Kernel Automotive Cybersecurity
NVIDIA
⚡ Postuler tôt India, Bengaluru Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 9 h
NVIDIA
Senior System Software Engineer, Software Defined Networking
NVIDIA
⚡ Postuler tôt US, CA, Santa Clara Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 9 h
NVIDIA
Senior Physical Design and Timing Engineer
NVIDIA
⚡ Postuler tôt US, MA, Westford Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 9 h
NVIDIA
Accelerated Compute Systems Performance Architect Intern - 2027
NVIDIA
⚡ Postuler tôt China, Shanghai Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 1 j
NVIDIA
Infrastructure Tool Development Intern - 2027
NVIDIA
⚡ Postuler tôt China, Shanghai Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 1 j
NVIDIA
Senior MLOps Engineer - DSX Enablement
NVIDIA
⚡ Postuler tôt Germany, Remote · lieu restreint
● Nouveau 👁 Vu ✓ Postulé il y a 1 j
NVIDIA
ASIC Design and Verification Intern, SOC - 2027
NVIDIA
⚡ Postuler tôt China, Shanghai Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 2 j
NVIDIA
Senior Chip Design Engineer
NVIDIA
⚡ Postuler tôt Israel, Tel Aviv Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 2 j
NVIDIA
Senior Business Intelligence Analyst
NVIDIA
⚡ Postuler tôt Israel, Yokneam Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 2 j

Inscrivez-vous pour des suggestions adaptées aux emplois que vous ouvrez et aux recherches que vous enregistrez.

Plus d’emplois chez NVIDIA

Voir tous les emplois chez NVIDIA →

Postuler maintenant
🤖

Doucement — un instant

JobsRadar a été conçu pour de vraies personnes qui traversent une période difficile dans leur recherche d’emploi — pas pour des requêtes automatisées. Vous cliquez beaucoup trop vite et vous êtes maintenant temporairement bloqué.

Revenez plus tard. Si vous cherchez réellement un emploi, nous sommes de votre côté — agissez simplement comme un être humain.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Prenez une longueur d’avance dans votre recherche d’emploi.

Rejoignez notre canal Telegram pour ce qui vous aide à décrocher le poste — références salariales, le pouls hebdomadaire du marché et les annonces de nouveautés. Pas de spam, que du signal.

Rejoindre le canal — c’est gratuit