Jobs Companies NVIDIA Senior Manager, Validation and HPC - NVIS

Über diese Senior Manager, Validation and HPC - NVIS Stelle bei NVIDIA

NVIDIA · Remote · US, TX, Remote

NVIDIA is in search of an HPC Deployment Manager to bolster our NVIDIA Infrastructure Specialists division! Across academia and industry, NVIDIA's products are driving ground-breaking advancements in deep learning, data analytics, and the optimization of data centers. Join our team, where we are at the forefront of constructing some of the globe's most expansive and rapid data centers! We seek an individual capable of supervising the deployment of cutting-edge InfiniBand and Ethernet technologies with a team comprising AI and HPC experts. This role demands dynamic interpersonal abilities and a customer-centric approach.

This Senior Manager will engage with clients, collaborators, and internal units to assess, delineate, and complete large-scale AI/HPC initiatives. They will orchestrate the day-to-day operations, guidance, and cultivation of a multi-layered team of HPC service professionals. This entails ensuring the timely delivery of a varied spectrum of AI HPC data center projects. Furthermore, this role offers an opportunity to thrive within a fast-paced, inventive, and technologically sophisticated atmosphere, emphasizing unparalleled performance and the exploration of an array of novel hardware and software technologies in AI supercomputing.

What you will be doing:

  • Directs and supervises the service HPC engineering functions in designing, developing, installing, and validating hardware and software for the Customer AI High-Performance Computing (HPC) systems.

  • Responsible for leading our HPC projects' planning, implementation, and performance. Improves the integrity of system services bring-up and related by applying groundbreaking technical and operational knowledge to configure and maintain HPC AI network and server platforms.

  • Drives HPC team hardware and software deployment, plans, develops, and deploys procedures for system validation.

  • Lead team activities and drive tests and plans for Customer's HPC AI systems implementations, custom scripts, and testing procedures to ensure operational reliability for the system.

  • Supports the HPC Engineering team, working with other internal collaborators to develop and run a well-rounded strategy for delivering service quality and continuous service improvement.

  • Leads team member development, helping them set and achieve goals for their career growth.

  • Build strong relationships with NVIDIA leaders, customers, partners, and collaborators. Works closely to identify, implement, and support leading NVIDIA's AI solutions engineering, maintaining currency with industry standards and innovations.

  • Be the domain authority with customers during planning calls through implementation.

What we need to see:

  • 10+ overall years' experience in IT, high-performance computing, or other related field; 3+ years of experience in a management or leadership role

  • Demonstrated expertise in HPC systems design configuration and planning, and solid knowledge of HPC storage

  • Proficiency with low latency/high-bandwidth interconnect infrastructure (Infiniband and Ethernet).

  • Expertise with HPC system software cluster management/provisioning tools, including job schedulers (Slurm, salt, xCAT).

  • Proficiency with shared and distributed memory parallelism (OpenMP, MPI, NCCL and HPL) and accelerators (GPUs).

  • Strong scripting ability (Bash, Perl, Python, etc.) and experience with programming fundamentals.

  • Expertise with administration, supervising and maintaining secure Linux/Unix operating systems (CentOS, Solaris).

  • Ability to understand and work with large, sophisticated systems, identify and resolve problems, handle performance, and troubleshoot network issues related to infrastructure.

  • Expertise with multi-vendor hardware/software management, security, and network/Internet protocols

  • Bachelor's degree in computer science, information systems, or a related field or equivalent experience

Ways to stand out from the crowd:

  • InfiniBand experience.

  • Experience with GPU-focused hardware/software.

  • Experience with MPI.

  • Automation tooling background (Ansible, Salt, Puppet, etc.).

  • Ethernet and Storage technologies such as Lustre or GPFS.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 216,000 USD - 345,000 USD for Level 4, and 248,000 USD - 396,750 USD for Level 5.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 1, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Bereit, sich bei NVIDIA zu bewerben?
Bei NVIDIA bewerben

Über NVIDIA

NVIDIA pioneered accelerated computing. Today, our AI infrastructure powers global intelligence, transforming every industry. Learn more about NVIDIA .

Alle Jobs bei NVIDIA ansehen →

Ähnliche Jobs

NVIDIA
Senior Software Engineer - Backend Platform
NVIDIA
⚡ Früh bewerben Israel, Raanana Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 13 Std.
NVIDIA
Marketing Specialist, Event and Marcom
NVIDIA
⚡ Früh bewerben Taiwan, Taipei Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 13 Std.
NVIDIA
Circuit Validation Engineer Intern - 2027
NVIDIA
⚡ Früh bewerben China, Shanghai Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 13 Std.
NVIDIA
Senior Software Engineer, AIOps
NVIDIA
⚡ Früh bewerben Israel, Raanana Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 13 Std.
NVIDIA
Senior AI Model Test Developer, SDET
NVIDIA
⚡ Früh bewerben China, Shanghai Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 13 Std.
NVIDIA
Senior Business Intelligence Analyst
NVIDIA
⚡ Früh bewerben Israel, Yokneam Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 1 Tg.
NVIDIA
Software Engineer – Networking Platforms, Diagnostics Tools and Performance
NVIDIA
⚡ Früh bewerben Israel, Tel Aviv Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 1 Tg.
NVIDIA
NPI Engineer, Chip
NVIDIA
⚡ Früh bewerben Israel, Yokneam Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 1 Tg.
NVIDIA
System Software Engineer
NVIDIA
⚡ Früh bewerben Israel, Yokneam Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 1 Tg.

Registrieren für Vorschläge, die auf die von Ihnen geöffneten Jobs und gespeicherten Suchen zugeschnitten sind.

Mehr Jobs bei NVIDIA

Alle Jobs bei NVIDIA ansehen →

Jetzt bewerben
🤖

Moment — langsam

JobsRadar wurde für echte Menschen gebaut, die eine schwere Zeit bei der Jobsuche haben — nicht für automatisierte Anfragen. Sie klicken viel zu schnell und sind jetzt vorübergehend blockiert.

Kommen Sie später wieder. Wenn Sie wirklich auf Jobsuche sind, stehen wir hinter Ihnen — verhalten Sie sich einfach wie ein Mensch.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Verschaffe dir einen Vorsprung bei der Jobsuche.

Tritt unserem Telegram-Kanal bei für das, was dir hilft, die Stelle zu bekommen — Gehaltsbenchmarks, den wöchentlichen Marktpuls und neue Feature-Drops. Kein Spam, nur Signal.

Dem Kanal beitreten — kostenlos