Jobs Companies NVIDIA Engineering Manager, Network System Validation

About this Engineering Manager, Network System Validation role at NVIDIA

NVIDIA · Onsite · Israel, Yokneam

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. 

Within NVIDIA, the Networking Business Unit (NBU) builds the high-speed interconnect — Ethernet, InfiniBand, NVLink, and BlueField DPUs — that switches thousands of GPUs into a single AI supercomputer, moving data at the scale and speed the most demanding workloads require. NVIDIA is looking for an Engineering Manager to join our Network System Validation group. You will work on system validating advanced networking solutions across NVIDIA complex AI cluster environments. The group is a high-performance engineering force that treats validation as a first-class software problem. We build systems, frameworks, and benchmarks that prove our network's correctness and performance at scale. In this role you will lead the validation direction and engineering excellence of one of our technology validation teams. This is a management role for a technology leader who can own the technical roadmap and execution, mentoring a team of high-performance engineers, and push NVIDIA's network to its speed-of-light limits. This role combines the development of methodologies and automation tools with system validation, performance analysis, and investigation of cutting-edge AI networking technologies at scale. 

What you’ll be doing:

  • Lead, mentor, and coach a team of software development and system validation engineers. 

  • Review system and product requirements, design validation methodologies, develop comprehensive test plans, functional and performance, for networking technologies in large-scale AI cluster solutions 

  • Develop and maintain benchmarks, automation tools and scripts for test execution, environment setup, log collection, and data analysis. 

  • Lead end-to-end investigation of complex issues by reproducing real-world scenarios, analyzing logs, telemetry, packet captures, and system metrics to identify functional issues and performance bottlenecks, triaging problems across the hardware and software stack, and driving them to root cause and resolution 

  • Collaborate deeply with software and hardware development teams to debug networking technologies, including NCCL, RoCE, RDMA, and related software components using targeted experiments and code inspection 

  • Profile and research AI training and inference workloads, correlating application behavior with network and system telemetry to identify scalability and performance limitations 

  • Document findings, communicate technical results, and continuously improve validation methodologies, automation environments, and engineering processes 

  • Foster a team culture centered on software quality, accountability, and technical excellence 

What we need to see:

  • B.Sc. / B.A. in Computer Science, Electrical Engineering, or equivalent experience 

  • 8+ overall years of experience in networking, system validation, or related domains 

  • 3+ years of experience leading software or system development team   

  • Proven experience debugging complex production systems by forming hypotheses, designing experiments, and driving issues to root cause 

  • Strong scripting and automation experience using Python, Bash, and/or Ansible 

  • Ability to read, debug, and reason about C/C++ code (Rust or Go a plus) 

  • Ability to drive technical alignment across teams, communicate tradeoffs clearly, and make high-quality architectural decisions at speed 

  • Advance AI-driven approaches to test automation: intelligent scenario generation, LLM-augmented root-cause analysis, and autonomous validation pipelines 

Ways to stand out from the crowd:

  • Experience with large-scale clusters or distributed systems 

  • Familiarity with NVIDIA networking solutions (ConnectX, SpecX, BlueField) 

  • Background in performance analysis, Kubernetes, or cloud environments 

  • Background in chaos testing, fault injection, or simulation systems 

We have some of the most forward-thinking and hardworking people working for us. If you're creative and autonomous, we want to hear from you! NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, disability status or any other characteristic protected by law.

Ready to apply to NVIDIA?
Apply to NVIDIA

About NVIDIA

NVIDIA pioneered accelerated computing. Today, our AI infrastructure powers global intelligence, transforming every industry. Learn more about NVIDIA .

See all jobs at NVIDIA →

Similar jobs

NVIDIA
Manager, Infrastructure Engineering and DevOps
NVIDIA
⚡ Apply early Israel, Yokneam Onsite
● New 👁 Seen ✓ Applied 3w ago
NVIDIA
Software Engineering Manager, Networking Tools Team
NVIDIA
⚡ Apply early Israel, Yokneam Onsite
● New 👁 Seen ✓ Applied 4w ago
NVIDIA
Manager, Networking Firmware Engineering
NVIDIA
⚡ Apply early Israel, Yokneam Onsite
● New 👁 Seen ✓ Applied 1mo ago
Workana
Engineering Manager
Workana
⚡ Apply early Mexico · location restricted
● New 👁 Seen ✓ Applied 1h ago
RateHawk
Business Development Manager, South Cone
RateHawk
⚡ Apply early Santiago, Santiago Metropolita... · location restricted
● New 👁 Seen ✓ Applied 1h ago
Wrapbook
Manager, Engineering - Frontend
Wrapbook
⚡ Apply early Remote - US & Canada · location restricted $183,000–$283,500
● New 👁 Seen ✓ Applied 1h ago
Sydecar
Senior Engineering Manager
Sydecar
⚡ Apply early San Francisco Office - Hybrid Hybrid $200,000–$230,000
● New 👁 Seen ✓ Applied 1h ago
Statista
Junior Corporate Development Manager (m/f/d)
Statista
⚡ Apply early Hamburg or Berlin Hybrid
● New 👁 Seen ✓ Applied 1h ago
Snowflake
Sr. Partner Development Manager, Google Cloud
Snowflake
⚡ Apply early US-CA-Menlo Park Onsite $196,000–$257,250
● New 👁 Seen ✓ Applied 1h ago

Sign up for suggestions tailored to the jobs you open and the searches you save.

More jobs at NVIDIA

See all jobs at NVIDIA →

Apply now
🤖

Whoa — hold up

JobsRadar was built for real people having a rough time in their job search — not for automated requests. You're clicking way too fast and you're now temporarily blocked.

Come back later. If you're genuinely job hunting, we've got your back — just act like a human.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Get an edge on your job hunt.

Join our Telegram channel for the stuff that helps you land the role — salary benchmarks, the weekly market pulse, and new-feature drops. No spam, just signal.

Join the channel — it's free