New The Skills of Tomorrow: how AI-exposed is every skill in 2026? See the data →
NVIDIA

Senior Software Engineer, Network System Validation

NVIDIA
Apply →
senior full-time Yokneam, Israel

First indexed 26 Aug 2026

Description

NVIDIA is seeking a Senior Software Engineer to join the Network System Validation group. The successful candidate will lead the validation of advanced networking solutions across complex AI cluster environments.

Job Summary: As a Technical Lead, you will develop validation methodologies and automation frameworks, debug, analyze performance, and work on cutting-edge AI networking technologies at scale.

Responsibilities:

  • Review system and product requirements, design validation methodologies, and develop comprehensive test plans for networking technologies in large-scale AI cluster solutions.
  • Develop and maintain benchmarks, automation tools, and scripts for test execution, environment setup, log collection, and data analysis.
  • Lead end-to-end investigation of complex issues, reproduce real-world scenarios, analyze logs, and drive issues to root cause and resolution.
  • Read and understand source code (C/C++/Python) to investigate defects, validate fixes, and improve logging, instrumentation, and debugging capabilities.
  • Collaborate with software and hardware development teams to debug networking technologies.
  • Profile and research AI training and inference workloads, correlating application behavior with network and system telemetry.
  • Document findings, communicate technical results, and continuously improve validation methodologies, automation environments, and engineering processes.

Requirements:

  • B.Sc. / B.A. in Computer Science, Electrical Engineering, or equivalent experience.
  • 8+ years of experience in networking, system validation, or related domains.
  • Proven experience debugging complex production systems.
  • Ability to read, debug, and reason about C/C++ code.
  • Strong scripting and automation experience using Python, Bash, and/or Ansible.
  • Deep understanding of distributed systems.
  • Ability to drive technical alignment across teams and make high-quality architectural decisions.

Nice to Have:

  • Experience with large-scale clusters or distributed systems.
  • Familiarity with NVIDIA networking solutions.
  • Background in performance analysis, Kubernetes, or cloud environments.
  • Background in chaos testing, fault injection, or simulation systems.