Description
NVIDIA is seeking a Manager, Infrastructure Engineering and DevOps to lead a high-impact infrastructure engineering team. The team develops and maintains engineering infrastructure solutions that enable R&D teams to provision, validate, test, debug, and recover complex server and networking environments at scale.
In this role, you will lead a team that provides infrastructure platforms and hands-on engineering support for internal customers across firmware, driver, hardware, software, and verification organizations. The position combines deep technical leadership with people management, execution ownership, customer support, and operational excellence in a fast-paced R&D environment.
Responsibilities:
- Lead and grow an infrastructure engineering team responsible for bare-metal provisioning, VM infrastructure, server fleet automation, CI/CD infrastructure, customer-facing debug support, and high-performance networking environments.
- Own the team's technical roadmap, priorities, execution plans, and delivery commitments across multiple infrastructure initiatives, while balancing long-term platform improvements with day-to-day customer needs.
- Drive ownership of VM box inventory and lifecycle management across many Linux distributions.
- Build infrastructure capabilities that enable engineering and verification teams to run provisioning, testing, validation, and debug workflows efficiently and reliably.
- Lead customer support, debug, and optimization of internal customer flows.
- Guide complex system debug and recovery in a firmware R&D environment.
- Provide technical leadership for Linux-based automation platforms.
- Partner closely with firmware, driver, hardware, software, cloud, and verification teams to define requirements and deliver infrastructure solutions.
Requirements:
- B.Sc. in Computer Engineering, Computer Science, Electrical Engineering, or a related technical field, or equivalent experience.
- 8+ overall years of experience in Linux systems administration, infrastructure automation, DevOps, system software, firmware infrastructure, lab infrastructure, or related engineering domains.
- 3+ years of experience leading or managing engineering teams, technical projects, or cross-functional infrastructure initiatives.
- Strong technical background in Linux environments.
- Hands-on experience designing, implementing, and debugging automation software using Python, scripting, CI/CD workflows, and modern software development practices.
- Experience managing infrastructure across multiple Linux distributions.
- Proven ability to support internal customers in complex technical environments.
- Strong people leadership skills.
Benefits:
- Competitive salaries
- Generous benefits package