Description
NVIDIA is seeking an Infrastructure Solutions Architect to help revolutionize computing. The successful candidate will serve as a senior technical point of contact for OEM Federal deployments of NVIDIA GPU-accelerated platforms.
Responsibilities:
- Serve as the senior technical point of contact for OEM Federal deployments of NVIDIA GPU-accelerated platforms, including PCIe GPUs, HGX, and MGX systems, addressing any blocking issues.
- Own issues end-to-end: take responsibility for customer issues from start to finish , prioritizing platform, firmware, and workload problems.
- Conduct failure analysis and establish diagnostic, RMA eligibility, and replacement procedures for Federal deployments.
- Navigate Federal security frameworks, including FedRAMP, IL5/IL6, and air-gapped environments.
- Partner with OEM Federal support and services teams to reduce blocking issues and unnecessary dispatches.
Requirements:
- U.S. citizen with an active U.S. Government TS/SCI clearance.
- BS or MS in Computer Engineering, Electrical Engineering, Computer Science, or equivalent experience.
- 5+ years of engineering experience on multi-GPU platforms, including at least 3 years in a customer-facing support, field engineering, or blocking issue role.
- Strong system software expertise across firmware, BIOS, kernel, drivers, and operating systems.
- Expert knowledge of data center infrastructure , x86/ARM systems, high-performance storage, and low-latency networking.
- Professional-level communication and organizational skills.
Preferred qualifications:
- Dell ecosystem expertise: proven experience with Dell PowerEdge GPU servers and Dell's management software.
- Federal mission experience: direct experience supporting DoD, Intelligence Community, or Civilian agency deployments.
- Programming depth: proficiency in Python for building custom diagnostic and analysis tooling, and C/C++ for platform OS, firmware, and driver work.
- Cluster & HPC technologies: containerized and scheduled environments, and upper layer protocols such as NCCL and MPI.
- Performance analysis: experience analyzing performance of distributed GPU-accelerated workloads.
Benefits:
- Competitive salary: $152,000 - $241,500 (Level 3) or $184,000 - $287,500 (Level 4).
- Equity and benefits package.
This listing is enriched and indexed by YubHub. To apply, use the employer's original posting:
https://nvidia.wd5.myworkdayjobs.com/en-US/NVIDIAExternalCareerSite/job/US-TX-Austin/Infrastructure-Solutions-Architect---OEM_JR2025080