# AI and ML Infra Software Engineer, GPU Clusters - New College Grad 2026

**Company**: NVIDIA
**Location**: Santa Clara, CA
**Experience**: entry
**Job type**: full-time
**Category**: Engineering
**Industry**: Technology

**Apply**: https://nvidia.wd5.myworkdayjobs.com/en-US/NVIDIAExternalCareerSite/job/US-CA-Santa-Clara/AI-and-ML-Infra-Software-Engineer--GPU-Clusters---New-College-Grad-2026_JR2021591?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_8769b4bc-e87

## Description

NVIDIA is seeking an AI/ML Infrastructure Software Engineer to join its Hardware Infrastructure team. The successful candidate will play a crucial role in boosting productivity for researchers by implementing advancements across the entire stack, focusing on GPU Clusters.

**Job Overview:** As an AI/ML Infrastructure Software Engineer at NVIDIA, you will work closely with customers to identify and resolve infrastructure gaps, enabling innovative AI and ML research. You will collaborate with diverse teams to build a seamless and coordinated AI/ML infrastructure ecosystem, monitor and optimize infrastructure performance, and stay up-to-date with the latest advancements in AI/ML technologies.

**Responsibilities:**

- Collaborate with AI and ML research teams to understand their infrastructure needs and obstacles.

- Monitor and optimize infrastructure performance, ensuring high availability, scalability, and efficient resource utilization.

- Define and improve measures of AI researcher efficiency, aligning actions with measurable results.

- Collaborate with researchers, data engineers, and DevOps professionals to build a coordinated AI/ML infrastructure ecosystem.

- Stay current with the latest advancements in AI/ML technologies and promote their implementation.

**Requirements:**

- Recent graduate with a MS, PhD or equivalent experience in Computer Science or related field.

- Proven experience in AI/ML and HPC workloads and infrastructure.

- Hands-on experience with High Performance Computing (HPC) grade infrastructure, accelerated computing, storage, scheduling & orchestration, high-speed networking, and container technologies.

- Expertise in running and optimizing large-scale distributed training workloads using PyTorch, NeMo, or JAX.

- Proficiency in programming & scripting languages such as Python, Go, Bash, and familiarity with cloud computing platforms.

- Excellent communication and collaboration skills.

**Benefits:** NVIDIA provides competitive salaries and a comprehensive benefits package, including equity.

## Skills

### Required
- AI/ML
- HPC
- GPU
- accelerated computing
- storage
- scheduling & orchestration
- high-speed networking
- container technologies
- PyTorch
- NeMo
- JAX
- Python
- Go
- Bash
- cloud computing

---

Source: [Apply at nvidia.wd5.myworkdayjobs.com](https://nvidia.wd5.myworkdayjobs.com/en-US/NVIDIAExternalCareerSite/job/US-CA-Santa-Clara/AI-and-ML-Infra-Software-Engineer--GPU-Clusters---New-College-Grad-2026_JR2021591?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
