# Senior System Software Engineer, AI Infrastructure

**Company**: NVIDIA
**Location**: Santa Clara, CA
**Experience**: senior
**Job type**: full-time
**Category**: Engineering
**Industry**: Technology

**Apply**: https://nvidia.wd5.myworkdayjobs.com/en-US/NVIDIAExternalCareerSite/job/US-CA-Santa-Clara/Senior-Technical-Marketing-Engineer--AI-Infrastructure_JR2001250?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_afd5c425-542

## Description

NVIDIA is seeking a Senior System Software Engineer to work on AI Infrastructure. You will collaborate with engineering, product, and marketing teams to assess and elevate the company's hardware and software solutions.

**Job Overview** As a Senior System Software Engineer, you will critically assess NVIDIA's hardware and software solutions, focusing on AI infrastructure and performance optimization. Your expertise will help improve the company's developer products and customer experience.

**Responsibilities**

- Run multi-node training/inference jobs on large GPU clusters to assess performance and validate usability.

- Design benchmark suites to showcase NVIDIA hardware, networking, and software stacks.

- Profile deep-learning workloads, identify bottlenecks, and provide optimization guidance.

- Create concise tutorials, scripts, and whitepapers for customers and technical press.

- Analyze competitive solutions and develop data-driven product positioning.

- Present live demos at global conferences such as GTC, CES, and SIGGRAPH.

**Requirements**

- Passion for AI infrastructure and performance optimization.

- 3+ years of experience in software development, technical marketing, or similar roles.

- Bachelor's/Master's degree in Computer Science, Computer Engineering, Electrical Engineering, or a related field.

- Strong Python and C++ skills for AI and HPC work.

- Hands-on experience with multi-node clusters using Slurm, Kubernetes, or cloud CSP clusters.

- Solid understanding of deep learning architectures, PyTorch, and distributed training methods.

- Knowledge of CPU/GPU architecture, CUDA, cuDNN, TensorRT-LLM, Triton, and NCCL.

- Excellent written and verbal communication skills for technical and executive audiences.

**Benefits**

- Equity eligibility

- Comprehensive benefits package

## Skills

### Required
- Python
- C++
- AI infrastructure
- Performance optimization
- Slurm
- Kubernetes
- Deep learning architectures
- PyTorch
- Distributed training methods
- CUDA
- cuDNN
- TensorRT-LLM
- Triton
- NCCL

### Nice to have
- Hands-on experience with HPC clusters
- Public technical blogs or talks
- Exceptional communication skills
- Familiarity with modern LLM architectures
- Expertise in InfiniBand, NVLink, RoCE, RDMA, and collective-comm libraries

---

Source: [Apply at nvidia.wd5.myworkdayjobs.com](https://nvidia.wd5.myworkdayjobs.com/en-US/NVIDIAExternalCareerSite/job/US-CA-Santa-Clara/Senior-Technical-Marketing-Engineer--AI-Infrastructure_JR2001250?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
