# Solutions Architect, AI Factory Infrastructure DevOps

**Company**: NVIDIA
**Work arrangement**: remote
**Experience**: senior
**Job type**: full-time
**Category**: IT
**Industry**: Technology

**Apply**: https://nvidia.wd5.myworkdayjobs.com/en-US/NVIDIAExternalCareerSite/job/US-CA-Remote/Solutions-Architect--AI-Factory-Infrastructure-DevOps_JR2018908?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_0c008820-d82

## Description

NVIDIA seeks a Solutions Architect to join its AI Factory infrastructure deployment team. You will play a key role in introducing advanced GPU products to deployments across data centers and edge computing.

**Job Overview:** As a Solutions Architect, you will help architect and scale high-performance, distributed AI infrastructure on-prem or in the cloud, built with the latest NVIDIA GPU supercomputers for new and existing customers. You will be a technical specialist on GPU and networking products, directly supporting sales account managers to secure build wins.

**Responsibilities:**

- Help architect and scale high-performance, distributed AI infrastructure on-prem or in the cloud, built with the latest NVIDIA GPU supercomputers for new and existing customers.

- Be a technical specialist on GPU and networking products, directly supporting sales account managers to secure build wins.

- Actively establish and nurture technical relationships with engineers, management, and architects at key customer accounts.

- Identify customer architectures and key product requirements in the CSP/OEM AI market to efficiently implement NVIDIA's solutions.

- Provide on-site support to solve hardware and software problems, with a focus on deep learning inference.

- Lead the product through its entire lifecycle, from design-in to end-of-life, ensuring detailed execution and customer satisfaction.

- Actively maintain the NVIDIA side of infrastructure components and collect findings at the customer site.

- Offer technical and sales training to direct sales teams and channel partners.

- The expected travel requirement is approximately 25-30%.

**Requirements:**

- BS or MS in Engineering, Electrical Engineering, Physics, or Computer Science (or equivalent experience).

- 5+ years of work-related experience in high-tech IT companies with experience in NCP, CSP, site reliability, and virtualization technologies (VMware, Linux KVM).

- 4+ years of working experience with Kubernetes, Slurm, Docker, etc.

- Proficiency with AI tools (Claud, Codex, Perplexity, etc.), Redfish, Grafana, and Prometheus.

- Remarkable talent for effectively handling multiple initiatives and priorities.

- Strong time-management and social skills for coordinating complex projects.

- Excellent written and oral communication skills in English, with the ability to collaborate effectively with both management and engineering teams.

**Benefits:**

- Equity

- Benefits

## Skills

### Required
- Kubernetes
- Slurm
- Docker
- AI tools
- Redfish
- Grafana
- Prometheus
- VMware
- Linux KVM

### Nice to have
- NVIDIA systems technology
- DGX
- GB200
- HGX systems
- OEMs in industrial, military, and ruggedized computing spaces

---

Source: [Apply at nvidia.wd5.myworkdayjobs.com](https://nvidia.wd5.myworkdayjobs.com/en-US/NVIDIAExternalCareerSite/job/US-CA-Remote/Solutions-Architect--AI-Factory-Infrastructure-DevOps_JR2018908?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
