# Senior Solutions Architect

**Company**: NVIDIA
**Experience**: senior
**Job type**: full-time
**Salary**: 292,500 PLN - 507,000 PLN
**Category**: IT
**Industry**: Technology

**Apply**: https://nvidia.wd5.myworkdayjobs.com/en-US/NVIDIAExternalCareerSite/job/France-Remote/Senior-Solutions-Architect---Large-Scale-AI-Training_JR2024427?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_7f982971-b49

## Description

We are seeking a Senior Solutions Architect with expertise in large-scale distributed training and alignment of neural networks. As a primary technical expert for AI organizations, you will support top EMEA AI natives and research institutions in training and post-training AI models, such as large Mixture-of-Experts (MoE) models.

**Responsibilities:**

- Build and manage strategic technical relationships with leading EMEA AI model builders developing large-scale foundation models.

- Collaborate closely with customers to define the software stack and infrastructure required for large-scale training and post-training workflows, including Reinforcement Learning (RL).

- Serve as the go-to expert on distributed training strategies, guiding customers through efficient large-scale training and post-training recipes using Deep Learning Frameworks such as PyTorch, Megatron-LM, or NeMo (RL/Gym).

- Help customers optimize training/fine-tuning efficiency at scale, covering GPU utilization, communication overlap, memory management at scale.

- Collaborate with NVIDIA's product and research teams to present customer needs.

- Animate the developer community by building or supporting hackathons, demos, technical conferences.

**Requirements:**

- MS or PhD in Computer Science, Engineering, or equivalent experience.

- Over 7 years of practical experience in distributed AI training, including direct involvement with HPC and/or AI environments with multi-node GPU clusters.

- Solid understanding of training infrastructure and how it affects efficiency and scalability.

- Strong proficiency with Megatron-LM, NeMo, or equivalent distributed training frameworks.

- Excellent communication skills with an ability to engage both research scientists and infrastructure engineers.

**Preferred Qualifications:**

- Experience in fine-tuning with Reinforcement Learning (RLVR, RLHF) at scale.

- Experience with LatentMoE, expert load balancing, and speculative decoding for MoE inference.

- Prior experience in an AI Datacenter/HPC center, national lab, or frontier AI lab environment.

- Published work or open-source contributions in distributed training.

NVIDIA offers highly competitive salaries and a comprehensive benefits package. The base salary range for this position in Poland is 292,500 PLN - 507,000 PLN.

## Skills

### Required
- Distributed AI training
- Deep Learning Frameworks
- GPU utilization
- Communication overlap
- Memory management
- Megatron-LM
- NeMo
- PyTorch

### Nice to have
- Reinforcement Learning
- LatentMoE
- Expert load balancing
- Speculative decoding
- AI Datacenter
- HPC center
- National lab
- Frontier AI lab environment
- Open-source contributions

---

Source: [Apply at nvidia.wd5.myworkdayjobs.com](https://nvidia.wd5.myworkdayjobs.com/en-US/NVIDIAExternalCareerSite/job/France-Remote/Senior-Solutions-Architect---Large-Scale-AI-Training_JR2024427?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
