# Senior Software Engineer II, Applied Training

**Company**: CoreWeave
**Location**: Sunnyvale, CA
**Experience**: senior
**Job type**: full-time
**Salary**: $182,000 - $242,000
**Category**: Engineering
**Industry**: Technology

**Apply**: https://job-boards.greenhouse.io/coreweave/jobs/4692028006?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_fe1d1ee6-d46

## Description

### Job Description

CoreWeave is The Essential Cloud for AI. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence.

### What You'll Do

As a Senior Software Engineer II, Applied Training, you will be an early member of a small team responsible for developing a Kubernetes-native research cluster platform or a sandbox client for agentic training and evaluation. Your goal will be to provide every CoreWeave customer with research infrastructure that currently only exists inside frontier labs.

### Responsibilities

- Contribute to the roadmap for Applied Training and identify what unlocks new workloads

- Work closely with customers and other teams to build cloud-native primitives

- Design and build a complete research cluster experience, including CLI, job configuration schema, and Kubernetes operators

- Own the Python SDK for sandbox infrastructure and enable RL training runs at scale

- Write documentation for running popular OSS training frameworks on CoreWeave

- Collaborate with infrastructure teams and customers to understand their needs

### Requirements

- 8–12+ years of experience building distributed systems, ML infrastructure, or developer platforms

- Real Kubernetes experience with custom controllers, operators, scheduling, and workload orchestration

- Understanding of what makes researchers productive, including code distribution and fast iteration cycles

- Familiarity with training, including distributed job scheduling and rank initialization

- Experience shipping infrastructure that others rely on daily

- Strong communication skills to work with customers and translate researcher complaints into system designs

### Preferred

- Experience building internal ML platforms or research clusters

- Familiarity with agentic AI, RL training, and sandbox isolation

- Background with Slurm, Ray, or similar workload orchestration

- Experience with container runtimes, isolation, or serverless platforms

- OSS contributions to Kubernetes SIGs, Ray, PyTorch, or similar

### Benefits

- Competitive salary ($182,000 - $242,000)

- Comprehensive benefits program, including medical, dental, and vision insurance

- Flexible spending account and health savings account

- Tuition reimbursement and employee stock purchase program

- Mental wellness benefits and family-forming support

- Paid parental leave and flexible PTO

- Catered lunch and casual work environment

## Skills

### Required
- Kubernetes
- distributed systems
- ML infrastructure
- Python
- cloud-native primitives

### Nice to have
- agentic AI
- RL training
- sandbox isolation
- Slurm
- Ray
- container runtimes
- serverless platforms

---

Source: [Apply at job-boards.greenhouse.io](https://job-boards.greenhouse.io/coreweave/jobs/4692028006?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
