# Technical Program Manager, Inference

**Company**: CoreWeave
**Location**: Livingston, NJ / New York, NY / Sunnyvale, CA / Bellevue, WA
**Experience**: senior
**Job type**: full-time
**Salary**: $198,000 to $264,000
**Category**: Engineering
**Industry**: Technology

**Apply**: https://job-boards.greenhouse.io/coreweave/jobs/4693164006?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_98038fea-817

## Description

CoreWeave is The Essential Cloud for AI. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence.

## What You'll Do:

The AI/ML TPM team owns delivery and execution across CoreWeave's AI/ML Platform Services organization. The team partners closely with Product, Engineering, Research, Infrastructure, and Go-to-Market teams to deliver scalable, reliable, and high-performance platforms that support the full AI lifecycle.

As a Technical Program Manager focused on inference, you will lead complex, cross-functional programs spanning inference platform delivery, customer onboarding, launch readiness, and runtime optimization.

## In this role, you will:

- Drive end-to-end program management for inference platform initiatives spanning reliability, customer onboarding, launch readiness, and runtime optimization

- Lead cross-functional programs for customer onboarding across dedicated and serverless inference offerings, ensuring clear ownership, launch criteria, and readiness for strategic customer use cases

- Drive launch readiness for new inference capabilities by aligning teams around real customer outcomes, supportability, and end-to-end validation

- Partner with engineering and product to define and deliver roadmap outcomes for latency, throughput, uptime, operational quality, and price-performance

- Coordinate multi-team execution across platform, infrastructure, and customer-facing teams to deliver reliable and scalable inference services

- Build and operationalize success metrics, dashboards, launch gates, and review cadences to measure service reliability, onboarding readiness, efficiency, and quality across the inference stack

- Establish repeatable processes for release validation, performance regression tracking, launch management, and postmortem follow-through

- Help unify operational processes, support mechanisms, and execution visibility across inference deployment models and customer onboarding paths

- Create strong communication channels between Engineering, Product, Infrastructure, and Go-to-Market teams to align priorities and deliver predictable, high-impact outcomes

## Who You Are:

- Bachelor's degree in a technical field or equivalent practical experience

- 8+ years of technical program management experience in distributed systems, cloud infrastructure, or AI/ML platform engineering

- Proven experience driving large-scale infrastructure or platform programs from concept to production in complex, cross-functional environments

- Strong technical fluency in distributed inference systems, GPU compute, cloud-native architectures, and performance optimization

- Demonstrated success driving measurable improvements in reliability, performance, operational readiness, or customer delivery

- Excellent written and verbal communication skills, with the ability to align engineering, product, infrastructure, and customer-facing stakeholders around shared goals

- Experience with inference-serving systems, model onboarding workflows, rollout strategies, and observability tooling

- Familiarity with launch readiness, supportability, incident follow-through, and release validation for production infrastructure or platform services

- Understanding of customer onboarding for technical products, especially where platform capabilities, infrastructure readiness, and support processes must align for launch

- Experience operating in high-growth environments where roadmap execution, reliability expectations, and customer commitments must be managed in parallel

## Why CoreWeave?

At CoreWeave, we work hard, have fun, and move fast! We're in an exciting stage of hyper-growth that you will not want to miss out on. We're not afraid of a little chaos, and we're constantly learning.

## What We Offer

- Medical, dental, and vision insurance - 100% paid for by CoreWeave

- Company-paid Life Insurance

- Voluntary supplemental life insurance

- Short and long-term disability insurance

- Flexible Spending Account

- Health Savings Account

- Tuition Reimbursement

- Ability to Participate in Employee Stock Purchase Program (ESPP)

- Mental Wellness Benefits through Spring Health

- Family-Forming support provided by Carrot

- Paid Parental Leave

- Flexible, full-service childcare support with Kinside

- 401(k) with a generous employer match

- Flexible PTO

- Catered lunch each day in our office and data center locations

- A casual work environment

- A work culture focused on innovative disruption

## Skills

### Required
- technical program management
- distributed systems
- cloud infrastructure
- AI/ML platform engineering
- inference-serving systems
- model onboarding workflows
- rollout strategies
- observability tooling

---

Source: [Apply at job-boards.greenhouse.io](https://job-boards.greenhouse.io/coreweave/jobs/4693164006?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
