# Sr. Engineering Manager, Inference

**Company**: CoreWeave
**Location**: Livingston, NJ / New York, NY / San Francisco, CA / Sunnyvale, CA / Bellevue, WA / Remote - US
**Work arrangement**: remote
**Experience**: senior
**Job type**: full-time
**Salary**: $188,000 to $303,000
**Category**: Engineering
**Industry**: Technology

**Apply**: https://job-boards.greenhouse.io/coreweave/jobs/4626644006?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_fe4f9354-a0d

## Description

CoreWeave is The Essential Cloud for AI. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence.

As a Senior Engineering Manager for the AI/ML Platform team, you will lead the group responsible for productizing and operating our inference offering. In partnership with CoreWeave AI platform team, you will ensure the service becomes a polished, reliable, developer-friendly product within the AI/ML Platform.

## Responsibilities

- Lead and grow the engineering team responsible for evolving and operating the W&B Inference product, focusing on service reliability, orchestration, operational maturity, and developer experience.

- Drive execution on roadmap initiatives in close partnership with Product, ensuring that platform capabilities are delivered predictably, robustly, and with measurable customer impact.

- Own the engineering processes for the inference productization layer, including incident response, operational readiness, release management, observability, and engineering quality.

- Partner with CoreWeave’s infrastructure teams to integrate capabilities from the underlying inference platform into a cohesive, user-facing product.

- Guide the design and delivery of application-layer enhancements such as tracing and tool calls handling.

- Ensure the inference offering meets high standards of reliability, usability, compliance, and performance, balancing trade-offs across cost, operational load, and architectural constraints.

- Create clarity in complex, cross-functional environments, ensuring strong communication, aligned priorities, and smooth execution across multiple teams.

- Build a culture of ownership, technical excellence, and continuous improvement within the engineering team.

## Who You Are:

- 7+ years of experience in software engineering, including 3+ years managing or leading engineering teams responsible for distributed systems, developer platforms, or large-scale services.

- Fluent in concepts related to distributed compute services, including observability, autoscaling patterns, reliability engineering, API design, and service operations.

- Comfortable with competing priorities and making principled trade-offs among latency, reliability, cost, and development velocity.

- Skilled at leading engineering teams through ambiguous, multi-stakeholder projects with strong communication and alignment.

- Deep empathy for ML practitioners and platform developers, with a drive to improve reliability, reduce friction, and elevate the developer experience.

## Preferred

- Background in high-scale systems, real-time APIs, or cloud infrastructure, ideally with exposure to model-serving or inference-adjacent domains.

- Background in related platform domains such as IAM, billing/metering, observability systems, or deployment tooling.

## Why Us?

We work hard, have fun, and move fast! We’re in an exciting stage of hyper-growth that you will not want to miss out on. We’re not afraid of a little chaos, and we’re constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values:

- Be Curious at Your Core

- Act Like an Owner

- Empower Employees

- Deliver Best-in-Class Client Experiences

- Achieve More Together

The base salary range for this role is $188,000 to $303,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location.

## What We Offer

- Medical, dental, and vision insurance - 100% paid for by CoreWeave

- Company-paid Life Insurance

- Voluntary supplemental life insurance

- Short and long-term disability insurance

- Flexible Spending Account

- Health Savings Account

- Tuition Reimbursement

- Ability to Participate in Employee Stock Purchase Program (ESPP)

- Mental Wellness Benefits through Spring Health

- Family-Forming support provided by Carrot

- Paid Parental Leave

- Flexible, full-service childcare support with Kinside

- 401(k) with a generous employer match

- Flexible PTO

- Catered lunch each day in our office and data center locations

- A casual work environment

- A work culture focused on innovative disruption

## Skills

### Required
- distributed systems
- developer platforms
- large-scale services
- observability
- autoscaling patterns
- reliability engineering
- API design
- service operations

### Nice to have
- high-scale systems
- real-time APIs
- cloud infrastructure
- model-serving
- inference-adjacent domains
- IAM
- billing/metering
- observability systems
- deployment tooling

---

Source: [Apply at job-boards.greenhouse.io](https://job-boards.greenhouse.io/coreweave/jobs/4626644006?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
