# Engineering Manager, Inference Infrastructure

**Company**: Anthropic
**Location**: San Francisco, CA
**Work arrangement**: hybrid
**Experience**: senior
**Job type**: full-time
**Salary**: $405,000-$625,000 USD
**Category**: Engineering
**Industry**: Technology
**Wikidata**: https://www.wikidata.org/wiki/Q116758847

**Apply**: https://job-boards.greenhouse.io/anthropic/jobs/5411560008?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_67a830e6-2ca

## Description

Anthropic's mission is to create reliable, interpretable, and steerable AI systems that are safe and beneficial for users and society.

As an Engineering Manager, Inference Infrastructure, you will lead a team of ML platform, infrastructure, and distributed-systems engineers responsible for building the control plane that coordinates Anthropic's inference fleet.

Key responsibilities:

- Own the technical roadmap for the inference fleet's coordination

- Partner with product, inference engine, performance, and capacity teams to identify and implement improvements

- Develop and maintain quantitative modeling capabilities

- Set technical strategy for the control plane's evolution across heterogeneous hardware and cloud providers

- Run the group's operational backbone, including on-call rotations and incident response

- Create clarity at the seam between API surface, inference engines, capacity planning, and cloud deployment teams

- Develop and retain strong teams, and hire against a high technical bar

Minimum qualifications:

- Engineering management experience leading teams on critical-path production infrastructure at scale

- Deep systems background with expertise in load balancing, scheduling, cluster orchestration, autoscaling, and high-performance networking

- Experience shipping performance or efficiency improvements in large-scale systems

- Experience running production infrastructure with real operational stakes

- Results-oriented, impact-driven approach

- Ability to build strong relationships across team boundaries

- Curiosity about machine learning systems

Preferred qualifications:

- 5+ years of engineering management experience

- Experience with LLM inference serving

- Background in cluster schedulers, autoscalers, load balancers, service meshes, or fleet control planes at scale

- Experience running workloads across multiple clouds or partner platforms

- Familiarity with heterogeneous accelerator fleets

- Experience leading teams at supercomputing or hyperscaler infrastructure scale

The annual compensation range for this role is $405,000-$625,000 USD.

## Skills

### Required
- load balancing
- scheduling
- cluster orchestration
- autoscaling
- high-performance networking
- distributed systems
- machine learning systems

### Nice to have
- LLM inference serving
- cluster schedulers
- autoscalers
- load balancers
- service meshes
- fleet control planes
- heterogeneous accelerator fleets

---

Source: [Apply at job-boards.greenhouse.io](https://job-boards.greenhouse.io/anthropic/jobs/5411560008?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
