New The Skills of Tomorrow: how AI-exposed is every skill in 2026? See the data →
Anthropic

Engineering Manager, Inference Infrastructure

Anthropic
Apply →
hybrid senior full-time $405,000-$625,000 USD San Francisco, CA

First indexed 2 Sept 2026

Description

Anthropic's mission is to create reliable, interpretable, and steerable AI systems that are safe and beneficial for users and society.

As an Engineering Manager, Inference Infrastructure, you will lead a team of ML platform, infrastructure, and distributed-systems engineers responsible for building the control plane that coordinates Anthropic's inference fleet.

Key responsibilities:

  • Own the technical roadmap for the inference fleet's coordination
  • Partner with product, inference engine, performance, and capacity teams to identify and implement improvements
  • Develop and maintain quantitative modeling capabilities
  • Set technical strategy for the control plane's evolution across heterogeneous hardware and cloud providers
  • Run the group's operational backbone, including on-call rotations and incident response
  • Create clarity at the seam between API surface, inference engines, capacity planning, and cloud deployment teams
  • Develop and retain strong teams, and hire against a high technical bar

Minimum qualifications:

  • Engineering management experience leading teams on critical-path production infrastructure at scale
  • Deep systems background with expertise in load balancing, scheduling, cluster orchestration, autoscaling, and high-performance networking
  • Experience shipping performance or efficiency improvements in large-scale systems
  • Experience running production infrastructure with real operational stakes
  • Results-oriented, impact-driven approach
  • Ability to build strong relationships across team boundaries
  • Curiosity about machine learning systems

Preferred qualifications:

  • 5+ years of engineering management experience
  • Experience with LLM inference serving
  • Background in cluster schedulers, autoscalers, load balancers, service meshes, or fleet control planes at scale
  • Experience running workloads across multiple clouds or partner platforms
  • Familiarity with heterogeneous accelerator fleets
  • Experience leading teams at supercomputing or hyperscaler infrastructure scale

The annual compensation range for this role is $405,000-$625,000 USD.

This listing is enriched and indexed by YubHub. To apply, use the employer's original posting: https://job-boards.greenhouse.io/anthropic/jobs/5411560008