# Engineering Manager, Infrastructure

**Company**: Scale AI
**Location**: London, UK
**Experience**: senior
**Job type**: full-time
**Category**: Engineering
**Industry**: Technology
**Wikidata**: https://www.wikidata.org/wiki/Q112629176

**Apply**: https://job-boards.greenhouse.io/scaleai/jobs/4719479005?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_ca4a91f5-b8b

## Description

At Scale AI, our mission is to accelerate the development of AI applications.

We are looking for an Infrastructure Engineering Manager to help shape the future of AI-powered applications. In this role, you'll bridge the gap between AI research and production, leading a team that is turning innovative prototypes into scalable, high-performance enterprise solutions.

Responsibilities:

- Define and execute the infrastructure roadmap aligned with business and engineering priorities.

- Lead the design and implementation of scalable, secure, and reliable infrastructure systems.

- Set and maintain SLAs/SLOs for platform uptime, performance, and developer experience.

- Manage the engineering team and drive technical delivery.

- Design, build, and optimize backend services for advanced AI-driven applications, focusing on AI agents, evaluation tooling, and automation.

- Influence the culture, values, and processes of a growing engineering team.

- Inspire and mentor engineers.

- Work closely with product, security, and engineering leadership to align on goals and priorities.

Requirements:

- At least 5 years of relevant experience and at least 2+ years of experience managing infrastructure or platform teams.

- Proven experience with cloud platforms such as AWS, GCP, Azure or OCI.

- Proven experience with kubernetes deployments on on-prem infrastructure.

- Deep understanding of CI/CD pipelines, infrastructure-as-code, and container orchestration.

- Experience managing production environments with high availability, reliability, and scalability requirements.

- Familiarity with monitoring, alerting, and incident response best practices.

- Experience working with modern developer platforms and internal tooling to improve engineering velocity.

- Solid foundation and real-world experience in network engineering.

## Skills

### Required
- cloud platforms
- kubernetes deployments
- CI/CD pipelines
- infrastructure-as-code
- container orchestration
- network engineering
- monitoring
- alerting
- incident response

---

Source: [Apply at job-boards.greenhouse.io](https://job-boards.greenhouse.io/scaleai/jobs/4719479005?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
