# Staff AI Infrastructure Engineer -

**Company**: Anduril Industries
**Location**: Costa Mesa, California, United States; Seattle, Washington, United States; Washington, District of Columbia, United States
**Experience**: staff
**Job type**: full-time
**Salary**: $220,000-$292,000 USD
**Category**: Engineering
**Industry**: Technology

**Apply**: https://job-boards.greenhouse.io/andurilindustries/jobs/5212860007?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_cc05130a-f0a

## Description

Anduril Industries is a defense technology company with a mission to transform U.S. and allied military capabilities with advanced technology.

We are looking for a founding Staff AI Infrastructure Engineer to architect, build, and scale the end-to-end machine learning platform that powers Anduril's autonomous systems.

**Responsibilities**

- Architect the ML Platform: Design, build, and maintain our foundational training, orchestration, and experimentation infrastructure to support state-of-the-art model development.

- Optimize AI Researcher Velocity: Actively identify, measure, and eliminate bottlenecks in the ML research lifecycle.

- Build highly automated tools for hyperparameter tuning, model profiling, and experimentation tracking.

- Build Data Pipelines at Scale: Design and scale robust, high-performance ETL pipelines capable of processing terabytes of multi-modal data.

- Deploy Model Serving Infrastructure: Architect high-throughput, low-latency model serving frameworks optimized for both scalable cloud environments and air-gapped, resource-constrained tactical edge environments.

- Implement Evaluation & Alignment Loops: Build robust, automated pipelines for continuous evaluation, model validation, and reinforcement learning alignment loops.

**Requirements**

- Production Infrastructure Experience: 7+ years of software engineering experience with a proven track record of designing, building, and operating production-scale machine learning systems and platforms (MLOps).

- Strong Programming & Systems Design: Proficient in Python, Go, C++, or similar backend languages.

- GPU & Compute Orchestration: Deep experience with containerized deployments, GPU scheduling/orchestration, and distributed training frameworks.

- Data Architecture: Hands-on experience building distributed data pipelines and managing massive datasets.

- Strategic Technical Leadership: Experience setting technical direction, leading complex system migrations, and mentoring senior engineers.

- Clearance Eligibility: Eligible to obtain and maintain an active U.S. Top Secret security clearance.

**Benefits**

- Competitive salary range: $220,000-$292,000 USD

- Highly competitive equity grants

- Top-tier benefits for full-time employees

## Skills

### Required
- Python
- Go
- C++
- MLOps
- GPU scheduling
- distributed training frameworks
- data architecture
- strategic technical leadership

### Nice to have
- Secure & Edge Environments
- LLM/GenAI Infra
- Hardware Profiling
- Multi-Cluster & Multi-Tenant Platforms
- Custom Accelerators
- Model Observability at Scale

---

Source: [Apply at job-boards.greenhouse.io](https://job-boards.greenhouse.io/andurilindustries/jobs/5212860007?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
