# Systems Generalist, GPT Infrastructure

**Company**: OpenAI
**Location**: San Francisco
**Work arrangement**: hybrid
**Experience**: senior
**Job type**: Full time
**Salary**: $293K - $385K
**Category**: Engineering
**Industry**: Technology
**Wikidata**: https://www.wikidata.org/wiki/Q124605186

**Apply**: https://jobs.ashbyhq.com/openai/78c2a68b-cc77-4c62-8891-96afb603650a?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_14706d55-98f

## Description

We are seeking an experienced systems generalist who can work comfortably across the stack to help build an automated inference optimization platform.

**Compensation**

- $293K – $385K • Offers Equity

The base pay offered may vary depending on multiple individualized factors, including market location, job-related knowledge, skills, and experience.

**Benefits**

- Medical, dental, and vision insurance for you and your family, with employer contributions to Health Savings Accounts

- Pre-tax accounts for Health FSA, Dependent Care FSA, and commuter expenses (parking and transit)

- 401(k) retirement plan with employer match

- Paid parental leave (up to 24 weeks for birth parents and 20 weeks for non-birthing parents), plus paid medical and caregiver leave (up to 8 weeks)

- Paid time off: flexible PTO for exempt employees and up to 15 days annually for non-exempt employees

- 13+ paid company holidays, and multiple paid coordinated company office closures throughout the year for focus and recharge, plus paid sick or safe time

- Mental health and wellness support

- Employer-paid basic life and disability coverage

- Annual learning and development stipend to fuel your professional growth

- Daily meals in our offices, and meal delivery credits as eligible

- Relocation support for eligible employees

**Key Responsibilities**

- Design, build, and operate durable APIs and control-plane services for multi-hour or multi-day optimization campaigns

- Build secure partner-side runner and grader software that can compile, execute, verify, and benchmark candidate artifacts on third-party accelerator hardware

- Integrate hardware profiles, ISA and toolchain context, compilers, runtimes, and inference-serving engines into a repeatable optimization workflow

- Turn research prototypes into reliable product surfaces with clear contracts, debuggable failure modes, reproducible outputs, and excellent developer ergonomics

- Develop correctness and performance evaluation systems spanning latency, throughput, memory use, utilization, and cost efficiency

- Build artifact, provenance, and qualification workflows that make optimized kernels, binaries, configurations, and reports safe to review and deploy

- Collaborate with Research, Inference Engineering, Infrastructure, Security, Product, and Strategic Partnerships to deliver production-ready solutions

- Drive technical architecture and execution across ambiguous, cross-functional initiatives that connect OpenAI systems with partner environments

**Basic Qualifications**

- 8+ years of professional software engineering experience building large-scale distributed systems, infrastructure platforms, or cloud services, or equivalent depth of experience

- Strong programming skills in one or more of C++, Python, Go, or Rust

- Experience designing and operating highly available backend systems, APIs, job orchestration systems, or durable workflows for production workloads

- Strong understanding of distributed systems, Linux, networking, storage, containers, and modern cloud architectures

- Experience debugging complex systems and using measurement, profiling, and benchmarks to guide engineering decisions

- Proven ability to lead complex technical initiatives as a senior individual contributor and work effectively across organizational boundaries

**Preferred Skills**

- Experience with AI infrastructure, inference-serving systems, or large-scale machine learning systems

- Experience with compilers, runtimes, kernel optimization, or performance engineering; familiarity with technologies such as LLVM, MLIR, Triton, CUDA, or ROCm is a plus

- Familiarity with GPUs, accelerators, hardware architecture, ISA concepts, or vendor toolchains

- Experience with inference-serving frameworks or engines such as vLLM, SGLang, Triton Inference Server, or similar systems

- Experience building developer platforms, external APIs, remote execution systems, or secure partner-facing infrastructure

- Experience working with strategic cloud, hardware, or infrastructure partners

## Skills

### Required
- C++
- Python
- Go
- Rust
- distributed systems
- Linux
- networking
- storage
- containers
- cloud architectures

### Nice to have
- AI infrastructure
- inference-serving systems
- compilers
- runtimes
- kernel optimization
- performance engineering
- LLVM
- MLIR
- Triton
- CUDA
- ROCm
- GPUs
- accelerators
- hardware architecture
- ISA concepts
- vendor toolchains
- inference-serving frameworks
- developer platforms
- external APIs
- remote execution systems
- secure partner-facing infrastructure

---

Source: [Apply at jobs.ashbyhq.com](https://jobs.ashbyhq.com/openai/78c2a68b-cc77-4c62-8891-96afb603650a?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
