Description
We are seeking an experienced systems generalist who can work comfortably across the stack to help build an automated inference optimization platform.
Compensation
- $293K – $385K • Offers Equity
The base pay offered may vary depending on multiple individualized factors, including market location, job-related knowledge, skills, and experience.
Benefits
- Medical, dental, and vision insurance for you and your family, with employer contributions to Health Savings Accounts
- Pre-tax accounts for Health FSA, Dependent Care FSA, and commuter expenses (parking and transit)
- 401(k) retirement plan with employer match
- Paid parental leave (up to 24 weeks for birth parents and 20 weeks for non-birthing parents), plus paid medical and caregiver leave (up to 8 weeks)
- Paid time off: flexible PTO for exempt employees and up to 15 days annually for non-exempt employees
- 13+ paid company holidays, and multiple paid coordinated company office closures throughout the year for focus and recharge, plus paid sick or safe time
- Mental health and wellness support
- Employer-paid basic life and disability coverage
- Annual learning and development stipend to fuel your professional growth
- Daily meals in our offices, and meal delivery credits as eligible
- Relocation support for eligible employees
Key Responsibilities
- Design, build, and operate durable APIs and control-plane services for multi-hour or multi-day optimization campaigns
- Build secure partner-side runner and grader software that can compile, execute, verify, and benchmark candidate artifacts on third-party accelerator hardware
- Integrate hardware profiles, ISA and toolchain context, compilers, runtimes, and inference-serving engines into a repeatable optimization workflow
- Turn research prototypes into reliable product surfaces with clear contracts, debuggable failure modes, reproducible outputs, and excellent developer ergonomics
- Develop correctness and performance evaluation systems spanning latency, throughput, memory use, utilization, and cost efficiency
- Build artifact, provenance, and qualification workflows that make optimized kernels, binaries, configurations, and reports safe to review and deploy
- Collaborate with Research, Inference Engineering, Infrastructure, Security, Product, and Strategic Partnerships to deliver production-ready solutions
- Drive technical architecture and execution across ambiguous, cross-functional initiatives that connect OpenAI systems with partner environments
Basic Qualifications
- 8+ years of professional software engineering experience building large-scale distributed systems, infrastructure platforms, or cloud services, or equivalent depth of experience
- Strong programming skills in one or more of C++, Python, Go, or Rust
- Experience designing and operating highly available backend systems, APIs, job orchestration systems, or durable workflows for production workloads
- Strong understanding of distributed systems, Linux, networking, storage, containers, and modern cloud architectures
- Experience debugging complex systems and using measurement, profiling, and benchmarks to guide engineering decisions
- Proven ability to lead complex technical initiatives as a senior individual contributor and work effectively across organizational boundaries
Preferred Skills
- Experience with AI infrastructure, inference-serving systems, or large-scale machine learning systems
- Experience with compilers, runtimes, kernel optimization, or performance engineering; familiarity with technologies such as LLVM, MLIR, Triton, CUDA, or ROCm is a plus
- Familiarity with GPUs, accelerators, hardware architecture, ISA concepts, or vendor toolchains
- Experience with inference-serving frameworks or engines such as vLLM, SGLang, Triton Inference Server, or similar systems
- Experience building developer platforms, external APIs, remote execution systems, or secure partner-facing infrastructure
- Experience working with strategic cloud, hardware, or infrastructure partners
This listing is enriched and indexed by YubHub. To apply, use the employer's original posting:
https://jobs.ashbyhq.com/openai/78c2a68b-cc77-4c62-8891-96afb603650a