New The Skills of Tomorrow: how AI-exposed is every skill in 2026? See the data →
OpenAI

Researcher, Safety Training, National Security

OpenAI
Apply →
remote senior Full time $380K - $500K San Francisco

First indexed 12 Sept 2026

Description

Compensation

$380K – $500K • Offers Equity

The base pay offered may vary depending on multiple individualized factors, including market location, job-related knowledge, skills, and experience. If the role is non-exempt, overtime pay will be provided consistent with applicable laws. In addition to the salary range listed above, total compensation also includes generous equity, performance-related bonus(es) for eligible employees, and the following benefits.

  • Medical, dental, and vision insurance for you and your family, with employer contributions to Health Savings Accounts
  • Pre-tax accounts for Health FSA, Dependent Care FSA, and commuter expenses (parking and transit)
  • 401(k) retirement plan with employer match
  • Paid parental leave (up to 24 weeks for birth parents and 20 weeks for non-birthing parents), plus paid medical and caregiver leave (up to 8 weeks)
  • Paid time off: flexible PTO for exempt employees and up to 15 days annually for non-exempt employees
  • 13+ paid company holidays, and multiple paid coordinated company office closures throughout the year for focus and recharge, plus paid sick or safe time (1 hour per 30 hours worked, or more, as required by applicable state or local law)
  • Mental health and wellness support
  • Employer-paid basic life and disability coverage
  • Annual learning and development stipend to fuel your professional growth
  • Daily meals in our offices, and meal delivery credits as eligible
  • Relocation support for eligible employees
  • Additional taxable fringe benefits, such as charitable donation matching and wellness stipends, may also be provided.

About the Team

The Safety Training research team aims to fundamentally advance our capabilities for precisely implementing safe behavior in AI models, and to leverage these advances to make OpenAI’s deployed models safe and beneficial. This requires a breadth of new ML research to address the growing set of safety challenges as AI becomes more powerful and used in more settings. Key focus areas include how to train nuanced safety behaviors, how to make the model robust to bad actors, how to address privacy and security risks, and how to make the model trustworthy in safety-critical situations.

About the Role

We’re seeking a researcher to train and evaluate models for U.S. government use, with a focus on national security applications. You’ll advance safety post-training and robustness, helping models follow nuanced policies while preserving their usefulness and capabilities.

In this role, you will:

  • Research and implement methods for safety training, reinforcement learning, and adversarial robustness.
  • Develop evaluations, identify model failure modes, and use findings to improve training.
  • Work with research, engineering, security, and policy partners to support safe, reliable deployment.

You might thrive in this role if you:

  • Bring 4+ years of relevant AI safety research experience, including RLHF, adversarial training, or robustness.
  • Have a degree in computer science, machine learning, or a related field, and strong deep learning research or engineering skills.
  • Have experience improving model safety for deployment and enjoy collaborative research.
  • Are motivated by OpenAI’s mission and the responsible use of AI in safety-critical settings.

Security Requirements

  • Active TS/SCI clearance or equivalent.
This listing is enriched and indexed by YubHub. To apply, use the employer's original posting: https://jobs.ashbyhq.com/openai/b60d3040-38e4-4abe-8af8-881286b10d50