Description
Compensation
The base pay for this role is $266K – $335K, with additional compensation including equity, performance-related bonuses, and benefits.
Benefits include:
- Medical, dental, and vision insurance for you and your family
- Pre-tax accounts for Health FSA, Dependent Care FSA, and commuter expenses
- 401(k) retirement plan with employer match
- Paid parental leave (up to 24 weeks for birth parents and 20 weeks for non-birthing parents)
- Paid medical and caregiver leave (up to 8 weeks)
- Flexible PTO for exempt employees and up to 15 days annually for non-exempt employees
- 13+ paid company holidays
- Mental health and wellness support
- Employer-paid basic life and disability coverage
- Annual learning and development stipend
- Daily meals in our offices, and meal delivery credits as eligible
- Relocation support for eligible employees
About the Team
The Safety Systems team manages the complete lifecycle of safety efforts for OpenAI’s frontier models, ensuring they are deployed responsibly and have a positive impact on society. The Model Policy team works to ensure that frontier models behave safely and reliably in real-world environments by designing policies that define safe model behavior.
About the Role
We’re hiring a Model Policy Manager to shape model behavior for U.S. government use, with a focus on national security applications. You’ll define nuanced policies and translate them into training and evaluation criteria, helping models navigate high-stakes scenarios while preserving their usefulness and capabilities.
Responsibilities
- Develop model policies that guide safe and useful behavior
- Build evaluations, identify policy gaps and model failures, and use findings to improve policies and training
- Work with research, engineering, and domain experts to support safe, reliable deployment
Requirements
- Relevant experience in AI safety, policy, or risk assessment
- Strong judgment and ability to turn complex safety questions into clear, practical policies
- Technical fluency to work hands-on with model data and evaluations
- Motivation by OpenAI’s mission and the responsible use of AI in safety-critical settings
- Active TS/SCI clearance or equivalent