Description
As a Safeguards Enforcement Analyst on the user well-being team, you'll build and execute enforcement workflows that keep Anthropic's products safe, focusing on detecting and mitigating potential harm.
Your initial focus will be on how Anthropic handles age, ensuring consumer products reach the right audiences, including detecting underage users and creating appeals workflows.
This position may expand into broader areas of user well-being enforcement over time, shaping policy enforcement to allow users to safely interact with and build on top of Anthropic's products.
Key Responsibilities:
- Design and architect automated enforcement systems and review workflows that scale effectively while maintaining high accuracy
- Partner with Engineering and Data Science teams to optimize detection models for policy violations and automated enforcement systems
- Review flagged content to drive enforcement and policy improvements
- Enforce usage policies with a focus on detecting and mitigating potential harmful use of AI systems
- Collaborate with Legal, Public Policy, and Privacy stakeholders to keep the age assurance approach proportionate and responsive to regulations
- Support the Safeguards policy design team by providing feedback on policy gaps based on real enforcement scenarios
- Stay updated with emerging AI policy enforcement best practices to inform decision-making and workflows
- Manage Anthropic's layered age assurance approach to keep consumer products safe
- Handle adjacent user well-being enforcement where age is a key factor
Minimum Qualifications:
- Experience in trust and safety, online child safety, age assurance, privacy, product policy, or a related field
- Subject matter expertise in age assurance, age verification systems, age-appropriate design, child online safety, or content classification for young people
- Experience driving cross-functional initiatives with Product, Engineering, Legal, and Policy partners
- Familiarity with evolving regulatory landscapes and enforcement best practices regarding age assurance, CSAM/CSEM, NCII, and digital well-being
- Strong written communication skills and ability to collaborate with team members
- Comfort using data to measure effectiveness and inform decisions
Preferred Qualifications:
- Experience building or operating age-gating flows, age estimation signals, or appeals workflows
- Experience advising or partnering with third-party platforms on deploying safely to younger users
- Experience working with or evaluating third-party age verification providers
- Interest in AI safety and responsible technology development
- Experience writing effective prompts for generative AI systems in a content review or enforcement context
The annual compensation range for this role is $245,000-$285,000 USD.
This listing is enriched and indexed by YubHub. To apply, use the employer's original posting:
https://job-boards.greenhouse.io/anthropic/jobs/5311234008