New The Skills of Tomorrow: how AI-exposed is every skill in 2026? See the data →
Anthropic

Head of Policy Design, Societal Harms

Anthropic
Apply →
hybrid senior full-time $330,000-$395,000 USD San Francisco, CA

First indexed 29 Aug 2026

Description

Job Overview

You will lead Anthropic's policy design team in the Safeguards organisation, focusing on societal harms. Your mission will be to define and enforce limits on how our AI model, Claude, can be used.

Key Responsibilities

  • Lead, develop, and grow the managers and teams responsible for the consumer harms portfolio, including child safety, user well-being, harmful manipulation, and election integrity.
  • Coordinate policy decisions across the portfolio and build mechanisms to track and enforce them.
  • Set the strategy for how mitigations built on top of the model complement what is trained into the model itself.
  • Prioritize across harm areas competing for the same resources and make tradeoffs clear to leadership.
  • Serve as the escalation point for high-severity and ambiguous consumer harms decisions.
  • Partner with engineering, data science, product, legal, and research across the model development cycle.
  • Engage external experts, civil society organisations, and regulators to inform policy and enforcement.

Requirements

  • Experience leading teams in AI safety, product policy, or a related field.
  • Deep, applied familiarity with consumer harm areas such as child safety, mental health and well-being, manipulation, or election integrity.
  • A track record of exceptional cross-team collaboration.
  • Working understanding of how frontier models are developed and deployed.
  • Experience translating policy positions into enforceable mechanisms and communicating reasoning to various audiences.

Preferred Qualifications

  • Subject-matter depth in one or more of the portfolio's harm areas.
  • Experience working directly with model training or research teams.
  • Experience with generative AI safety systems.
  • Experience engaging external stakeholders in these domains.
  • Experience using agentic AI tools to scale a team's analysis and operations.

Logistics

  • Annual Salary: $330,000-$395,000 USD
  • Bachelor’s degree or an equivalent combination of education, training, and/or experience.
  • Currently, we expect all staff to be in one of our offices at least 25% of the time.
  • We do sponsor visas and encourage applications from diverse candidates.
This listing is enriched and indexed by YubHub. To apply, use the employer's original posting: https://job-boards.greenhouse.io/anthropic/jobs/5407418008