Description
Job Overview
The Safeguards team at Anthropic is responsible for ensuring models and products are developed and deployed safely. We are looking for a Staff+ Software Engineer to join our Review Tooling team, which builds systems used by humans and AI to investigate potential harms and take enforcement actions across Anthropic's products and third-party cloud platforms.
Key Responsibilities
- Build investigation, review, and enforcement tooling for first-party and third-party platform surfaces
- Develop a platform layer of reusable APIs, data storage, and backend services
- Scale review through automation, including enabling reviewers to use Claude effectively
- Partner with policy, operations, legal, privacy, and data science stakeholders to translate enforcement and investigation needs into reliable systems
- Build guardrails for sensitive internal tools, including granular permissions and audit trails
- Instrument tools to surface metrics on queue health, reviewer throughput, and decision quality
Requirements
- Technical background in full-stack or platform engineering
- Experience shipping internal tools or platforms with demanding operational users
- Experience working cross-functionally with non-engineering partners
- Excellent communication skills
- Care about the societal impacts of AI
Preferred Qualifications
- 8+ years of industry software engineering experience
- Experience building trust and safety, integrity, fraud, or abuse-prevention tooling
- Experience designing systems under strict privacy, compliance, or data governance constraints
- Experience integrating LLMs or agentic systems into operational workflows
- Experience building developer platforms or extensible tooling frameworks
Logistics
- Annual salary: $320,000 - $485,000 USD
- Location-based hybrid policy: 25% office time
- Visa sponsorship available
This listing is enriched and indexed by YubHub. To apply, use the employer's original posting:
https://job-boards.greenhouse.io/anthropic/jobs/5342935008