New The Skills of Tomorrow: how AI-exposed is every skill in 2026? See the data →
Anthropic

Staff+ Researcher, Cybersecurity Products

Anthropic
Apply →
hybrid staff full-time $405,000-$485,000 USD San Francisco, CA

First indexed 11 Aug 2026

Description

Anthropic's mission is to create reliable, interpretable, and steerable AI systems that are safe and beneficial for users and society.

We're looking for a Capabilities Researcher to join the team building Claude Security. In this role, you'll identify which security capabilities in frontier models are ready to build on, measure how well they perform, and work out how to make them useful to customers who are not security experts.

Responsibilities

  • Prototype rapidly to define the AI frontier for cybersecurity work
  • Design evaluations that measure model performance on the work security teams actually do
  • Build the datasets, harnesses, and scoring those evaluations depend on
  • Engage with the cybersecurity community to help define where AI can make the most impact
  • Work with engineers and researchers to operationalize promising capabilities into something customers can rely on
  • Track how model capabilities for security are changing, and what that means for what we build next
  • Share findings that inform product direction, and partner with product leadership on priorities

Requirements

  • Have deep expertise in one or more security domains, such as vulnerability research, exploit development, reverse engineering, malware analysis, incident response, or offensive security
  • Have built AI-powered tools or capabilities for security work
  • Can get from an idea to a working prototype quickly, and abandon the ones that don't hold up
  • Are comfortable designing rigorous evaluations and interpreting the results honestly
  • Can write and communicate clearly about technical findings
  • Have 7+ years of experience in security research, security engineering, or a closely related field

Strong candidates may also have:

  • Published research, CTF results, CVEs, or open source security tooling
  • Experience with model evaluation, benchmarking, or red teaming
  • Experience building agentic applications
  • Familiarity with the safety considerations of AI in security contexts

The annual compensation range for this role is $405,000-$485,000 USD.

This listing is enriched and indexed by YubHub. To apply, use the employer's original posting: https://job-boards.greenhouse.io/anthropic/jobs/5385217008