# Cyber Evaluations Engineer

**Company**: Anthropic
**Location**: San Francisco, CA
**Work arrangement**: hybrid
**Experience**: senior
**Job type**: full-time
**Salary**: $300,000-$405,000 USD
**Category**: Engineering
**Industry**: Technology
**Wikidata**: https://www.wikidata.org/wiki/Q116758847

**Apply**: https://job-boards.greenhouse.io/anthropic/jobs/5406367008?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_5d5aef88-08f

## Description

## About the role

We're hiring Cyber Evaluations Engineers to build and run evaluations that measure cyber-relevant capabilities and safeguard robustness in our models. You will design new evaluations, run per-release robustness testing, and analyze data on jailbreaks and prompt bypasses. You will also design probes to detect cyber abuse in production and help shape the overall detection architecture.

## Key responsibilities

- Design and run capability, uplift, and safety evaluations to assess cyber-relevant risk in new models

- Execute per-release safeguard-robustness testing ahead of major model launches

- Analyze evaluation results and communicate findings clearly to the team and stakeholders

- Design, prototype, and tune detection probes for cyber misuse

- Work with the cyber policy team to turn policy lines into a layered, robust abuse-detection architecture, and measure its precision and coverage over time

- Build and maintain internal tooling used to run and score evaluations

- Collaborate with policy and engineering partners to translate evaluation findings into safeguard improvements

## Minimum qualifications

- Experience building or running evaluations, benchmarks, or test suites for software or ML systems

- Hands-on cybersecurity experience

- Proficiency in Python

- Strong ability to communicate evaluation results with multiple stakeholders

## Preferred qualifications

- Deep offensive-security or security-research experience

- Experience analyzing adversarial or abuse data

- Experience working onsite with government partners

- Experience with AI/ML evaluation frameworks

- Familiarity with coordinated vulnerability disclosure practices

- Experience testing pre-release or pre-deployment software or models under confidentiality constraints

- Experience authoring detection content or building ML-based abuse detection

- Active secret security clearance or higher, or eligibility to obtain one

## Logistics

- Annual Salary: $300,000-$405,000 USD

## Skills

### Required
- Python
- cybersecurity
- evaluations
- benchmarks
- test suites

### Nice to have
- offensive-security
- security-research
- AI/ML evaluation frameworks
- coordinated vulnerability disclosure
- detection content
- ML-based abuse detection

---

Source: [Apply at job-boards.greenhouse.io](https://job-boards.greenhouse.io/anthropic/jobs/5406367008?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
