# Software Engineer, Reinforcement learning

**Company**: Cursor
**Category**: Engineering
**Industry**: Technology

**Apply**: https://cursor.com/careers/software-engineer-rl-data?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_6f0b5abf-e83

## Description

Cursor is building the future of coding by training frontier coding agents and scaling RL on real user data to make them increasingly effective.

As a Software Engineer on the RL Data team at Cursor, you will create the tasks, rewards, and environments that train our coding agents. The team owns the data that goes into training: what the model is asked to do, how we score it, and the setups it learns in.

Responsibilities:

- Design a task set that teaches a specific agent capability, then iterate on it from traces and evals until the model actually gets better.

- Read agent traces, find a failure mode or a surprising behavior, and build a system that surfaces more of the same.

- Turn a one-off recipe into something other teams can reuse: better rewards, cleaner environments, tighter data quality.

- Partner with research on whether a dataset is actually teaching the thing we think it is.

Requirements:

- You write careful, fast code and have strong software engineering fundamentals.

- You like setting tasks: breaking a fuzzy capability into something concrete you can measure.

- You have an infra, data, or distributed systems background. RL experience is a plus, not a requirement.

- You enjoy looking at messy real-world agent behavior and turning it into a dataset or a tool.

## Skills

### Required
- software engineering
- reinforcement learning
- data analysis
- distributed systems

### Nice to have
- RL experience

---

Source: [Apply at cursor.com](https://cursor.com/careers/software-engineer-rl-data?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
