AI Red Team Engineer

London & San Franciscoonsitemid

Posted 2w ago · via Lever

About this role

THE OPPORTUNITY We are currently building Watcher,  a monitoring tool for coding agents. Our monitoring research agenda attempts to translate compute into safety at scale. Red-teaming previously sat inside the RS (Control) role as a partial responsibility. As it's grown from a single pilot into a recurring need, it now needs a dedicated owner. As the AI Red Team Engineer, you will help build the practice of red-teaming AI monitors (both Watcher's own defenses and frontier labs' monitoring systems (see our pilot campaign red-teaming Anthropic's auto mode). You will hunt for attack surfaces monitors that haven't been tested against yet and turn what you find into fixes.You'll work closely with Marius (CEO & currently leads the monitoring efforts), control researchers and product engineers.…

Read the full description on Apollo Research's site →

What we'd score you on

reqspace match rubric

Five dimensions, recruiter-grade. Upload your resume and we'll generate a written explanation of where you fit and where the gaps are.

1

Skills match

For this role: mode, anthropic

2

Level fit

This role is mid-level. We check your trajectory against it.

3

Domain experience

Your work in the role's domain matters more than your years total. We weight recent and direct experience.

4

Recency

A skill you used last quarter weighs more than one from five years ago. We grade on recency, not lifetime.

5

Location fit

This role is based in London & San Francisco. We weight your proximity and willingness to relocate.

Score yourself on this role.
Free · no card · written explanation included
See if I'm a fit →

Skills in this role

Pulled from the job description. These are the keywords we'll weight when scoring your fit.

modeanthropic

More at Apollo Research

See all open jobs at Apollo Research