Red Teaming jobs
Red teaming AI systems is professional creative misuse: jailbreaks, prompt injection, data extraction, and chaining harmless-looking capabilities into harmful ones. The best practitioners approach a model the way a penetration tester approaches a network — systematically, with notes — then write it up so engineers can actually fix things.
Hiring spans frontier labs, enterprises deploying AI into regulated workflows, and a growing consultancy scene. Security backgrounds transfer well; so does the particular stubbornness of good QA. Published jailbreaks and injection research function as a portfolio here, which keeps this a corner of AI safety where demonstrated work outweighs credentials.
12open roles right now
- Lead, Frontier Red Team (Cyber)AnthropicHybrid · Lead / Manager · AI Safety$485k–$755k
- Red Team Engineer, SafeguardsAnthropicRemote (US) · Mid-level · AI Safety$320k–$405k
- CyberSecurity, Offensive Security EngineerMistral AIOn-site · Senior · AI Safety
- Pentester, Offensive Forward Deployment EngineerMistral AIOn-site · Senior · AI Safety
- Strategic Projects Lead, Red TeamScale AIOn-site · Mid-level · AI Safety$151k–$189k
- Staff Software Engineer, FraudReplitHybrid · Staff+ · AI Safety$250k–$315k
- Senior Software Engineer, FraudReplitHybrid · Senior · AI Safety$210k–$265k
- Staff Software Engineer, Anti-Abuse & SecurityReplitHybrid · Staff+ · AI Safety$190k–$240k
- Research Program Manager – Adversarial Model ResearchOpenAIHybrid · Mid-level · AI Safety$239k–$328k
- Research Engineer / Scientist, Frontier Red Team (Cyber)AnthropicOn-site · Senior · AI Safety$320k–$485k
- Offensive Security Engineer, Agent ProductsOpenAIRemote (US) · Staff+ · AI Safety$347k–$490k
- Staff Security Software Engineer, AI SecurityDatabricksRemote (US) · Staff+ · AI Safety