— Making sure the capable thing is also a careful thing.
AI Safety & Alignment jobs
AI safety and alignment has grown from a niche into a working discipline: dangerous-capability evals, red-teaming, interpretability research, and the policy-adjacent roles that translate all of it for regulators. The titles to watch for are Alignment Researcher, Safety Engineer, and Model Policy Lead, with trust-and-safety hybrids appearing at product companies.
The surprise for many applicants is how much of the work is engineering — safety teams build eval harnesses and monitoring infrastructure, not just position papers, so RLHF literacy and solid ML plumbing both matter. It's a small field concentrated at a handful of labs, which keeps openings scarce and competition real. Remote policies vary lab by lab; pay is strong and rarely printed.
— Specialisms
115open roles right now
- Researcher, Recursive Self-Improvement SafetyOpenAIOn-site · Senior · AI Safety$295k–$445k
- Safety Engineer - Free Tier AbuseElevenLabsRemote, worldwide · Senior · AI Safety
- Researcher, Alignment CoT MonitorabilityOpenAIHybrid · Senior · AI Safety$250k–$445k
- Researcher, Frontier Cybersecurity RisksOpenAIOn-site · Senior · AI Safety$295k–$445k
- Safeguards Enforcement Analyst, User Well-beingAnthropicRemote (US) · Mid-level · AI Safety$245k–$285k
- Researcher, Multimodal SafetyOpenAIHybrid · Senior · AI Safety$295k–$445k
- Strategic Narrative and Impact Lead, Intelligence and InvestigationsOpenAIOn-site · Senior · AI Safety$288k–$425k
- Member of Technical Staff (Secure Intelligence Institute)PerplexityOn-site · Senior · AI Safety$220k–$405k
- Safety Transparency Editor, Safety SystemsOpenAIOn-site · Mid-level · AI Safety$220k–$245k
- Lead, Frontier Red Team (Cyber)AnthropicHybrid · Lead / Manager · AI Safety$485k–$755k
- Safeguards Enforcement Analyst, Violence & ExtremismAnthropicRemote (US) · Mid-level · AI Safety$285k–$330k
- Engineering Manager, Agent & Product SecurityCursor (Anysphere)On-site · Lead / Manager · AI Safety
- Research Engineer, PrivacyOpenAIOn-site · Mid-level · AI Safety$380k–$445k
- Software Security Engineering Manager, Secure FrameworksAnthropicHybrid · Lead / Manager · AI Safety$405k–$485k
- Quantitative Intelligence AnalystOpenAIHybrid · Mid-level · AI Safety$198k–$320k
- Safeguards Enforcement Analyst, Ban Evasion & RecidivismAnthropicRemote (US) · Mid-level · AI Safety$245k–$285k
- Safeguards Enforcement Analyst, Account Takeover & Credential AbuseAnthropicRemote (US) · Mid-level · AI Safety$245k–$285k
- Safeguards Enforcement Analyst, Access Controls & IdentityAnthropicRemote (US) · Mid-level · AI Safety$285k–$330k
- Staff+ Software Engineer, Safeguards Review ToolingAnthropicHybrid · Staff+ · AI Safety$320k–$485k
- Safeguards Enforcement Analyst, Radiological & Nuclear HarmsAnthropicRemote (US) · Mid-level · AI Safety$245k–$285k
- Safeguards Enforcement Analyst, Bio HarmsAnthropicRemote (US) · Mid-level · AI Safety$245k–$285k
- Safeguards Enforcement Analyst, Chem & Explosives HarmsAnthropicRemote (US) · Mid-level · AI Safety$245k–$285k
- Safeguards Enforcement Analyst, Age-Appropriate DesignAnthropicRemote (US) · Mid-level · AI Safety$245k–$285k
- Safeguards Enforcement Analyst, Cyber HarmAnthropicRemote (US) · Mid-level · AI Safety$285k–$330k
- Safeguards Enforcement Analyst, Integrity & AuthenticityAnthropicRemote (US) · Mid-level · AI Safety$285k–$330k
- Red Team Engineer, SafeguardsAnthropicRemote (US) · Mid-level · AI Safety$320k–$405k
- Engineering Manager, Safeguards InterventionsAnthropicHybrid · Lead / Manager · AI Safety$405k–$485k
- Partner Manager, Global HealthAnthropicHybrid · Senior · AI Safety$215k–$300k
- CyberSecurity, Offensive Security EngineerMistral AIOn-site · Senior · AI Safety
- Pentester, Offensive Forward Deployment EngineerMistral AIOn-site · Senior · AI Safety
- Threat Intel Manager, CBRN-E & Advanced WeaponsAnthropicHybrid · Senior · AI Safety$375k–$455k
- Threat Intel Manager, Influence Operations & SurveillanceAnthropicHybrid · Lead / Manager · AI Safety$375k–$455k
- Product Manager, Safeguards (Child Safety)AnthropicHybrid · Senior · AI Safety$305k–$385k
- Technical Program Manager, Strategic InitiativesOpenAIHybrid · Senior · AI Safety$257k–$445k
- Senior Security Engineer (AI Safety), London, LausanneIsomorphic LabsHybrid · Senior · AI Safety
- System Safety EngineerApplied IntuitionOn-site · Mid-level · AI Safety
- Senior Research Engineer - Safety Tooling and DataCohereRemote (US, Canada, UK, …) · Senior · AI SafetyUSD/CAD 230k–USD/CAD 535k
- Product Security EngineerVercelHybrid · Senior · AI Safety$208k–$312k
- Principal AI Security EngineerCerebrasRemote (US, Canada) · Staff+ · AI Safety
- Lead Safety Engineer, RoboticsOpenAIOn-site · Lead / Manager · AI Safety$342k–$468k
- Engineering Manager, Sensitive DeploymentsOpenAIOn-site · Lead / Manager · AI Safety$347k–$490k
- Product Manager, Safeguards Rare HarmsAnthropicHybrid · Senior · AI Safety$305k–$385k
- Agentic Risk AnalystOpenAIOn-site · Senior · AI Safety$288k–$425k
- Engineering Manager, Safeguards Review ToolingAnthropicHybrid · Lead / Manager · AI Safety$405k–$485k
- Staff+ Software Engineer, Safeguards EvalsAnthropicHybrid · Staff+ · AI Safety$320k–$485k
- Staff+ Security Engineer, Risk EngineeringAnthropicHybrid · Staff+ · AI Safety$320k–$405k
- Strategic Projects Lead, Red TeamScale AIOn-site · Mid-level · AI Safety$151k–$189k
- Data Scientist, SafeguardsAnthropicHybrid · Senior · AI Safety$285k–$380k
- Member of Technical Staff - Imagine SafetyxAIOn-site · Senior · AI Safety$180k–$440k
- Technical Threat Investigator, Threat Intel Engineering - UKOpenAIRemote (UK, US) · Mid-level · AI Safety