Trust & Safety jobs
Trust and safety at AI companies is the operational front line: abuse detection, content policy, moderation tooling, and the escalation queues where policy meets an actual user doing an actual strange thing. Generative products changed the job — the platform now produces content rather than just hosting it, which rewrites the classic playbooks.
Days mix policy writing, classifier development, and vendor management for human review, with LLMs increasingly moderating LLMs. Experience from platform T&S transfers well, and product-side pragmatism beats abstraction. Nearly every consumer AI company staffs this, not just the labs, making it the widest on-ramp in the safety category.
30open roles right now
- Fullstack Software Engineer, Child Safety Tools & SystemsOpenAIHybrid · Mid-level · AI Safety$266k–$385k
- Enforcement Operations Lead, Cloud PartnersAnthropicHybrid · Lead / Manager · AI Safety$285k–$330k
- Engineering Manager, SafeguardsAnthropicHybrid · Lead / Manager · AI Safety£325k–£390k
- Software Engineer, Integrity Foundations - LondonOpenAIOn-site · Senior · AI Safety$266k–$445k
- Safeguards Enforcement Lead, Cyber HarmsAnthropicHybrid · Lead / Manager · AI Safety$285k–$330k
- Product Manager, Multi-Cloud Trust & SafetyAnthropicHybrid · Senior · AI Safety$305k–$385k
- Senior / Staff Data Scientist, Trust & SafetySunoOn-site · Senior · AI Safety$220k–$320k
- Safety Engineer - Free Tier AbuseElevenLabsRemote, worldwide · Senior · AI Safety
- Safeguards Enforcement Analyst, User Well-beingAnthropicRemote (US) · Mid-level · AI Safety$245k–$285k
- Safeguards Enforcement Analyst, Account Takeover & Credential AbuseAnthropicRemote (US) · Mid-level · AI Safety$245k–$285k
- Safeguards Enforcement Analyst, Access Controls & IdentityAnthropicRemote (US) · Mid-level · AI Safety$285k–$330k
- Staff+ Software Engineer, Safeguards Human Review ToolingAnthropicHybrid · Staff+ · AI Safety$320k–$485k
- Safeguards Enforcement Analyst, Bio HarmsAnthropicRemote (US) · Mid-level · AI Safety$245k–$285k
- Safeguards Enforcement Analyst, Chem & Explosives HarmsAnthropicRemote (US) · Mid-level · AI Safety$245k–$285k
- Safeguards Enforcement Analyst, Age-Appropriate DesignAnthropicRemote (US) · Mid-level · AI Safety$245k–$285k
- Safeguards Enforcement Analyst, Cyber HarmAnthropicRemote (US) · Mid-level · AI Safety$285k–$330k
- Safeguards Enforcement Analyst, Integrity & AuthenticityAnthropicRemote (US) · Mid-level · AI Safety$285k–$330k
- Principal AI Security EngineerCerebrasRemote (US, Canada) · Staff+ · AI Safety
- Engineering Manager, Safeguards Review ToolingAnthropicHybrid · Lead / Manager · AI Safety$405k–$485k
- Data Scientist, SafetyOpenAIOn-site · Mid-level · AI Safety$265k–$380k
- Protection Scientist Engineer, IntegrityOpenAIOn-site · Mid-level · AI Safety$198k–$425k
- Senior Analyst, Safety Operations (Child Safety)xAIOn-site · Senior · AI Safety
- Senior Analyst, Safety Operations (Child Safety)xAIOn-site · Senior · AI Safety$91k–$130k
- Staff+ Software Engineer, Safeguards InfrastructureAnthropicHybrid · Staff+ · AI Safety£325k–£395k
- Full Stack Engineer, Intelligence SystemsOpenAIHybrid · Mid-level · AI Safety$266k–$385k
- Staff+ Software Engineer, SafeguardsAnthropicHybrid · Staff+ · AI Safety$320k–$485k
- Protection Scientist Engineer, IntegrityOpenAIOn-site · Mid-level · AI Safety
- Software Engineer, Trust & SafetyVercelHybrid · Senior · AI Safety$196k–$294k
- Software Engineer, AI SafetyOpenAIOn-site · Mid-level · AI Safety$207k–$385k
- Software Engineer, Integrity FoundationsOpenAIOn-site · Mid-level · AI Safety$266k–$445k