"trust safety" Jobs
47 open tech roles matching “trust safety”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: AI/ML, Python, SQL. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 47 results
Join Anthropic's Safeguards team as a Red Team Engineer to enhance the safety of AI systems through adversarial testing.
Own the risk management of mission-critical and highest-risk vendor portfolios at a leading AI lab.
Join Anthropic as a Safeguards Enforcement Analyst to lead fraud and scams enforcement in a mission-driven AI company.
Join Anthropic as a Safeguards Enforcement Analyst to enhance access controls and identity verification in AI systems.
Join Anthropic as a Safeguards Enforcement Analyst to enhance AI safety by managing account compromise and credential abuse.
Join Anthropic as a Safeguards Enforcement Analyst to enhance child safety in AI systems while managing content review workflows.
Join Anthropic as a Safeguards Enforcement Analyst to protect against AI misuse in radiological and nuclear contexts.
Join Anthropic as a Safeguards Analyst to build enforcement workflows for AI systems against misuse and ensure user safety.
Join Anthropic as a Safeguards Enforcement Analyst to protect against AI misuse in chemical and explosives contexts.
Join Anthropic as a Safeguards Enforcement Analyst to enhance AI safety by investigating ban evasion and developing enforcement workflows.
Join Anthropic as a Safeguards Enforcement Analyst to ensure AI systems are safe and age-appropriate for users.
Join Anthropic as a Safeguards Enforcement Analyst to help detect and mitigate misuse of AI systems in a hybrid work environment.
Join Anthropic as a Safeguards Enforcement Analyst to protect against AI misuse in biological contexts.
Join Anthropic as a Staff+ Application Security Engineer to lead security due diligence and integration for acquisitions in a hybrid work environment.
Join Anthropic as a Safeguards Enforcement Analyst to combat violence and extremism in AI systems.