"safeguards" Jobs
112 open tech roles matching “safeguards”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: AI/ML, Python, SQL. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 112 results
Join Anthropic as a Safeguards Enforcement Analyst to enhance child safety in AI systems while managing content review workflows.
Join Anthropic as a Safeguards Enforcement Analyst to help detect and mitigate misuse of AI systems in a hybrid work environment.
Join Anthropic as a Safeguards Analyst to build enforcement workflows for AI systems against misuse and ensure user safety.
Join Anthropic as a Safeguards Enforcement Analyst to combat violence and extremism in AI systems.
Join Anthropic as a Safeguards Enforcement Analyst to protect against AI misuse in chemical and explosives contexts.
Support cyber product policy work at Anthropic, ensuring compliance with usage policies and safety standards.
Lead the Review Tooling team to enhance safety investigations and enforcement systems for AI products at Anthropic.
Join Anthropic as a Safeguards Enforcement Analyst to protect against AI misuse in biological contexts.
Join Anthropic as a Staff Software Engineer to build scalable ML infrastructure for AI safety systems.
Lead enforcement actions to mitigate misuse of AI systems against cyber threats while managing a team of Cyber Enforcement Analysts.
Join Anthropic as a Product Policy Manager to assess product safety risks and drive responsible AI deployment.
Design and operate backend systems for Claude's safety systems, ensuring low latency and high reliability.
Join Anthropic as a Safety & Security Counsel to shape legal frameworks for responsible AI development and deployment.
Manage and define policies for conventional weapons within AI systems at Anthropic, ensuring safety and compliance.
Lead a research engineering team focused on biological safety evaluations and classifiers in a mission-driven AI organization.
Lead the policy design team at Anthropic to manage consumer harms related to AI systems.