"ai safety" Jobs

910 open tech roles matching “ai safety”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: AI/ML, Python, AWS. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 910 results

Decagon

Join Decagon as a Research Engineer to enhance AI safety and reliability in conversational agents.

Decagon San Francisco $200k–$400k/yr Published 1 week ago
Flexible on stack
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to enhance child safety in AI systems while managing content review workflows.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 2 months ago
Anthropic

Support the design and deployment of mental health guardrails as a Safeguards Enforcement Analyst at Anthropic.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 1 month ago
Anthropic

Join Anthropic as a Product Policy Manager to assess product safety risks and drive responsible AI deployment.

Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY $245k–$285k/yr Published 2 weeks ago
Sierra

Join Sierra as a Security Engineer to build secure systems that protect customer data and trust while enabling innovation.

Sierra San Francisco, CA Published 1 year ago
AI-first team
Anthropic

Join Anthropic's Safeguards team as a Red Team Engineer to enhance the safety of AI systems through adversarial testing.

Anthropic Remote-Friendly (Travel Required) | San Francisco, CA $320k–$405k/yr Published 2 months ago
Flexible on stack
Anthropic

Join Anthropic as a Product Manager to lead the development of AI Safeguards systems ensuring safe and beneficial AI for users.

Anthropic San Francisco, CA $305k–$385k/yr Published 2 months ago
Anthropic

Join Anthropic as a Product Manager to ensure AI systems are safe and beneficial for users across various platforms.

Anthropic San Francisco, CA $305k–$385k/yr Published 3 months ago
Anthropic

Join Anthropic as a Product Manager to ensure AI systems are safe and beneficial for users across various platforms.

Anthropic San Francisco, CA $305k–$385k/yr Published 7 months ago
Faire

Lead the physical safety and security function at Faire during a pivotal growth period, impacting employee trust and workplace safety.

Faire San Francisco, CA $160k–$220k/yr Published 2 weeks ago
Anthropic

Join Anthropic as a Staff+ Software Engineer to build critical review tooling for AI safety and enforcement.

Anthropic San Francisco, CA $320k–$485k/yr Published 2 months ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to enhance access controls and identity verification in AI systems.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $285k–$330k/yr Published 2 months ago
Anthropic

Join Anthropic as a Data Scientist to build a data-driven culture and enhance AI safety and product strategy.

Anthropic New York City, NY; San Francisco, CA; Seattle, WA $275k–$370k/yr Published 3 months ago
Flexible on stack
Anthropic

Join Anthropic as a Physical Security Design Lead to enhance contract documents and ensure quality in global office facilities.

Anthropic Boston, MA; Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY; Washington, DC $180k–$230k/yr Published 3 weeks ago
Flexible on stack
Anthropic

Own the technical needs for executive security systems, managing design, integration, and vendor relationships in a hybrid role.

Anthropic Boston, MA; Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY; Washington, DC $245k–$305k/yr Published 1 month ago
Sierra

Join Sierra as a Software Engineer in Security to build secure AI systems and enhance customer trust.

Sierra San Francisco, CA Published 8 months ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to help detect and mitigate misuse of AI systems in a hybrid work environment.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $285k–$330k/yr Published 2 months ago
Flexible on stack
Anthropic

Drive endpoint security engineering at Anthropic, ensuring safe and efficient access for a growing team.

Anthropic San Francisco, CA | Seattle, WA | New York City, NY | Washington, DC $320k–$405k/yr Published 3 weeks ago
Flexible on stack AI-first team
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to mitigate misuse of AI systems related to conventional weapons.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $245k–$330k/yr Published 1 week ago
Anthropic
Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY $300k–$405k/yr Published 1 year ago