"safeguards" Jobs

506 open tech roles matching “safeguards”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, SQL. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 506 results

Anthropic

Join Anthropic as a Staff+ Software Engineer to build safety mechanisms for AI systems with a focus on data governance and integrity.

Anthropic San Francisco, CA | New York City, NY $320k–$485k/yr Published 4 weeks ago
Flexible on stack
Anthropic

Join Anthropic as a Staff+ Site Reliability Engineer to ensure safe AI model launches and automate deployment processes.

Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY $320k–$485k/yr Published 1 week ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA | New York City, NY $320k–$425k/yr Published 11 months ago
Anthropic

Join Anthropic as a Staff+ Software Engineer to build critical review tooling for AI safety and enforcement.

Anthropic San Francisco, CA $320k–$485k/yr Published 2 months ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to enhance access controls and identity verification in AI systems.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $285k–$330k/yr Published 2 months ago
Anthropic

Support the design and deployment of mental health guardrails as a Safeguards Enforcement Analyst at Anthropic.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 1 month ago
Anthropic
Anthropic San Francisco, CA | New York City, NY | Washington, DC $230k–$270k/yr Published 6 months ago
Anthropic

Join Anthropic as a Product Manager to ensure AI systems are safe and beneficial for users across various platforms.

Anthropic San Francisco, CA $305k–$385k/yr Published 3 months ago
Anthropic

Own the infrastructure for machine learning research to detect and mitigate misuse of AI models at Anthropic.

Anthropic San Francisco, CA | New York City, NY $350k–$500k/yr Published 1 month ago
Flexible on stack
Anthropic

Join Anthropic as a Data Engineer to build data infrastructure that ensures AI systems are safe and beneficial.

Anthropic San Francisco, CA | New York City, NY $320k–$405k/yr Published 2 months ago
Flexible on stack
Revolut

Manage regulatory compliance and safeguarding initiatives in a dynamic financial technology environment.

Revolut Barcelona, Spain Published 2 weeks ago
Anthropic

Join Anthropic as a Product Manager to ensure AI systems are safe and beneficial for users across various platforms.

Anthropic San Francisco, CA $305k–$385k/yr Published 7 months ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to mitigate misuse of AI systems related to conventional weapons.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $245k–$330k/yr Published 1 week ago
Anthropic

Join Anthropic as a Product Manager to lead the development of AI Safeguards systems ensuring safe and beneficial AI for users.

Anthropic San Francisco, CA $305k–$385k/yr Published 2 months ago
Anthropic

Join Anthropic's Safeguards team as a Red Team Engineer to enhance the safety of AI systems through adversarial testing.

Anthropic Remote-Friendly (Travel Required) | San Francisco, CA $320k–$405k/yr Published 2 months ago
Flexible on stack
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to lead fraud and scams enforcement in a mission-driven AI company.

Anthropic Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 2 months ago
Anthropic

Join Anthropic as a Data Scientist to build a data-driven culture and enhance AI safety and product strategy.

Anthropic New York City, NY; San Francisco, CA; Seattle, WA $275k–$370k/yr Published 3 months ago
Flexible on stack
Anthropic

Lead the Review Tooling team to develop systems for safe AI model investigation and enforcement at Anthropic.

Anthropic London, UK £325k–£390k/yr Published 1 week ago
Heavy meetings