"safeguards" Jobs

96 open tech roles matching “safeguards”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, SQL, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 96 results

Anthropic

Join Anthropic as a Staff+ Software Engineer to build safety mechanisms for AI systems with a focus on data governance and integrity.

Anthropic San Francisco, CA | New York City, NY $320k–$485k/yr Published 1 month ago
Flexible on stack
Anthropic

Join Anthropic as a Staff+ Site Reliability Engineer to ensure safe AI model launches and automate deployment processes.

Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY $320k–$485k/yr Published 1 week ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA | New York City, NY $320k–$425k/yr Published 11 months ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to enhance access controls and identity verification in AI systems.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $285k–$330k/yr Published 2 months ago
Anthropic

Support the design and deployment of mental health guardrails as a Safeguards Enforcement Analyst at Anthropic.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 1 month ago
Anthropic
Anthropic San Francisco, CA | New York City, NY | Washington, DC $230k–$270k/yr Published 6 months ago
Anthropic

Own the infrastructure for machine learning research to detect and mitigate misuse of AI models at Anthropic.

Anthropic San Francisco, CA | New York City, NY $350k–$500k/yr Published 1 month ago
Flexible on stack
Anthropic

Join Anthropic as a Data Engineer to build data infrastructure that ensures AI systems are safe and beneficial.

Anthropic San Francisco, CA | New York City, NY $320k–$405k/yr Published 2 months ago
Flexible on stack
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to mitigate misuse of AI systems related to conventional weapons.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $245k–$330k/yr Published 1 week ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to lead fraud and scams enforcement in a mission-driven AI company.

Anthropic Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 2 months ago
Anthropic

Join Anthropic as a Data Scientist to build a data-driven culture and enhance AI safety and product strategy.

Anthropic New York City, NY; San Francisco, CA; Seattle, WA $275k–$370k/yr Published 3 months ago
Flexible on stack
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to ensure AI systems are safe and age-appropriate for users.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 2 months ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to enhance AI safety by managing account compromise and credential abuse.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 2 months ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to enhance child safety in AI systems while managing content review workflows.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 2 months ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to help detect and mitigate misuse of AI systems in a hybrid work environment.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $285k–$330k/yr Published 2 months ago
Flexible on stack
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to enhance AI safety by investigating ban evasion and developing enforcement workflows.

Anthropic Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 2 months ago
Anthropic

Join Anthropic as a Safeguards Analyst to build enforcement workflows for AI systems against misuse and ensure user safety.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $285k–$330k/yr Published 2 months ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to combat violence and extremism in AI systems.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $285k–$330k/yr Published 2 months ago
Flexible on stack
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to protect against AI misuse in chemical and explosives contexts.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 2 months ago