Direct from source · No middlemen

Ai Safety Jobs (Remote)

621 open positions · Updated 3 months ago

Average salary: 282.4k–377.3k/yr

Showing 20 of 621 positions

Search with filters →
Anthropic
Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Washington, DC; San Francisco, CA | New York City, NY $230k–$270k/yr Published 4 months ago
Discord
Discord San Francisco Bay Area or Los Angeles Area $248k–$310k/yr Published 7 months ago
Anthropic

Join Anthropic as a Data Engineer to build data infrastructure that ensures AI systems are safe and beneficial.

Anthropic San Francisco, CA | New York City, NY $1–$2/yr Published 1 week ago
Flexible on stack
Anthropic

Lead the Interventions team to ensure safe deployment of AI systems at Anthropic.

Anthropic San Francisco, CA $405k–$485k/yr Published 2 weeks ago
Heavy meetings
Anthropic

Join Anthropic as a Data Engineer to build data infrastructure that ensures AI systems are safe and beneficial.

Anthropic San Francisco, CA | New York City, NY $320k–$405k/yr Published 3 weeks ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA | New York City, NY $350k–$850k/yr Published 3 months ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to protect against AI misuse in chemical and explosives contexts.

Anthropic Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 2 weeks ago
Anthropic
Anthropic San Francisco, CA | New York City, NY $320k–$425k/yr Published 9 months ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to enhance child safety in AI systems while managing content review workflows.

Anthropic Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 2 weeks ago
Anthropic

Join Anthropic's Safeguards team as a Red Team Engineer to enhance the safety of AI systems through adversarial testing.

Anthropic Remote-Friendly (Travel Required) | San Francisco, CA $320k–$405k/yr Published 2 weeks ago
Flexible on stack
Anthropic

Join Anthropic as a Product Manager to lead the development of AI Safeguards systems ensuring safe and beneficial AI for users.

Anthropic San Francisco, CA $305k–$385k/yr Published 3 weeks ago
Anthropic

Join Anthropic as a Product Manager to ensure AI systems are safe and beneficial for users across various platforms.

Anthropic San Francisco, CA $305k–$385k/yr Published 1 month ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to protect against AI misuse in radiological and nuclear contexts.

Anthropic Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 2 weeks ago
Anthropic

Join Anthropic as a Product Manager to ensure AI systems are safe and beneficial for users across various platforms.

Anthropic San Francisco, CA $305k–$385k/yr Published 5 months ago
Anthropic

Join Anthropic as a Staff+ Software Engineer to build critical review tooling for AI safety and enforcement.

Anthropic San Francisco, CA $320k–$485k/yr Published 2 weeks ago
Anthropic

Build evaluation infrastructure for AI safety systems at Anthropic, focusing on real-world misuse detection.

Anthropic San Francisco, CA | New York City, NY $320k–$485k/yr Published 1 month ago
Flexible on stack
Page 1 of 32