Direct from source · No middlemen

Trust Safety Jobs in Friendly (Remote)

47 open positions · Updated 2 weeks ago

Average salary: 261.5k–315k/yr

Showing 20 of 47 positions

Search with filters →
Anthropic

Join Anthropic's Safeguards team as a Red Team Engineer to enhance the safety of AI systems through adversarial testing.

Anthropic Remote-Friendly (Travel Required) | San Francisco, CA $320k–$405k/yr Published 2 weeks ago
Flexible on stack
Anthropic

Own the risk management of mission-critical and highest-risk vendor portfolios at a leading AI lab.

Anthropic Remote-Friendly (Travel Required) | San Francisco, CA $255k–$270k/yr Published 6 days ago
Anthropic
Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Washington, DC; San Francisco, CA | New York City, NY $230k–$270k/yr Published 4 months ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to lead fraud and scams enforcement in a mission-driven AI company.

Anthropic Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 2 weeks ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to enhance access controls and identity verification in AI systems.

Anthropic Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC $285k–$330k/yr Published 2 weeks ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to enhance AI safety by managing account compromise and credential abuse.

Anthropic Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 2 weeks ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to enhance child safety in AI systems while managing content review workflows.

Anthropic Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 2 weeks ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to protect against AI misuse in radiological and nuclear contexts.

Anthropic Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 2 weeks ago
Anthropic

Join Anthropic as a Safeguards Analyst to build enforcement workflows for AI systems against misuse and ensure user safety.

Anthropic Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC $285k–$330k/yr Published 2 weeks ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to protect against AI misuse in chemical and explosives contexts.

Anthropic Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 2 weeks ago
Anthropic
Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | New York City, NY $245k–$285k/yr Published 3 months ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to enhance AI safety by investigating ban evasion and developing enforcement workflows.

Anthropic Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 2 weeks ago
Anthropic
Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY $300k–$405k/yr Published 1 year ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to ensure AI systems are safe and age-appropriate for users.

Anthropic Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 2 weeks ago
Anthropic
Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Washington, DC $230k–$290k/yr Published 6 months ago
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to help detect and mitigate misuse of AI systems in a hybrid work environment.

Anthropic Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC $285k–$330k/yr Published 2 weeks ago
Flexible on stack
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to protect against AI misuse in biological contexts.

Anthropic Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 2 weeks ago
Anthropic
Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Washington, DC $230k–$290k/yr Published 6 months ago
Anthropic

Join Anthropic as a Staff+ Application Security Engineer to lead security due diligence and integration for acquisitions in a hybrid work environment.

Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY $320k–$485k/yr Published 2 weeks ago
Flexible on stack
Anthropic

Join Anthropic as a Safeguards Enforcement Analyst to combat violence and extremism in AI systems.

Anthropic Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC $285k–$330k/yr Published 2 weeks ago
Flexible on stack
Page 1 of 3