Direct from source · No middlemen

Interpretability Jobs

1327 open positions · Updated 3 months ago

Average salary: 326.6k–487.8k/yr

Showing 20 of 1327 positions

Search with filters →
Anthropic

Join Anthropic as a Staff Software Engineer to drive reinforcement learning efforts and design systems for coding capabilities.

Anthropic San Francisco, CA | New York City, NY | Seattle, WA $405k–$625k/yr Published 3 days ago
Stripe

Join Stripe as an Operations Insights analyst to enhance Tax operations through data analysis and reporting infrastructure.

Stripe US Remote Published 1 month ago
Flexible on stack
Anthropic
Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | New York City, NY $320k–$485k/yr Published 3 months ago
Anthropic

Join Anthropic as a Performance Engineer to optimize AI inference systems for throughput, latency, reliability, and correctness.

Anthropic San Francisco, CA | New York City, NY | Seattle, WA $350k–$850k/yr Published 2 months ago
Flexible on stack
Anthropic
Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Washington, DC; San Francisco, CA | New York City, NY $230k–$270k/yr Published 4 months ago
Anthropic

Join Anthropic as a Safeguards Analyst to build enforcement workflows for AI systems against misuse and ensure user safety.

Anthropic Remote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC $285k–$330k/yr Published 2 weeks ago
Anthropic

Lead the revenue accounting team at Anthropic, ensuring compliance and efficiency as the company scales rapidly.

Anthropic San Francisco, CA | New York City, NY | Seattle, WA $300k–$385k/yr Published 1 week ago
Anthropic

Join Anthropic as a Data Scientist to drive data-informed decision-making for our Developer Platform in a mission-driven environment.

Anthropic New York City, NY | Seattle, WA; San Francisco, CA $275k–$370k/yr Published 2 months ago
Flexible on stack
Anthropic

Lead investor relations at Anthropic, shaping engagement and narratives for a growing AI company.

Anthropic San Francisco, CA $425k–$600k/yr Published 2 weeks ago
Anthropic

Lead the Model Exploitation & Fraud team at Anthropic to combat large-scale exploitation of AI systems.

Anthropic San Francisco, CA $375k–$455k/yr Published 2 weeks ago
Anthropic

Build evaluation infrastructure for AI safety systems at Anthropic, focusing on real-world misuse detection.

Anthropic San Francisco, CA | New York City, NY $320k–$485k/yr Published 1 month ago
Flexible on stack
Page 1 of 67