"interpretability" Jobs

761 open tech roles matching “interpretability”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: AI/ML, Python, SQL. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 761 results

Anthropic

Join Anthropic as a Staff+ Software Engineer to build and maintain the RL Data Platform for reliable AI systems.

Anthropic San Francisco, CA | New York City, NY $320k–$405k/yr Published 2 weeks ago
Flexible on stack AI-first team
Anthropic

Join Anthropic as a Performance Engineer to optimize AI inference systems for throughput, latency, reliability, and correctness.

Anthropic San Francisco, CA | New York City, NY | Seattle, WA $350k–$850k/yr Published 3 months ago
Flexible on stack
Perplexity AI

Join Perplexity AI as a senior data scientist to shape product decisions through user behavior analysis and AI-driven insights.

Perplexity AI San Francisco Published 6 months ago
Anthropic

Support the design and deployment of mental health guardrails as a Safeguards Enforcement Analyst at Anthropic.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $245k–$285k/yr Published 1 month ago
Anthropic

Join Anthropic as a Staff+ Research Engineer to build and maintain the RL Data Platform for reliable AI systems.

Anthropic San Francisco, CA | New York City, NY $500k–$850k/yr Published 2 weeks ago
Flexible on stack AI-first team
Anthropic
Anthropic San Francisco, CA | New York City, NY | Washington, DC $230k–$270k/yr Published 6 months ago
Anthropic

Join Anthropic as a Safeguards Analyst to build enforcement workflows for AI systems against misuse and ensure user safety.

Anthropic San Francisco, CA | New York City, NY | Washington, DC $285k–$330k/yr Published 2 months ago
Anthropic

Join Anthropic as a Staff Software Engineer to shape privacy engineering in AI systems at scale.

Anthropic San Francisco, CA | New York City, NY | Seattle, WA $405k–$485k/yr Published 1 month ago
Flexible on stack
Goodfire

Lead the events strategy at Goodfire, creating high-signal experiences for technical audiences in AI.

Goodfire San Francisco, CA Published 3 months ago
Anthropic

Lead the revenue accounting team at Anthropic, ensuring compliance and efficiency as the company scales rapidly.

Anthropic San Francisco, CA | New York City, NY | Seattle, WA $300k–$385k/yr Published 1 month ago
Anthropic

Join Anthropic as a Data Scientist to drive data-informed decision-making for our Developer Platform in a mission-driven environment.

Anthropic New York City, NY | Seattle, WA; San Francisco, CA $275k–$370k/yr Published 3 months ago
Flexible on stack
Anthropic

Lead investor relations at Anthropic, shaping engagement and narratives for a growing AI company.

Anthropic San Francisco, CA $425k–$600k/yr Published 2 months ago
Goodfire

Lead Goodfire's model training organization to develop interpretable AI systems with a world-class team.

Goodfire San Francisco, CA $400k–$600k/yr Published 1 month ago
AI-first team
Anthropic

Lead the Model Exploitation & Fraud team at Anthropic to combat large-scale exploitation of AI systems.

Anthropic San Francisco, CA $375k–$455k/yr Published 2 months ago
Anthropic

Join Anthropic as a Staff Software Engineer to build scalable ML infrastructure for AI safety systems.

Anthropic San Francisco, CA $320k–$485k/yr Published 4 days ago
Flexible on stack
Anthropic

Join Anthropic as a Staff+ Software Engineer to build safety mechanisms for AI systems with a focus on data governance and integrity.

Anthropic San Francisco, CA | New York City, NY $320k–$485k/yr Published 1 month ago
Flexible on stack
Insitro

Join insitro as a Staff Software Engineer to enhance imaging software and ML infrastructure in drug discovery.

Insitro South San Francisco, CA $219k–$233k/yr Published 2 weeks ago
Flexible on stack