"evaluation framework" Jobs

253 open tech roles matching “evaluation framework”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: AI/ML, Python, TypeScript. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 253 results

Harvey AI

Join Harvey as a Senior Product Operations Manager to build and scale the evaluation engine for a global AI platform.

Harvey AI San Francisco Published 2 months ago
AI-first team
Figma

Lead the evaluation of Figma's AI-powered experiences to ensure quality and effectiveness in product features.

Figma San Francisco, CA • New York, NY • United States $258k–$348k/yr Published 1 month ago
Anthropic

Build evaluation infrastructure for AI safety systems at Anthropic, focusing on real-world misuse detection.

Anthropic San Francisco, CA | New York City, NY $320k–$485k/yr Published 2 months ago
Flexible on stack
Anthropic
Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Washington, DC; San Francisco, CA | New York City, NY $230k–$270k/yr Published 5 months ago
Anthropic

Lead the Secure Frameworks Research team at Anthropic to build high-leverage security frameworks for AI systems.

Anthropic San Francisco, CA | Seattle, WA $405k–$485k/yr Published 1 month ago
Heavy meetings
Cursor

Lead a team of engineers to build infrastructure for training and evaluating ML models in a flat, innovative organization.

Cursor San Francisco Published 1 month ago
Heavy meetings
Asana

Lead Asana's AI Platform organization to drive strategy and execution for AI experiences across the company.

Asana San Francisco $306k–$360k/yr Published 2 weeks ago
Anthropic
Anthropic New York City, NY; San Francisco, CA; Seattle, WA $350k–$850k/yr Published 7 months ago
Anthropic

Drive the clinical AI research agenda for Anthropic's global health work, ensuring AI tools are safe and effective in low-resource settings.

Anthropic San Francisco, CA | New York City, NY $215k–$300k/yr Published 1 month ago
Anthropic

Join Anthropic's Safeguards team as a Red Team Engineer to enhance the safety of AI systems through adversarial testing.

Anthropic Remote-Friendly (Travel Required) | San Francisco, CA $320k–$405k/yr Published 1 month ago
Flexible on stack
Anthropic

Join Anthropic as a Research Engineer to enhance AI capabilities in finance, healthcare, and legal domains through applied research and data sourcing.

Anthropic San Francisco, CA | New York City, NY | Seattle, WA $350k–$850k/yr Published 2 months ago