"evaluation frameworks" Jobs

581 open tech roles matching “evaluation frameworks”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, SQL. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 581 results

Anthropic

Build evaluation infrastructure for AI safety systems at Anthropic, focusing on real-world misuse detection.

Anthropic San Francisco, CA | New York City, NY $320k–$485k/yr Published 2 months ago
Flexible on stack
Figma

Lead the evaluation of Figma's AI-powered experiences to ensure quality and effectiveness in product features.

Figma San Francisco, CA • New York, NY • United States $258k–$348k/yr Published 1 month ago
Anthropic
Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Washington, DC; San Francisco, CA | New York City, NY $230k–$270k/yr Published 5 months ago
Anthropic

Lead the Secure Frameworks Research team at Anthropic to build high-leverage security frameworks for AI systems.

Anthropic San Francisco, CA | Seattle, WA $405k–$485k/yr Published 1 month ago
Heavy meetings
Asana

Lead the Agent Context team at Asana to enhance AI-driven retrieval systems at enterprise scale.

Asana New York City $264k–$300k/yr Published 1 month ago
Heavy meetings
Cockroach Labs

Join Cockroach Labs as a Value Engineer to drive financial narratives and ROI models for enterprise customers adopting CockroachDB.

Cockroach Labs London, UK; New York, NY $145k–$190k/yr Published 2 months ago
Figma
Figma San Francisco, CA • New York, NY • United States $153k–$376k/yr Published 9 months ago
Anthropic

Drive the clinical AI research agenda for Anthropic's global health work, ensuring AI tools are safe and effective in low-resource settings.

Anthropic San Francisco, CA | New York City, NY $215k–$300k/yr Published 1 month ago
Anthropic
Anthropic New York City, NY; San Francisco, CA; Seattle, WA $350k–$850k/yr Published 7 months ago
Anthropic

Join Anthropic's Safeguards team as a Red Team Engineer to enhance the safety of AI systems through adversarial testing.

Anthropic Remote-Friendly (Travel Required) | San Francisco, CA $320k–$405k/yr Published 1 month ago
Flexible on stack
Vanta

Lead the development of federal compliance frameworks and automated GRC solutions for Vanta's public sector initiatives.

Vanta Remote U.S. Published 1 week ago
Elastic

Join Elastic as an RFP Proposal Analyst to streamline proposal processes and support sales operations in a hybrid work environment.

Elastic United States $61.9k–$97.9k/yr Published 3 weeks ago
AI-first team
Asana

Lead Asana's AI Platform organization to drive strategy and execution for AI experiences across the company.

Asana San Francisco $306k–$360k/yr Published 2 weeks ago