"ai model evaluation" Jobs

3713 open tech roles matching “ai model evaluation”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, SQL. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 3713 results

Perplexity AI

Build specialized evals to improve answer quality across Perplexity's products in a high-impact data science role.

Perplexity AI San Francisco Published 2 months ago
Flexible on stack
OpenRouter

Join OpenRouter as an AI Provider Operations & Support Engineer to drive model launches and provider onboarding in a collaborative environment.

OpenRouter Remote (US) Published 7 months ago
Flexible on stack
Block

Join Block as a Senior Machine Learning Engineer to lead model risk management and validate AI systems in a remote-friendly environment.

Block New York, NY, United States of America $194.5k–$343.1k/yr Published 3 months ago
Flexible on stack
Inceptive

Join Inceptive to pioneer AI-designed drugs as part of a collaborative team focused on computational biology and machine learning.

Inceptive Palo Alto, CA $135k–$240k/yr Published 1 month ago
Flexible on stack
Arena

Join Arena as an Applied AI Engineer to integrate AI models and build solutions for cutting-edge AI teams.

Arena Bay Area Published 2 days ago
Flexible on stack
Fireworks AI

Join Fireworks AI as a Member of Technical Staff to enhance model evaluation and fine-tuning workflows in a fast-paced generative AI environment.

Fireworks AI San Mateo Published 10 months ago
Anthropic

Join Anthropic as a Product Designer to build and evaluate prompts for AI systems, ensuring alignment with user expectations and safety.

Anthropic San Francisco, CA $305k–$385k/yr Published 1 week ago
Flexible on stack
Arena

Join Arena as a Site Reliability Engineer to build core infrastructure for AI model evaluations at scale.

Arena Bay Area Published 1 month ago
Flexible on stack
Cartesia

Join Cartesia as a Product and Research Operations Manager to design and scale a global evaluation workforce for AI.

Cartesia *HQ - San Francisco, CA Published 2 weeks ago
AI-first team
Inworld AI

Join Inworld AI as a Staff/Principal Research Scientist to innovate in real-time voice models and AI applications.

Inworld AI Mountain View, California, USA $270k–$500k/yr Published 3 years ago
Protege

Join Protege as a Forward Deployed Machine Learning Engineer to build the technical foundation for AI training data evaluations.

Protege Remote Published 1 month ago
Snorkel AI

Lead a team of researchers to enhance data evaluation and analysis methods for AI model performance at Snorkel AI.

Snorkel AI New York City, NY (Hybrid); San Francisco, CA (Hybrid); United States (Remote) $275k–$425k/yr Published 3 months ago
AI-first team
Abridge

Lead the product strategy for Abridge's AI/ML evaluation platform, ensuring quality and efficiency across multiple product teams.

Abridge SF Office Published 2 months ago
Periodic Labs

Join Periodic Labs as a Midtraining Research Engineer to enhance scientific reasoning in AI models for groundbreaking discoveries.

Periodic Labs Menlo Park, CA $250k–$350k/yr Published 1 month ago
Harvey

Lead the design and development of systems powering AI requests at Harvey, ensuring high reliability and operational excellence.

Harvey San Francisco $193.4k–$290k/yr Published 6 days ago
Flexible on stack 60% coding
Reflection AI

Join Reflection AI as a Research Software Engineer to build secure infrastructure for sensitive model evaluations in a fast-paced startup environment.

Reflection AI New York, NY Published 1 month ago
Flexible on stack
Cartesia

Join Cartesia as a Researcher to enhance multimodal models through innovative post-training methods and alignment techniques.

Cartesia *HQ - San Francisco, CA Published 10 months ago
Figma

Join Figma as a Design Program Manager to enhance AI evaluation processes and drive design quality.

Figma London, England Published 3 days ago
Reflection AI

Lead the post-training and evaluation capabilities for large language models in a dynamic AI research lab.

Reflection AI New York, NY Published 10 months ago