"ai model evaluation" Jobs
1268 open tech roles matching “ai model evaluation”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: AI/ML, Python, TypeScript. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 1268 results
Join Mirendil as a research engineer to build evaluation infrastructure for frontier AI models.
Join Cartesia as a lead researcher to design evaluation frameworks for next-generation AI models.
Build eval systems and quality metrics for a personalized AI assistant at a startup in San Francisco.
Conduct critical analysis and develop evaluation frameworks to improve AI model capabilities in a fast-paced startup environment.
Join Baseten as a Software Engineer to drive model performance systems at the intersection of HPC and LLM engineering.
Join Exa as an ML evals engineer to design and build evaluation frameworks for a groundbreaking AI search engine.
Join Anthropic as a Cyber Evaluations Engineer to design and run evaluations for AI systems, ensuring their safety and robustness.
Join Krea as an ML Researcher to finetune diffusion models and enhance AI creative tools in a collaborative environment.
Join Distyl AI as a Senior AI Engineer to design evaluation frameworks that enhance AI systems in production.
Drive model launches and improve coding performance as a Product Manager on Claude Code's model performance team.
Join Cartesia as a Researcher to advance neural network architecture design in a collaborative, innovative environment.
Lead the Evals team at Cursor to create high-signal evaluation datasets and tools for coding agents.
Join Benchling as a Research Engineer to improve AI models for scientific applications in a fast-paced, collaborative environment.
Join Preference Model as a Research Engineer to advance self-directed learning in large language models within a fast-paced startup.
Join Anthropic as a Tech Lead to build reliable AI evaluation systems in a hybrid work environment.