"ai model evaluation" Jobs
1266 open tech roles matching “ai model evaluation”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: AI/ML, Python, TypeScript. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 1266 results
Join Mirendil as a research engineer to build evaluation infrastructure for frontier AI models.
Join Cartesia as a lead researcher to design evaluation frameworks for next-generation AI models.
Build eval systems and quality metrics for a personalized AI assistant at a startup in San Francisco.
Conduct critical analysis and develop evaluation frameworks to improve AI model capabilities in a fast-paced startup environment.
Join Baseten as a Software Engineer to drive model performance systems at the intersection of HPC and LLM engineering.
Join Exa as an ML evals engineer to design and build evaluation frameworks for a groundbreaking AI search engine.
Join Anthropic as a Cyber Evaluations Engineer to design and run evaluations for AI systems, ensuring their safety and robustness.
Join Distyl AI as a Senior AI Engineer to design evaluation frameworks that enhance AI systems in production.
Join Krea as an ML Researcher to finetune diffusion models and enhance AI creative tools in a collaborative environment.
Drive model launches and improve coding performance as a Product Manager on Claude Code's model performance team.
Lead the Evals team at Cursor to create high-signal evaluation datasets and tools for coding agents.
Join Cartesia as a Researcher to advance neural network architecture design in a collaborative, innovative environment.
Join Benchling as a Research Engineer to improve AI models for scientific applications in a fast-paced, collaborative environment.
Join Preference Model as a Research Engineer to advance self-directed learning in large language models within a fast-paced startup.
Join Anthropic as a Tech Lead to build reliable AI evaluation systems in a hybrid work environment.