"ai model evaluation" Jobs
3713 open tech roles matching “ai model evaluation”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, SQL. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 3713 results
Build specialized evals to improve answer quality across Perplexity's products in a high-impact data science role.
Join OpenRouter as an AI Provider Operations & Support Engineer to drive model launches and provider onboarding in a collaborative environment.
Join Block as a Senior Machine Learning Engineer to lead model risk management and validate AI systems in a remote-friendly environment.
Join Inceptive to pioneer AI-designed drugs as part of a collaborative team focused on computational biology and machine learning.
Join Arena as an Applied AI Engineer to integrate AI models and build solutions for cutting-edge AI teams.
Join Fireworks AI as a Member of Technical Staff to enhance model evaluation and fine-tuning workflows in a fast-paced generative AI environment.
Join Anthropic as a Product Designer to build and evaluate prompts for AI systems, ensuring alignment with user expectations and safety.
Join Cartesia as a Product and Research Operations Manager to design and scale a global evaluation workforce for AI.
Join Inworld AI as a Staff/Principal Research Scientist to innovate in real-time voice models and AI applications.
Join Protege as a Forward Deployed Machine Learning Engineer to build the technical foundation for AI training data evaluations.
Lead a team of researchers to enhance data evaluation and analysis methods for AI model performance at Snorkel AI.
Lead the product strategy for Abridge's AI/ML evaluation platform, ensuring quality and efficiency across multiple product teams.
Join Periodic Labs as a Midtraining Research Engineer to enhance scientific reasoning in AI models for groundbreaking discoveries.
Join Reflection AI as a Research Software Engineer to build secure infrastructure for sensitive model evaluations in a fast-paced startup environment.
Join Cartesia as a Researcher to enhance multimodal models through innovative post-training methods and alignment techniques.
Join Figma as a Design Program Manager to enhance AI evaluation processes and drive design quality.
Lead the post-training and evaluation capabilities for large language models in a dynamic AI research lab.