"eval frameworks" Jobs
2718 open tech roles matching “eval frameworks”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, AWS. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 2718 results
Join Macroscope as an Applied ML Engineer to enhance machine learning systems in a collaborative startup environment.
Join Reflection AI as a Research Program Manager to build foundational infrastructure for model evaluations and safety in AI.
Join Exa as an ML evals engineer to design and build evaluation frameworks for a groundbreaking AI search engine.
Join Fireworks AI as a senior AI Field Engineer to build production systems and engage with enterprise customers on generative AI solutions.
Join Edra as an AI Engineer to build complex LLM-based systems that enhance enterprise AI processes.
Join Rox as a Founding Applied Research Engineer to shape the future of applied AI with a focus on real-world production challenges.
Lead the Secure Frameworks Research team at Anthropic to build high-leverage security frameworks for AI systems.
Join Fireworks AI as an AI Field Engineer to build production systems for generative AI with large organizations across EMEA.
Join MongoDB as a Software Engineer to lead a team in building automated testing frameworks for application modernization.
Build secure AI agents for enterprise internal systems in a 12-week residency in Mountain View, California.
Conduct critical analysis and develop evaluation frameworks to improve AI model capabilities in a fast-paced startup environment.
Lead the product strategy for Abridge's AI/ML evaluation platform, ensuring quality and efficiency across multiple product teams.
Join Triomics as an ML Evaluation Engineer to ensure model quality and stability in clinical AI systems.
Join Fireworks AI as a senior AI Field Engineer to build production systems for innovative AI-native companies.
Join Edra as an AI Engineer to build LLM-powered systems that enhance enterprise AI processes in a collaborative, onsite environment.
Build eval systems and quality metrics for a personalized AI assistant at a startup in San Francisco.