"evaluation framework" Jobs
1140 open tech roles matching “evaluation framework”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, AWS. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 1140 results
Join DEFCON AI as a Model Test and Measurement Engineer to ensure AI systems perform reliably and transparently in a fully remote role.
Join Arena as a Software Engineer to build core infrastructure for AI model evaluation in a fast-paced startup environment.
Build eval systems and quality metrics for a personalized AI assistant at a startup in San Francisco.
Join Reflection AI as a Research Program Manager to build foundational infrastructure for model evaluations and safety in AI.
Join Distyl AI as an Applied AI Researcher to redefine software usage and drive innovative benchmarking in AI systems.
Join Abridge as a Research Scientist to evaluate the impact of ambient AI on healthcare outcomes in a fast-paced startup environment.
Own the full lifecycle of revenue generation in a senior role at Arena, focusing on AI model evaluation.
Serve as a Subject Matter Expert on election integrity and fraud, ensuring Reflection's models meet safety and compliance standards.
Join Fieldguide as a Senior AI Engineer to build reliable AI agents for audit workflows in a remote-first environment.
Join Reflection AI as a Sr. Enterprise Risk Governance Specialist to shape risk management frameworks in a dynamic AI environment.
Join Arena as an Associate General Counsel to lead privacy compliance and advise on AI product development in a mission-driven startup.
Join Macroscope as an Applied ML Engineer to enhance machine learning systems in a collaborative startup environment.
Join Redpanda as a Senior Product Engineer to build a policy engine that enhances security for AI agents.
Build production-grade AI systems end-to-end at Hilbert, a fast-growing company in San Francisco.
Join Rox as a Founding Applied Research Engineer to shape the future of applied AI with a focus on real-world production challenges.
Join Fieldguide as a Senior AI Engineer to build AI agents for complex audit workflows in a remote-first environment.
Build autonomous AI agents for go-to-market operations in a senior engineering role at Anthropic.
Lead quality management for AI deliverables at Lilt, ensuring high standards in a hybrid work environment.