"llm evaluation" Jobs
1533 open tech roles matching “llm evaluation”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, TypeScript. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 1533 results
Join Krea as an ML Researcher to finetune diffusion models and enhance AI creative tools in a collaborative environment.
Join LILT as a Forward Deployed Engineer to integrate AI solutions for complex clients and enhance global communication.
Join Crosby as an AI Developer Experience Engineer to enhance developer productivity in building AI features for legal systems.
Join Langfuse as a Senior Product Engineer to build impactful open-source AI applications in a small, dynamic team.
Lead the AgentControl Evaluations team at LaunchDarkly, focusing on offline evaluations and AI-powered systems.
Join Edra as a Forward Deployed AI Engineer to build LLM-powered systems for enterprise customers in a hands-on engineering role.
Join Preference Model as a Senior ML Infrastructure Engineer to build scalable infrastructure for post-training research on large language models.
Join Preference Model as a Senior ML Infrastructure Engineer to build scalable infrastructure for post-training research on large language models.
Lead the engineering team building LangSmith, an observability and evaluation platform for LLM applications.
Lead the product strategy for Abridge's AI/ML evaluation platform, ensuring quality and efficiency across multiple product teams.
Join Protege as a Forward Deployed Machine Learning Engineer to build the technical foundation for AI training data evaluations.
Join Langfuse as a Senior Software Engineer to shape the developer experience of our widely-used SDKs in the AI space.