"human preference model" Jobs
2254 open tech roles matching “human preference model”, taken straight from company career pages — not reposted from other job boards. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 2254 results
Own internal products that enhance how Anthropic works with Claude, collaborating with various teams to drive adoption and impact.
Join Anthropic as a Staff Software Engineer to design APIs and frameworks for reinforcement learning environments in a collaborative AI research team.
Lead a team of engineers to build AI-powered cybersecurity products in a fast-paced environment.
Join SpaceX as an ML Engineer to develop AI surrogate models that enhance engineering simulations for launch vehicles and spacecraft.
Join Anthropic as a Staff+ Software Engineer to lead the development of enterprise AI products that enhance organizational workflows.
Join Anthropic as a Safeguards Enforcement Analyst to combat violence and extremism in AI systems.
Join Anthropic as a Research Engineer to design and run large-scale experiments in Reinforcement Learning for AI systems.
Join Beacon Software as an HR Business Partner to optimize HR processes in a fast-paced technology company with significant growth potential.
Own the roadmap for coding data modalities at Handshake AI, navigating ambiguity and leading cross-functional teams.
Join Anthropic as a Product Manager to ensure AI systems are safe and beneficial for users across various platforms.
Join OfferUp as a Principal Product Designer to enhance user experience and brand relevance in a hands-on role.
Lead marketing efforts for Dexterity's Physical AI solutions, driving enterprise demand and category positioning.
Lead Anthropic's cybersecurity policy efforts, engaging with government and regulatory stakeholders to ensure safe AI systems.
Own the technical strategy for config and experimentation infrastructure at Anthropic, enhancing developer productivity in AI systems.