"reinforcement learning from human feedback" Jobs
60 open tech roles matching “reinforcement learning from human feedback”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: AI/ML, Python, reinforcement learning. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 60 results
Join Hippocratic AI as an Applied Scientist to enhance AI safety and clinical reasoning through reinforcement learning.
Join Handshake as a Software Engineer II to build Reinforcement Learning Environments for AI model training in a growing team in India.
Join Genmo as a Research Scientist to lead initiatives in alignment and post-training for large-scale video generation models.
Join Anthropic as a Research Engineer to enhance AI capabilities in finance, healthcare, and legal domains through applied research and data sourcing.
Join Arena as a Machine Learning Scientist to evaluate AI models and contribute to impactful research in a collaborative environment.
Join Reflection AI as a hands-on technical staff member to enhance model performance through data-driven evaluations and feedback loops.
Join Synthesia as a Staff Research Engineer to shape the future of interactive multimodal systems in AI video communication.
Join DeepL as a Senior Research Scientist to innovate in reinforcement learning and shape the future of AI technology.
Join Airbnb as a Senior Machine Learning Engineer to innovate customer service with cutting-edge AI technologies.
Join Fireworks AI as an Applied Machine Learning Engineer to bridge AI research and real-world applications in a fast-growing team.
Join Fireworks AI as an Applied Machine Learning Engineer to bridge AI research and real-world applications in a fast-growing team.
Own the roadmap for coding data modalities at Handshake AI, navigating ambiguity and leading cross-functional teams.
Join Doctronic as a Senior AI Engineer to build intelligent systems that enhance clinical decision-making in healthcare.
Conduct original research on large language models to advance understanding and routing optimization at OpenRouter.
Lead research on fine-tuning and steerability of LLM-based translation models in a collaborative AI-focused environment.
Lead open-source research efforts at Arena Intelligence to advance AI model evaluation and transparency.
Join Distyl AI as a Research Engineer to bridge AI research and production systems, enhancing AI behavior and reliability.
Related searches