"reinforcement learning from human feedback" Jobs
6 open tech roles matching “reinforcement learning from human feedback”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: PyTorch, Large Language Models, AI/ML. Every listing is re-checked daily and closed roles are removed.
Showing 6 of 6 results
Join Synthesia as a Staff Research Engineer to shape the future of interactive multimodal systems in AI video communication.
Join DeepL as a Senior Research Scientist to innovate in reinforcement learning and shape the future of AI technology.
Lead research on fine-tuning and steerability of LLM-based translation models in a collaborative AI-focused environment.
Join hyperexponential as the first Head of Revenue Enablement to build a dedicated function that enhances seller effectiveness and drives revenue.
Lead the global reseller enablement strategy for Anthropic, designing and launching training programs to ensure successful partner outcomes.
Support the store's leadership team in creating exceptional experiences and fostering an inclusive environment as a Key Lead at Glossier in London.
Related searches