"reinforcement learning from human feedback" Jobs

6 open tech roles matching “reinforcement learning from human feedback”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: PyTorch, Large Language Models, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 6 of 6 results

Synthesia

Join Synthesia as a Staff Research Engineer to shape the future of interactive multimodal systems in AI video communication.

Synthesia Europe Published 1 month ago
Flexible on stack
DeepL

Join DeepL as a Senior Research Scientist to innovate in reinforcement learning and shape the future of AI technology.

DeepL London Published 1 month ago
Flexible on stack
DeepL

Lead research on fine-tuning and steerability of LLM-based translation models in a collaborative AI-focused environment.

DeepL London Published 1 month ago
Flexible on stack
hyperexponential

Join hyperexponential as the first Head of Revenue Enablement to build a dedicated function that enhances seller effectiveness and drives revenue.

hyperexponential London (hybrid) Published 3 weeks ago
Anthropic

Lead the global reseller enablement strategy for Anthropic, designing and launching training programs to ensure successful partner outcomes.

Anthropic London, UK £160k–£215k/yr Published 2 weeks ago
Glossier

Support the store's leadership team in creating exceptional experiences and fostering an inclusive environment as a Key Lead at Glossier in London.

Glossier London, UK £15–£16/hr Published 4 months ago