"reinforcement learning from human feedback" Jobs

60 open tech roles matching “reinforcement learning from human feedback”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: AI/ML, Python, reinforcement learning. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 60 results

Hippocratic AI

Join Hippocratic AI as an Applied Scientist to enhance AI safety and clinical reasoning through reinforcement learning.

Hippocratic AI Menlo Park, CA Published 1 month ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA | New York City, NY $300k–$405k/yr Published 5 months ago
Handshake

Join Handshake as a Software Engineer II to build Reinforcement Learning Environments for AI model training in a growing team in India.

Handshake India Published 1 month ago
Flexible on stack
Genmo

Join Genmo as a Research Scientist to lead initiatives in alignment and post-training for large-scale video generation models.

Genmo San Francisco HQ Published 4 days ago
Flexible on stack
Anthropic

Join Anthropic as a Research Engineer to enhance AI capabilities in finance, healthcare, and legal domains through applied research and data sourcing.

Anthropic San Francisco, CA | New York City, NY | Seattle, WA $350k–$850k/yr Published 2 months ago
Arena

Join Arena as a Machine Learning Scientist to evaluate AI models and contribute to impactful research in a collaborative environment.

Arena Bay Area Published 8 months ago
Flexible on stack 70% coding
Reflection AI

Join Reflection AI as a hands-on technical staff member to enhance model performance through data-driven evaluations and feedback loops.

Reflection AI New York, NY Published 2 weeks ago
Anthropic
Anthropic New York City, NY; San Francisco, CA; Seattle, WA $350k–$850k/yr Published 7 months ago
Synthesia

Join Synthesia as a Staff Research Engineer to shape the future of interactive multimodal systems in AI video communication.

Synthesia Europe Published 1 month ago
Flexible on stack
DeepL

Join DeepL as a Senior Research Scientist to innovate in reinforcement learning and shape the future of AI technology.

DeepL London Published 1 month ago
Flexible on stack
Airbnb

Join Airbnb as a Senior Machine Learning Engineer to innovate customer service with cutting-edge AI technologies.

Airbnb Remote-USA $196k–$227k/yr Published 2 months ago
Flexible on stack
Fireworks AI

Join Fireworks AI as an Applied Machine Learning Engineer to bridge AI research and real-world applications in a fast-growing team.

Fireworks AI San Mateo Published 1 year ago
Flexible on stack
Fireworks AI

Join Fireworks AI as an Applied Machine Learning Engineer to bridge AI research and real-world applications in a fast-growing team.

Fireworks AI Singapore Published 1 month ago
Flexible on stack
Handshake

Own the roadmap for coding data modalities at Handshake AI, navigating ambiguity and leading cross-functional teams.

Handshake San Francisco, CA Published 2 months ago
Doctronic

Join Doctronic as a Senior AI Engineer to build intelligent systems that enhance clinical decision-making in healthcare.

Doctronic New York City Published 1 month ago
OpenRouter

Conduct original research on large language models to advance understanding and routing optimization at OpenRouter.

OpenRouter Remote (US) Published 2 months ago
Flexible on stack
DeepL

Lead research on fine-tuning and steerability of LLM-based translation models in a collaborative AI-focused environment.

DeepL London Published 1 month ago
Flexible on stack
Arena

Lead open-source research efforts at Arena Intelligence to advance AI model evaluation and transparency.

Arena Bay Area Published 8 months ago
Flexible on stack
Distyl AI

Join Distyl AI as a Research Engineer to bridge AI research and production systems, enhancing AI behavior and reliability.

Distyl AI San Francisco $150k–$250k/yr Published 2 months ago