"rlhf" Jobs

71 open tech roles matching “rlhf”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Machine Learning, PyTorch. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 71 results

Hippocratic AI

Join Hippocratic AI as an Applied Scientist to enhance AI safety and clinical reasoning through reinforcement learning.

Hippocratic AI Menlo Park, CA Published 1 month ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA | New York City, NY | Seattle, WA $500k–$850k/yr Published 11 months ago
Anthropic

Join Anthropic as a Research Engineer to enhance AI's coding capabilities through reinforcement learning in a collaborative environment.

Anthropic San Francisco, CA | New York City, NY $500k–$850k/yr Published 3 months ago
Flexible on stack
Anthropic

Join Anthropic as a Staff+ Software Engineer to build and maintain the RL Data Platform for reliable AI systems.

Anthropic San Francisco, CA | New York City, NY $320k–$405k/yr Published 2 weeks ago
Flexible on stack AI-first team
Anthropic

Join Anthropic as a Staff+ Research Engineer to build and maintain the RL Data Platform for reliable AI systems.

Anthropic San Francisco, CA | New York City, NY $500k–$850k/yr Published 2 weeks ago
Flexible on stack AI-first team
Anthropic

Join Anthropic as a Technical Program Manager to drive progress in reinforcement learning research and enhance AI systems.

Anthropic San Francisco, CA | New York City, NY $365k–$435k/yr Published 1 month ago
Lyft

Join Lyft as a Machine Learning Engineer to develop AI agents that enhance safety and customer care for riders and drivers.

Lyft Toronto, Canada CA$118.8k–CA$148.5k/yr Published 1 month ago
Flexible on stack 70% coding
Reflection AI

Join Reflection AI as a Forward Deployed Engineer to fine-tune models and work directly with enterprise customers in a dynamic startup environment.

Reflection AI San Francisco, CA Published 4 months ago
Reflection AI

Own the red-teaming and adversarial evaluation pipeline for Reflection’s models to ensure safety and reliability.

Reflection AI San Francisco, CA Published 8 months ago
DeepL

Lead research on fine-tuning and steerability of LLM-based translation models in a collaborative AI-focused environment.

DeepL London Published 1 month ago
Flexible on stack
DeepL

Join DeepL as a Senior Research Scientist to innovate in reinforcement learning and shape the future of AI technology.

DeepL London Published 1 month ago
Flexible on stack
Iambic Therapeutics

Join Iambic Therapeutics as a Machine Learning Scientist to innovate AI-based drug discovery with multimodal models.

Iambic Therapeutics UK Office Published 2 weeks ago
Flexible on stack
Hedra

Join Hedra as a Research Scientist to lead innovative research in generative AI and physical systems with access to large-scale compute.

Hedra San Francisco Published 5 months ago
Iambic Therapeutics

Join Iambic Therapeutics as a Machine Learning Scientist to innovate AI-based drug discovery with multimodal models.

Iambic Therapeutics Boston Office Published 2 weeks ago
Flexible on stack
DeepL

Lead fine-tuning and development of multimodal models for document translation at DeepL.

DeepL London Published 1 month ago
Flexible on stack
Harvey AI

Join Harvey AI as a Research Engineer to drive post-training experiments and enhance legal AI models.

Harvey AI San Francisco $231k–$340k/yr Published 2 months ago
Flexible on stack
Genesis Molecular AI

Join Genesis Molecular AI as an ML Research Engineer to develop cutting-edge foundation models for drug discovery.

Genesis Molecular AI San Mateo, CA Published 1 year ago
Flexible on stack 70% coding