"rl training" Jobs

110 open tech roles matching “rl training”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, reinforcement learning, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 110 results

Mirendil

Join Mirendil as a research engineer to build the post-training stack for frontier reasoning models in AI.

Mirendil San Francisco $300k–$400k/yr Published 2 months ago
Preference Model

Join Preference Model as a Research Engineer to advance self-directed learning in large language models within a fast-paced startup.

Preference Model San Francisco, United States Published 2 days ago
Flexible on stack
Preference Model

Join Preference Model as a Research Engineer to advance self-directed learning in machine learning with a focus on RL environments.

Preference Model San Francisco Published 2 weeks ago
Flexible on stack
Reflection AI

Join Reflection AI to build systems that transform pre-trained models into aligned agents in a fast-paced startup environment.

Reflection AI San Francisco, CA Published 1 month ago
Cursor

Join Cursor as a Software Engineer on the RL Data team to create and improve tasks for training coding agents.

Cursor San Francisco Published 2 weeks ago
baseten

Join Baseten as a senior software engineer to develop cutting-edge AI training products and enhance user workflows.

baseten San Francisco Published 7 months ago
Flexible on stack
krea.ai

Join Krea as an ML Researcher to finetune diffusion models and enhance AI creative tools in a collaborative environment.

krea.ai San Francisco Published 1 week ago
Flexible on stack
Mirendil

Join Mirendil as a staff engineer to build the post-training stack for frontier reasoning models in a tech-first startup.

Mirendil San Francisco $300k–$400k/yr Published 2 months ago
Harvey AI

Join Harvey AI as a Research Engineer to drive post-training experiments and enhance legal AI models.

Harvey AI San Francisco $231k–$340k/yr Published 2 months ago
Flexible on stack
Distyl AI

Join Distyl AI as a Senior Applied AI Researcher to redefine AI utilization in enterprise with cutting-edge research and technology.

Distyl AI San Francisco $150k–$250k/yr Published 11 months ago
Flexible on stack
Preference Model

Join Preference Model as a Senior ML Infrastructure Engineer to build scalable infrastructure for post-training research on large language models.

Preference Model San Francisco, United States Published 2 days ago
Flexible on stack
Preference Model

Join Preference Model as a Senior ML Infrastructure Engineer to build scalable infrastructure for post-training research on large language models.

Preference Model San Francisco Published 3 days ago
Flexible on stack
Reflection AI

Join Reflection AI as a Data Quality Engineer to ensure high data standards for AI model training and evaluation.

Reflection AI San Francisco, CA Published 8 months ago
Flexible on stack
Reflection AI

Join Reflection AI as a Forward Deployed Engineer to fine-tune models and work directly with enterprise customers in a dynamic startup environment.

Reflection AI San Francisco, CA Published 4 months ago
Anthropic

Join Anthropic as a Technical Program Manager to drive progress in reinforcement learning research and enhance AI systems.

Anthropic San Francisco, CA | New York City, NY $365k–$435k/yr Published 1 month ago
Anthropic

Join Anthropic as a Research Engineer to enhance AI's coding capabilities through reinforcement learning in a collaborative environment.

Anthropic San Francisco, CA | New York City, NY $500k–$850k/yr Published 3 months ago
Flexible on stack
Reflection AI

Design and operate large-scale GPU infrastructure for model inference and mid-training workloads at Reflection AI.

Reflection AI San Francisco, CA Published 5 months ago
Flexible on stack
Labelbox

Join Labelbox as a Staff ML Engineer to shape AI training environments and systems in a high-impact, fast-paced setting.

Labelbox San Francisco Bay Area $250k–$280k/yr Published 1 month ago
Flexible on stack