"rl training" Jobs
220 open tech roles matching “rl training”, taken straight from company career pages — not reposted from other job boards. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 220 results
Join Mirendil as a research engineer to build the post-training stack for frontier reasoning models in AI.
Join Preference Model as a Research Engineer to advance self-directed learning in large language models within a fast-paced startup.
Join Preference Model as a Research Engineer to advance self-directed learning in machine learning with a focus on RL environments.
Lead the post-training and evaluation capabilities for large language models in a dynamic AI research lab.
Join Periodic Labs as a Midtraining Research Engineer to enhance scientific reasoning in AI models for groundbreaking discoveries.
Join Reflection AI to build systems that transform pre-trained models into aligned agents in a fast-paced startup environment.
Join Cursor as a Software Engineer on the RL Data team to create and improve tasks for training coding agents.
Join Baseten as a senior software engineer to develop cutting-edge AI training products and enhance user workflows.
Join Krea as an ML Researcher to finetune diffusion models and enhance AI creative tools in a collaborative environment.
Join Mirendil as a staff engineer to build the post-training stack for frontier reasoning models in a tech-first startup.
Join Harvey AI as a Research Engineer to drive post-training experiments and enhance legal AI models.
Join Distyl AI as a Senior Applied AI Researcher to redefine AI utilization in enterprise with cutting-edge research and technology.
Join Hippocratic AI as an Applied Scientist to enhance AI safety and clinical reasoning through reinforcement learning.
Join Preference Model as a Senior ML Infrastructure Engineer to build scalable infrastructure for post-training research on large language models.
Join Preference Model as a Senior ML Infrastructure Engineer to build scalable infrastructure for post-training research on large language models.
Join Reflection AI as a Data Quality Engineer to ensure high data standards for AI model training and evaluation.
Join Reflection AI as a Research Software Engineer to bridge research and production in cutting-edge AI training systems.
Join Reflection AI as a Forward Deployed Engineer to fine-tune models and work directly with enterprise customers in a dynamic startup environment.
Join DeepL as a Senior Research Scientist to innovate in reinforcement learning and shape the future of AI technology.