"rl training" Jobs

110 open tech roles matching “rl training”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, reinforcement learning, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 110 results

Cartesia

Join Cartesia as a Researcher to enhance multimodal models through innovative post-training methods and alignment techniques.

Cartesia *HQ - San Francisco, CA Published 10 months ago
Anthropic
Anthropic San Francisco, CA | New York City, NY | Seattle, WA $500k–$850k/yr Published 11 months ago
Twelve Labs

Drive technical direction for training infrastructure and operations within Pegasus at a growing AI company focused on video understanding.

Twelve Labs Seoul, South Korea Published 1 week ago
Mirendil

Join Mirendil as a research engineer to build data systems and execution environments for reinforcement learning.

Mirendil San Francisco $300k–$400k/yr Published 2 months ago
Anthropic
Anthropic San Francisco, CA | New York City, NY $300k–$405k/yr Published 5 months ago
Preference Model

Join Preference Model as a Senior Machine Learning Engineer to design RL environments for advancing ML capabilities.

Preference Model San Francisco, United States Published 2 days ago
Flexible on stack
Preference Model

Join Preference Model as a senior ML Engineer to design RL environments for advancing machine learning capabilities.

Preference Model San Francisco Published 2 weeks ago
Flexible on stack
Anthropic

Join Anthropic as a Research Engineer to advance AI models' silicon design capabilities within a hybrid work environment.

Anthropic San Francisco, CA | New York City, NY $500k–$850k/yr Published 2 months ago
Flexible on stack
Together AI

Join Together AI as a Research Engineer to develop a platform for customizing open-source models with user data.

Together AI San Francisco $200k–$290k/yr Published 2 months ago
Flexible on stack
Preference Model

Join Preference Model as a new graduate Machine Learning Engineer to design and build reinforcement learning environments.

Preference Model San Francisco Published 4 months ago
Flexible on stack
Cursor

Lead a team of engineers to build infrastructure for training and evaluating ML models in a flat, innovative organization.

Cursor San Francisco Published 2 months ago
Heavy meetings
Anthropic
Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | New York City, NY $500k–$850k/yr Published 4 months ago
Anthropic

Join Anthropic as a Staff+ Research Engineer to build and maintain the RL Data Platform for reliable AI systems.

Anthropic San Francisco, CA | New York City, NY $500k–$850k/yr Published 2 weeks ago
Flexible on stack AI-first team
Anthropic

Join Anthropic as a Staff+ Software Engineer to build and maintain the RL Data Platform for reliable AI systems.

Anthropic San Francisco, CA | New York City, NY $320k–$405k/yr Published 2 weeks ago
Flexible on stack AI-first team
Anthropic

Join Anthropic as a Staff Software Engineer to drive reinforcement learning efforts and design systems for coding capabilities.

Anthropic San Francisco, CA | New York City, NY | Seattle, WA $405k–$625k/yr Published 1 month ago