"rl training" Jobs

220 open tech roles matching “rl training”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, reinforcement learning, PyTorch. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 220 results

Inflection AI

Lead model training and post-training strategies for emotionally intelligent AI at Inflection AI.

Inflection AI Palo Alto, California, United States $400k–$550k/yr Published 2 months ago
Traversal

Join Traversal as an AI Researcher to enhance AI agents for diagnosing and resolving production incidents in complex environments.

Traversal New York $160k–$300k/yr Published 1 year ago
Flexible on stack 70% coding
Fireworks AI

Join Fireworks AI as a Product Manager to shape the future of AI training products for enterprises.

Fireworks AI San Mateo Published 3 weeks ago
Anthropic

Join Anthropic as a Research Engineer to advance AI models' silicon design capabilities within a hybrid work environment.

Anthropic San Francisco, CA | New York City, NY $500k–$850k/yr Published 2 months ago
Flexible on stack
Coreweave

Join CoreWeave as a Senior Software Engineer II to build cutting-edge research infrastructure for AI labs.

Coreweave Sunnyvale, CA / Bellevue, WA $182k–$242k/yr Published 2 months ago
Flexible on stack
Sarvam AI

Join Sarvam AI as an ML Researcher to drive foundational model research with high autonomy and impact.

Sarvam AI Bengaluru Published 3 months ago
Flexible on stack
Together AI

Join Together AI as a Research Engineer to develop a platform for customizing open-source models with user data.

Together AI San Francisco $200k–$290k/yr Published 2 months ago
Flexible on stack
Bugcrowd

Join Bugcrowd as a Reinforcement Learning Engineer to build AI systems that enhance cybersecurity through innovative training environments.

Bugcrowd Remote - US $176.4k–$242.6k/yr Published 1 month ago
Flexible on stack
Lyft

Join Lyft as a Machine Learning Engineer to develop AI agents that enhance safety and customer care for riders and drivers.

Lyft Toronto, Canada CA$118.8k–CA$148.5k/yr Published 1 month ago
Flexible on stack 70% coding
Preference Model

Join Preference Model as a new graduate Machine Learning Engineer to design and build reinforcement learning environments.

Preference Model San Francisco Published 4 months ago
Flexible on stack
Cursor

Lead a team of engineers to build infrastructure for training and evaluating ML models in a flat, innovative organization.

Cursor San Francisco Published 2 months ago
Heavy meetings
Anthropic
Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | New York City, NY $500k–$850k/yr Published 4 months ago
Ricursive Intelligence

Join Ricursive Intelligence to conduct novel AI research and work on LLM modeling and scaling in a hands-on startup environment.

Ricursive Intelligence Palo Alto Published 7 months ago
Anthropic

Join Anthropic as a Staff+ Research Engineer to build and maintain the RL Data Platform for reliable AI systems.

Anthropic San Francisco, CA | New York City, NY $500k–$850k/yr Published 2 weeks ago
Flexible on stack AI-first team
Anthropic

Join Anthropic as a Research Engineer to design and run large-scale experiments in Reinforcement Learning for AI systems.

Anthropic London, UK £375k–£640k/yr Published 2 months ago
Flexible on stack
Skild AI

Join Skild AI as a Machine Learning Engineer to develop cutting-edge reinforcement learning algorithms for robotic applications.

Skild AI San Mateo, Pittsburgh $100k–$300k/yr Published 1 year ago
Flexible on stack