Join Anthropic as a Technical Program Manager to drive progress in reinforcement learning research and enhance AI systems.
Posted by employer 2 months ago
First seen on Joblaze 2 months ago
Last verified on the company career page 23 hours ago
Skills & Technologies
Requirements
Not disclosed in this posting: years of experience.
Benefits
Joblaze summary
In the role of Technical Program Manager for the reinforcement learning team at Anthropic, the individual will oversee the systems and processes that accelerate research and production in AI. This position requires a strong background in machine learning engineering or research, along with the ability to manage complex data pipelines and coordinate across various teams. Ideal candidates are those who thrive in fast-paced environments and possess excellent communication skills to influence technical stakeholders. Anthropic emphasizes collaboration and aims to push the boundaries of AI research, making this role pivotal in their mission.
Joblaze insights
Quick facts
From the original posting
Our Reinforcement Learning teams are central to advancing our AI systems, contributing to every Claude model and driving the autonomy and coding gains in our latest releases. The work spans computer use, code generation through RL, fundamental RL research for large language models, scalable infrastructure and training methodologies, and model reasoning.
Research TPM team supports the full model development lifecycle, from pre-training through post-training, operating at the frontier of AI development.
As a Technical Program Manager on the reinforcement learning team, you will own the systems and programs that determine how fast our research moves: a trustworthy read on the state of RL research, the review and prioritization processes that turn that read into critical decision for production RL runs. Strong candidates should have an ML engineering or research background and have grown into program leadership. You'll need real technical depth: the ability to debug data pipelines, read RL transcripts to spot issues, and make allocation and quality decisions in real time when research or production runs hit problems. You'll need organizational effectiveness in equal measure: the ability to navigate a fast-growing organization, quickly identify the critical people and teams across research, infrastructure, product, and data operations, and coordinate across them without losing velocity.
Join us in our mission to build AI systems that are safe, reliable, and beneficial to humanity.
The annual compensation range for this role is listed below.
Standard company text repeated across Anthropic's postings is omitted here.