← Back to results

Engineering Manager, ML

Lead a team of engineers to build infrastructure for training and evaluating ML models in a flat, innovative organization.

Location
San Francisco
Compensation
Not disclosed
Level
lead
Type
full time

Role intensity

20% coding — mostly leadership/strategy

AI in the day-to-day

Engineers use Cursor to move fast and debug alongside their team.

Joblaze summary

The Engineering Manager for ML at Cursor will oversee a team dedicated to developing the infrastructure necessary for training, testing, and evaluating machine learning models. This role requires a strong foundation in infrastructure and distributed systems, along with the ability to navigate the complexities of model behavior and system performance. Ideal candidates will have experience leading engineering teams in ML environments and a passion for technical work, as they will remain hands-on with coding and debugging. Cursor's flat organizational structure fosters a collaborative atmosphere where innovative ideas and spirited discussions are encouraged.

Joblaze insights

Quick facts

What's the tech stack?
Joblaze extracted these technologies from the posting: AI/ML, distributed systems, reinforcement learning, infrastructure.
What seniority level is this role?
Cursor targets lead candidates for this position.
Is this full-time or contract?
Full-time for this Engineering Manager, ML role at Cursor.

From the original posting

Our mission is to automate coding. The first step in our journey is to build the best tool for professional programmers, using a combination of inventive research, design, and engineering. Our organization is very flat, and our team is small and talent dense. We particularly like people who are truth-seeking, passionate, and creative. We enjoy spirited debate, crazy ideas, and shipping code.

About the Role

You will lead a team of engineers building the infrastructure used to train, test, and evaluate our models. This is one of the few places at Cursor where infrastructure and model behavior meet directly: when something breaks, it's rarely obvious whether it's a systems bug or the model doing exactly what it was trained to do, and your team has to be good at telling the difference before they can fix it.

You'll set technical direction for how we train and evaluate models at scale, stay close enough to the code to debug alongside your team, and work daily with researchers to turn tradeoffs in latency, quality, and cost into infrastructure that actually gets built. We're hiring across a range of scope for this role, depending on experience and the size of problem you're ready to own.

Example projects include..

  • Building the rollout infrastructure that lets researchers run RL experiments at scale without fighting the plumbing.

  • Designing eval pipelines that catch regressions before they ship, and give researchers fast, trustworthy signal on whether a change actually helped.

  • Owning the environments in which models are trained and tested: sandboxed, reproducible, and fast enough that iteration speed isn't the bottleneck.

  • Bringing rigor to how the team measures quality and progress, in places where "did it ship" isn't the same as "did it work?"

  • Partnering with research to translate model-level tradeoffs (latency, quality, cost) into concrete infrastructure decisions.

  • Hiring and growing the team: sourcing, interviewing, and closing exceptional infrastructure engineers, while developing your engineers through coaching, mentorship, and high-leverage project assignments.

You may be a fit if

  • You've led engineering teams building infrastructure that trains, evaluates, or serves ML models in production.

  • You have strong infrastructure and distributed systems fundamentals: you know what reliability and performance look like under real load, not just in a design doc.

  • You genuinely want to stay technical: you're comfortable writing code, reviewing PRs with depth, and using tools like Cursor itself to move fast.

  • You’re comfortable operating in ambiguity: you ask the right questions, make sound decisions with incomplete information, and help the team find a path forward.

  • You have a track record of hiring and developing engineers who are better than you were at their stage.

  • You can talk fluently with researchers about model behavior and with engineers about systems design, and you know when a problem is actually the other team's.

  • Bonus: hands-on experience with RL training infrastructure, eval frameworks, or building and maintaining simulated environments for model training or testing.

Similar positions

Cursor
Software Engineer, ML Research
Cursor · San Francisco
Cursor
Cursor
Cursor
Engineering Manager, Evals
Cursor · San Francisco