← Back to results

Software Engineer, Research Tools

Join a small, passionate team to build tools for researchers in reinforcement learning at a startup focused on automating coding.

Location
San Francisco, United States
Compensation
Not disclosed
Level
mid
Type
full time

Posted by employer 18 hours ago

First seen on Joblaze 46 minutes ago

Last verified on the company career page 46 minutes ago

Skills & Technologies

What you'll build

  • Create workflows for vendor task creation and iteration
  • Build review tools for inspecting rollouts and task quality
  • Develop environment-health and versioning experiences
  • Establish a shared component kit for self-serve interfaces
  • Build authoring interfaces for researchers and domain experts

Must have

  • Shipped full-stack products
  • Experience with TypeScript and React
  • Experience with Node, Python, or Go

Nice to have

  • Built data-facing tools like transcript viewers or dashboards
  • Experience with design systems or component libraries
  • Experience with evaluations or data-quality systems

Not disclosed in this posting: compensation, years of experience, work arrangement, visa sponsorship.

Joblaze summary

In the role of Software Engineer on the RL Data team, the individual will design and develop tools that facilitate the creation, review, and monitoring of reinforcement-learning tasks. Proficiency in full-stack development using technologies like TypeScript, React, Node, Python, or Go is essential, along with experience in building data-centric tools and workflows. This position is well-suited for someone with a strong background in product ownership and a keen interest in data quality, particularly in a collaborative research environment. The team operates in a flat structure, encouraging open dialogue and innovative ideas.

Joblaze insights

Quick facts

What's the tech stack?
Joblaze extracted these technologies from the posting: Go, Node, Python, React, TypeScript.
What seniority level is this role?
Cursor targets mid-level candidates for this position.
Is this full-time or contract?
Full-time for this Software Engineer, Research Tools role at Cursor.

From the original posting

Our mission is to automate coding. The first step in our journey is to build the best tool for professional programmers, using a combination of inventive research, design, and engineering. Our organization is very flat, and our team is small and talent dense. We particularly like people who are truth-seeking, passionate, and creative. We enjoy spirited debate, crazy ideas, and shipping code.

About the role

As a Software Engineer on the RL Data team, you’ll design and build the tools that researchers and external contributors use to create, review, submit, and monitor the environments and tasks behind Cursor’s reinforcement-learning runs. This is a full-stack product-engineering role embedded in a research team.

You’ll own the review and acceptance experience end to end: from rollout and transcript inspection, task-quality signals grader and reward-hacking analysis, to the workflows that move a submission into training. From there, you’ll build authoring interfaces that let researchers, vendors, and domain experts create and improve environments and tasks quickly and confidently.

Your work will significantly shorten the loop from a task idea or data sources, to candidate task, to trusted training data.

What you’ll work on

  • Create fast, trustworthy workflows for vendors and research team to interact effectively with each other — vendor task creation and iteration, vendor submissions, and task acceptance into training.

  • Build review tools for inspecting and comparing rollouts, transcripts, grader outputs, and other signals of task quality.

  • Develop environment-health, failure-search, versioning, and catalog experiences that make training data easy to understand, manage, and extend.

  • Establish a shared component kit, then use it to build self-serve interfaces for creating and improving tasks with quality checks inline.

You may be a fit if

  • You’ve shipped full-stack products and owned systems from user interface through storage or services, using technologies such as TypeScript and React alongside Node, Python, or Go.

  • You’ve built dense, data-facing tools such as transcript viewers, diffing systems, review queues, observability products, or operational dashboards—and you have strong opinions about how structured data should be rendered.

  • You’ve built or maintained a design system or component library and can establish durable product and engineering conventions for a fast-moving team.

  • You’ve designed review, QA, moderation, fraud, or acceptance workflows where users had an incentive to get past the checks, and you know how to keep those systems honest.

  • You care about data quality, and are willing to inspect raw data. Experience with evaluations, graders, reinforcement learning, or data-quality systems is helpful but not required.

  • You move quickly under ambiguity, collaborate closely with researchers and domain experts, and take open-ended problems from rough need to reliable product.

Applying

If there appears to be a fit, we’ll schedule two or three short technical interviews focused on frontend craft for dense data and system design for a review-and-acceptance workflow. After that, we’ll invite you onsite to work on a small project using real rollouts, discuss ideas, and meet the team.

#LI-DNI

Similar positions

Cursor
Software Engineer, RL Data
Cursor · San Francisco
Cursor
Software Engineer, ML Research
Cursor · San Francisco
Cursor
Engineering Manager, ML
Cursor · San Francisco