hirly

Cursor

Full Stack Engineer, ML Research Tools

San Francisco

See how you match this job — and similar ones. Free.

Upload your resume and hirly scores it against this role at Cursor first, then against similar open jobs, and shows where you fit and why.

PDF or DOCX, up to 12MB. No sign-up to see your matches.

Get past the screening software and onto a recruiter's desk

hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.

  • Keywords matched to this posting
  • Fit score before you apply
  • Cover letter included

Matched against 2.4M live jobs from 200,000+ employers in 200+ countries.

Tailor my resume for this job →

Apply from your AI assistant

Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.

Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.

hirly's read of this role

Role family
Engineering
Seniority
Mid level
Country
US
Work mode
On-site / unstated
First seen by hirly
16 Sept 2026

Derived automatically from the posting. Upload your resume above to see how the role scores against it.

the posting

Our mission is to automate coding. The first step in our journey is to build the best tool for professional programmers, using a combination of inventive research, design, and engineering. Our organization is very flat, and our team is small and talent dense. We particularly like people who are truth-seeking, passionate, and creative. We enjoy spirited debate, crazy ideas, and shipping code.

About the role

As a Software Engineer on the RL Data team, you’ll design and build the tools that researchers and external contributors use to create, review, submit, and monitor the environments and tasks behind Cursor’s reinforcement-learning runs. This is a full-stack product-engineering role embedded in a research team.

You’ll own the review and acceptance experience end to end: from rollout and transcript inspection, task-quality signals grader and reward-hacking analysis, to the workflows that move a submission into training. From there, you’ll build authoring interfaces that let researchers, vendors, and domain experts create and improve environments and tasks quickly and confidently.

Your work will significantly shorten the loop from a task idea or data sources, to candidate task, to trusted training data.

What you’ll work on

Create fast, trustworthy workflows for vendors and research team to interact effectively with each other — vendor task creation and iteration, vendor submissions, and task acceptance into training.

Build review tools for inspecting and comparing rollouts, transcripts, grader outputs, and other signals of task quality.

Develop environment-health, failure-search, versioning, and catalog experiences that make training data easy to understand, manage, and extend.

Establish a shared component kit, then use it to build self-serve interfaces for creating and improving tasks with quality checks inline.

You may be a fit if

You’ve shipped full-stack products and owned systems from user interface through storage or services, using technologies such as TypeScript and React alongside Node, Python, or Go.

You’ve built dense, data-facing tools such as transcript viewers, diffing systems, review queues, observability products, or operational dashboards—and you have strong opinions about how structured data should be rendered.

You’ve built or maintained a design system or component library and can establish durable product and engineering conventions for a fast-moving team.

You’ve designed review, QA, moderation, fraud, or acceptance workflows where users had an incentive to get past the checks, and you know how to keep those systems honest.

You care about data quality, and are willing to inspect raw data. Experience with evaluations, graders, reinforcement learning, or data-quality systems is helpful but not required.

You move quickly under ambiguity, collaborate closely with researchers and domain experts, and take open-ended problems from rough need to reliable product.

Applying

If there appears to be a fit, we’ll schedule two or three short technical interviews focused on frontend craft for dense data and system design for a review-and-acceptance workflow. After that, we’ll invite you onsite to work on a small project using real rollouts, discuss ideas, and meet the team.

#LI-DNI

Original posting on Cursor's site ↗

Browse similar roles

Want this one?

Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.

Tailor my resume for this job