hirly

Clera

Lead Research Engineer, Data Quality

San Francisco

See how you match this job — and similar ones. Free.

Upload your resume and hirly scores it against this role at Clera first, then against similar open jobs, and shows where you fit and why.

PDF or DOCX, up to 12MB. No sign-up to see your matches.

Get past the screening software and onto a recruiter's desk

hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.

  • Keywords matched to this posting
  • Fit score before you apply
  • Cover letter included

Matched against 2.3M live jobs from 200,000+ employers in 200+ countries.

Tailor my resume for this job →

hirly's read of this role

Seniority
Lead / management
Stated salary
$150,000 – $180,000 per year
Country
US
Work mode
On-site / unstated
First seen by hirly
29 Sept 2026

Derived automatically from the posting. Upload your resume above to see how the role scores against it.

the posting

About the Role

This is a senior technical leadership role on the data quality team at an early-stage AI infrastructure company focused on building and scaling reinforcement learning environments for frontier model training. You will own the strategy and systems that measure and improve training data quality, shaping research culture around what makes agent data genuinely useful rather than just superficially correct.

What You'll Do

Lead the data quality team in building systems that evaluate thousands of tasks across RL environments, synthetic data, benchmarks, and domain-specific workflows.

Define the data quality strategy by building QC systems, enforcing standards, and designing experiments to grade agent outputs.

Develop new methods for validating synthetic data at scale, including failure-mode analysis, task mutation checks, and trajectory auditing.

Partner with research engineers, domain experts, and data vendors to diagnose quality issues and improve data generation workflows.

Turn qualitative research insights into production systems: internal tools, dashboards, validation pipelines, and feedback loops.

Mentor research engineers to maintain a high bar for technical rigor, clarity, and execution speed.

What We're Looking For

5+ years of relevant engineering or research experience, specifically building systems for AI/ML data evaluation or data quality.

Proven track record leading technical teams on ambiguous projects from problem definition through implementation and iteration.

Advanced proficiency in Python, Docker, and Linux environments.

Experience building QC systems, evals, benchmarks, synthetic data pipelines, validation workflows, or model evaluation infrastructure.

Deep intuition for what makes training tasks realistic, learnable, diverse, reliable, and useful for AI agents.

Research-oriented understanding of AI evals and post-training, beyond surface-level agent tooling projects.

Comfort designing metrics, experiments, and QA/QC processes, not just executing them.

Strong written communication skills, with the ability to explain methodology clearly to diverse audiences.

Experience working with subject-matter experts to capture domain judgment and convert it into scalable review or generation systems.

Early-stage startup experience and the ability to work independently in fast-paced environments.

Compensation & Benefits

Salary range: $150,000 to $180,000 USD annually. Visa sponsorship is available.

Location

On-site in San Francisco, CA, USA. The team also has a presence in Singapore.

Original posting on Clera's site ↗

Browse similar roles

Want this one?

Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.

Tailor my resume for this job