Clera
Research Engineer, Synthetic Data
Singapore
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.3M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →hirly's read of this role
- Seniority
- Mid level
- Stated salary
- $150,000 – $250,000 per year
- Country
- SG
- Work mode
- On-site / unstated
- First seen by hirly
- 28 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
About the Role
This is a Research Engineer role focused on synthetic data, sitting within a roughly 15-person engineering team of Olympiad medalists and published researchers. You will build the pipelines that turn domain-specific workflows into scalable, high-quality training tasks for AI agents, directly shaping what models learn and how well they perform.
What You'll Do
Build end-to-end synthetic data pipelines that transform domain-specific workflows into realistic, structured, and challenging training tasks.
Collaborate with subject-matter experts to create synthetic tasks for AI agents across professional and technical domains.
Design task generation methods that produce diverse, realistic, and learnable outputs at scale.
Build tooling to mutate, validate, and continuously improve synthetic tasks.
Analyze model and agent performance on synthetic tasks to identify what they teach and where they break down.
Develop metrics to quantify synthetic task diversity, realism, learnability, and overall quality.
What We're Looking For
2 to 4 years of experience in software engineering, machine learning engineering, or AI research, with hands-on work building data pipelines, ML infrastructure, or synthetic data systems.
Proficiency in Python and experience developing in Linux environments using containerization tools such as Docker.
Demonstrated experience applying synthetic data research methods to build end-to-end data generation pipelines for AI/ML applications.
Strong understanding of synthetic data quality criteria and evaluation metrics, including diversity, realism, and learnability, as well as their inherent limitations.
Experience designing, implementing, or maintaining evaluation frameworks, benchmarks, or testing environments for AI agents or large language models.
Track record of independently owning and delivering technical projects end-to-end with minimal predefined requirements.
Experience building automated systems to generate, validate, mutate, or process structured datasets at scale.
Sharp eye for edge cases, inconsistencies, and quality issues in synthetic or algorithmically generated data.
Familiarity with reinforcement learning training paradigms, agentic AI workflows, or LLM post-training pipelines is a plus.
Comfortable operating in unstructured, early-stage environments and reasoning from first principles.
Strong communication skills for asynchronous, cross-timezone collaboration.
Compensation & Benefits
Salary range: $150,000 to $250,000 USD annually. Visa sponsorship is available.
Location
On-site in Singapore .
Similar jobs
- Applied Science - Research Engineer (Robotics)Mistral · SingaporeFirst seen yesterdayremote
- Research Engineer – AI AgentsJumptrading · SingaporeFirst seen 8d ago
- Research Engineer, GeneralistExa · SingaporeFirst seen 19d ago
- Research Engineer (Psychology / Neuropsychology / Cognitive Neuroscience)Ntu · NTU Main Campus, SingaporeFirst seen today
- Research Engineer, Robotics DataHud · San FranciscoFirst seen today
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job