hirly

Metis, Inc.

Research Engineer

San Francisco, California

See how you match this job — and similar ones. Free.

Upload your resume and hirly scores it against this role at Metis, Inc. first, then against similar open jobs, and shows where you fit and why.

PDF or DOCX, up to 12MB. No sign-up to see your matches.

Get past the screening software and onto a recruiter's desk

hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.

  • Keywords matched to this posting
  • Fit score before you apply
  • Cover letter included

Matched against 2.3M live jobs from 200,000+ employers in 200+ countries.

Tailor my resume for this job →

hirly's read of this role

Seniority
Mid level
Stated salary
$200,000 – $1,000,000 per year
Country
US
Work mode
Remote-friendly
First seen by hirly
28 Sept 2026

Derived automatically from the posting. Upload your resume above to see how the role scores against it.

the posting

About the Role

As a Research Engineer at Metis, you’ll work on building the next generation of autonomous post-training systems that leverage our Mantis platform. You’ll operate at the intersection of cutting-edge ML research and scalable engineering, designing, implementing, and deploying algorithms that improve how AI agents learn from feedback, synthetic data, and real-world interactions.

You’ll move seamlessly between papers and production, leading large-scale experiments, creating optimized training pipelines, and helping shape the future of post-training autonomy. You’ll have significant ownership, high compute budgets, and the mandate to push the state of the art in applied reinforcement and preference optimization.

What You'll Do

Research and help build an autonomous post-training agent leveraging the Mantis platform

Design and execute large-scale experiments on synthetic data generation and algorithmic architecture

Develop and refine methods for reinforcement learning, reward modeling, and human feedback integration

Collaborate cross-functionally with Core and Platform Engineering to deploy and evaluate models in production settings

Publish or contribute to leading-edge research in the post-training domain

Use tooling and compute efficiently to iterate on experimental pipelines and accelerate research velocity

Requirements

Deep experience in machine learning, preferably reinforcement learning, post-training, or alignment research

Demonstrated research contributions; ideally published papers (ICML, NeurIPS) or public implementations

Strong proficiency in Python and ML frameworks (PyTorch, JAX, or TensorFlow)

Comfort with distributed training, high-throughput data pipelines, and large-scale experiment management

Ability to reason independently, formulate hypotheses, and run experiments from idea → insight → product impact

Compensation & Benefits

Base: $200,000–$1,000,000

Significant Equity

Full medical, dental, and vision

Wellness & L&D stipend

Equinox membership

Breakfast, lunch, and dinner provided (Unlimited Doordash)

$25,000 housing stipend

About Metis

Metis helps enterprises and labs build the most reliable AI agents by leveraging post-training. Our platform enables the creation, improvement, and deployment of the most capable frontier agents designed for rigorous, real-world workflows.

Momentum

0 → six-figure monthly revenue in the last six weeks

Working with several Fortune 500 enterprises & frontier AI labs

Growing 150%+ MoM

Backed by

Y Combinator, CRV, and executives from OpenAI, Google, Mercor, NVIDIA, and others.

Original posting on Metis, Inc.'s site ↗

Browse similar roles

Want this one?

Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.

Tailor my resume for this job