Periodic Labs
Research Scientist/Research Engineer, Midtraining
Menlo Park, CA
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.3M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →hirly's read of this role
- Role family
- Data & ML
- Seniority
- Mid level
- Stated salary
- $250,000 – $350,000 per year
- Country
- US
- Work mode
- On-site / unstated
- First seen by hirly
- 1 Oct 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
We're an AI and physical sciences company building state-of-the-art models to accelerate breakthroughs across materials, energy, and beyond. Backed by world-class investors and growing rapidly, we operate at the pace the frontier requires. Our team brings deep expertise, genuine ownership, and a drive to push the boundaries of what's scientifically possible.
About the Role
We're training frontier models to develop deep scientific knowledge and reasoning for scientific discovery. As a Midtraining Research Engineer, you'll take base models and improve their scientific reasoning: curating and generating data, building evals, and running large-scale training experiments. Your work will also lay the groundwork for our pre-training efforts down the line.
What You'll Do
Identify, process, and curate novel sources of scientific data for large-scale model training.
Generate high-quality synthetic data to fill gaps in scientific knowledge and reasoning.
Build evaluations that correlate with downstream scientific task performance, working closely with RL researchers, physicists, and chemists.
Develop and apply techniques such as self-distillation and on-policy distillation to improve model capability.
Design and run large-scale training experiments, partnering with supercompute engineers to scale efficiently across thousands of GPUs.
Build tools for yourself and the team to investigate how data choices shape model intelligence.
You Will Thrive in This Role If You Have
Experience training LLMs on curated mixes of trillions of tokens.
Experience on a dedicated evals team supporting a large production training run.
Hands-on use of self-distillation, on-policy distillation, or similar methods in a real training pipeline.
Experience with scaling laws and compute-optimal hyperparameters.
Comfort working across data, evals, and training infrastructure.
Especially Strong Candidates May Also Have
Experience optimizing throughput and reliability for large-scale distributed training runs.
A background in AI for science or training on specialized domain data (e.g., protein, materials, or other scientific datasets).
Experience creating evals or synthetic data for non verifiable tasks and tracking performance over live runs.
Mechanics
Minimum education: Bachelor's degree or similar experience
Location: Menlo Park, CA
Compensation: $250,000–$350,000 + equity
Visa sponsorship: Yes, we sponsor visas and will do everything we can to assist in this process.
Similar jobs
- Basic Life Research Scientist (1-Year Fixed-Term, 80% FTE)Stanford University · Stanford, CA, United StatesFirst seen today
- Research Scientist- Lott LabMayo US · Scottsdale, AZ, United StatesFirst seen today
- Research Scientist - FacultyTamus · Prairie View, TXFirst seen today
- Research Scientist - FacultyTamus · Prairie View, TXFirst seen today
- Research ScientistMeharrymedicalcollege · Main CampusFirst seen today
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job