Periodic Labs
Research Scientist, Scaling RL
Menlo Park, CA
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.3M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →hirly's read of this role
- Role family
- Data & ML
- Seniority
- Mid level
- Stated salary
- $250,000 – $350,000 per year
- Country
- US
- Work mode
- On-site / unstated
- First seen by hirly
- 1 Oct 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
About Periodic Labs
We're an AI and physical sciences company building state-of-the-art models to accelerate breakthroughs across materials, energy, and beyond. Backed by world-class investors and growing rapidly, we operate at the pace the frontier requires. Our team brings deep expertise, genuine ownership, and a drive to push the boundaries of what's scientifically possible.
About the Role
We're training frontier models to develop deep scientific knowledge and reasoning for scientific tasks. You’ll study how RL scales with training compute, develop better algorithms, and take ideas from controlled experiments to our largest runs like Periodic Neon .
What You'll Do
Design experiments to understand how RL performance scales with compute, model size, data, and reward quality, building on work such as ScaleRL
Develop better RL algorithms, spanning policy optimization, advantage estimation, exploration, and credit assignment for long-horizon RL tasks
Build adaptive sampling and curriculum methods that adjust task difficulty, problem selection, and the number of rollouts as models improve
Study bias and stability during RL training, including importance-sampling corrections and methods to tackle policy staleness and training–inference mismatch, as discussed here .
Improve compute efficiency across training and inference through experiments with hyperparameters, such as length penalties, rollout counts, batch sizes, and update schedules.
You Will Thrive in This Role If You Have
Hands-on experience training LLMs with reinforcement learning
Strong attention to detail and rigorous approach to answer questions scientifically.
Coming up with small-scale RL setups that transfers to large-scale training runs.
Comfort working across a complex training stack to implement, debug, and test new research ideas.
Mechanics
- Minimum experience: 5+ years
- Minimum education: Bachelor’s degree or similar experience
Location: Menlo Park, CA
Compensation: $250,000-$350,000 base + equity
Visa sponsorship: Yes, we sponsor visas and will do everything we can to assist in this process.
Similar jobs
- ML Research Scientist - Computational Biologist/BioinformaticsMerge Labs · San Francisco Bay AreaFirst seen today
- Research Scientist - Computational NeuroscienceAstera · Emeryville HQFirst seen todayremote
- Research Scientist, Applied AI ResearchEdmentum · United StatesFirst seen todayremote
- AI Research Scientist IAxon · Washington, United StatesFirst seen todayremote
- High Temp Composite Lab Research ScientistUniversity of Tulsa Student · Tulsa, OK, United StatesFirst seen today
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job