Generalist
Research Scientist: Pretraining
San Francisco Bay Area (San Mateo) or Boston (Somerville)
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.5M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Role family
- Data & ML
- Seniority
- Mid level
- Stated salary
- $240,000 – $350,000 per year
- Country
- US
- Work mode
- On-site / unstated
- First seen by hirly
- 28 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
About Generalist
At Generalist, we are on a mission to build general intelligence for the physical world and make it useful to everyone. We believe the industries and homes of the future will depend on humans and machines working together in new ways. Robots can help us build more and get more done.
We build embodied foundation models, starting with a focus on dexterity. This requires advancing the frontiers of data, models, and hardware, to enable robots to intelligently interact with the physical world.
The company embraces both large-scale AI and robotics as core to its DNA. Our team of researchers, roboticists, and company builders come from OpenAI, Boston Dynamics, Google DeepMind, and other frontier labs—with a track record of shipping AI breakthroughs. Before Generalist, we pioneered large embodied multimodal models and vision-language-action models (PaLM-E, RT-2 , Gemini Robotics ), launched and scaled ChatGPT and GPT-4 to hundreds of millions of users, engineered the foundations of autonomous driving, built next-generation robots ( Atlas , Spot , Stretch ) and pushed the limits of what they can do (from parkour to manipulation , and testing robustness ).
We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.
About the Role
You will build the base intelligence layer for robotics. We train large-scale robot foundation models from massive multimodal datasets spanning video, proprioception, action traces, language, and more. You will design and run the core large-scale training efforts that give our models fundamentally new general capabilities across embodiments, tasks, and environments. You will “live and breathe” all forms of robot data.
If you have worked on pretraining large-scale ML systems or generative foundation models e.g. multimodal, language, audio, etc. (not only robotics), then this may be the role for you.
You’ll be responsible for:
Designing and executing large-scale pretraining runs for robot foundation models (transformer- and diffusion-based architectures)
Defining model architectures, objectives, and training curricula across multimodal robotic data (vision, action, state, language)
Developing scalable data mixtures and sampling strategies across petabyte-scale datasets
Guiding data collection operations towards new directions, as well as sourcing new datasets
Running ablations to understand scaling laws, data quality effects, and architecture tradeoffs
Collaborating closely with ML Infra and Systems to push cluster utilization, throughput, and reliability
Turning raw robotic interaction data into generalizable model capabilities
You might thrive in this role if you:
Have worked on large-scale ML systems with senior/staff (L6+) experience.
Have deep experience training large transformer or diffusion models at scale (for generative models e.g. including language models, audio models, or video models)
Have led or significantly contributed to multi-node, multi-GPU distributed training efforts
Have worked on scaling laws, optimization dynamics, and large-model failure modes
Have strong PyTorch fundamentals and comfort debugging at every layer of the stack
Care about both empirical rigor and raw iteration speed
Are excited about building general-purpose robot intelligence from first principles
Similar jobs
- Research Scientist - Bioanalytical LCMS Research and DevelopmentThermofisher · Richmond, Virginia, USAFirst seen today
- Assistant Research Scientist-Calcium Carbonate Digital TwinUmd · UMCES Horn Point LaboratoryFirst seen today
- Research Scientist, Fundamental Generative AI - New College Grad 2026Nvidia · US, CA, Santa ClaraFirst seen today
- Research Scientist level 3/4Ngc · United States-West Virginia-Rocket CenterFirst seen today
- Research ScientistJobgether · USFirst seen todayremote
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job