hirly

Figureai

Helix AI Engineer, Generative AI

San Jose, CA

See how you match this job — and similar ones. Free.

Upload your resume and hirly scores it against this role at Figureai first, then against similar open jobs, and shows where you fit and why.

PDF or DOCX, up to 12MB. No sign-up to see your matches.

Get past the screening software and onto a recruiter's desk

hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.

  • Keywords matched to this posting
  • Fit score before you apply
  • Cover letter included

Matched against 2.6M live jobs from 190,000+ employers in 200+ countries.

Tailor my resume for this job →

Apply from your AI assistant

Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.

Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.

hirly's read of this role

Seniority
Mid level
Stated salary
$200,000 – $400,000 per year
Country
US
Work mode
Remote-friendly
First seen by hirly
4 Sept 2026

Derived automatically from the posting. Upload your resume above to see how the role scores against it.

the posting

Figure is an AI robotics company developing autonomous general-purpose humanoid robots. Our goal is to build embodied AI systems that can perceive, reason, and act in the real world. Figure is headquartered in San Jose, CA, and this role requires 5 days/week in-office collaboration.

Our Helix team is responsible for developing the core AI systems that power humanoid autonomy. We are looking for a Helix AI Engineer, Generative AI to build and scale generative models that enable robots to understand, simulate, and interact with the physical world. This role focuses on training and deploying diffusion and generative models across vision, video, and multimodal domains, with applications spanning perception, data generation, and model-based reasoning.

Responsibilities

Design, train, and deploy large-scale generative models, with a focus on diffusion-based approaches for vision, video, and multimodal data

Develop models that improve robot perception, world modeling, and prediction from raw sensory inputs

Build generative systems for synthetic data creation, augmentation, and dataset scaling for robot learning

Explore and implement state-of-the-art techniques in diffusion, generative modeling, and multimodal foundation models

Optimize training pipelines for large-scale generative models across distributed systems

Work closely with data, training infrastructure, and agent teams to integrate generative models into the full autonomy stack

Evaluate model quality, robustness, and generalization across real-world scenarios

Contribute to the design of scalable experimentation frameworks for generative model development

Requirements

Experience training and deploying generative models (diffusion, autoregressive, or related approaches) at scale

Strong understanding of modern deep learning techniques for vision and/or multimodal systems

Proficiency in Python and deep learning frameworks such as PyTorch

Experience working with large-scale datasets and distributed training systems

Strong experimental rigor and ability to iterate quickly on model performance

Solid software engineering skills and ability to build reliable, maintainable systems

Ability to operate independently and own ambiguous, high-impact technical problems

Bonus Qualifications

Experience with diffusion models for image or video generation

Experience with multimodal foundation models (vision-language or vision-language-action)

Background in synthetic data generation or simulation for robotics or embodied AI

Experience optimizing large-scale training (multi-node, GPU clusters, etc.)

Familiarity with 3D, video prediction, or world models

Prior work in robotics, embodied AI, or real-world ML systems

Publication record in machine learning, computer vision, or generative modeling

The US base salary range for this full-time position is between $200,000 - $400,000

The pay offered for this position may vary based on several individual factors, including job-related knowledge, skills, and experience. The total compensation package may also include additional components/benefits depending on the specific role. This information will be shared if an employment offer is extended.

Original posting on Figureai's site ↗

Listed on hirly, a job board. hirly is not the employer: Figureai is hiring for this role.

Want this one?

Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.

Tailor my resume for this job