hirly
Likely filled

This role has closed. Adexis has taken the posting down.

hirly last saw it live on 30 September 2026. See similar open roles below, or browse all jobs in London.

Adexis

Member of Technical Staff – Post-Training

London

This one has closed. See which open jobs fit you. Free.

Upload your resume and hirly scores it against open jobs in London, then shows your best matches and why.

PDF or DOCX, up to 12MB. No sign-up to see your matches.

hirly's read of this role

Seniority
Lead / management
Stated salary
£150,000 – £225,000 per year
Country
GB
Work mode
On-site / unstated
First seen by hirly
28 Sept 2026

Derived automatically from the posting.

the posting

About Adexis

Adexis is a London-based frontier research lab solving dexterous manipulation. Our mission is to bring machines the power of touch on the path to physical general intelligence, bringing about a future of abundance.

We are building a lean, talent dense organisation, inspired by history's great groups — the Manhattan Project, Apollo, Bell Labs, Skunkworks — where aggregating outlier individuals drove collective output that changed history.

We are backed by leading investors, and researchers from Anthropic, ElevenLabs, Google Deepmind and more.

Member of Technical Staff — Post-Training

As a founding hire, you will own how our foundation model becomes a policy that acts — the post-training recipe that turns a world representation into dexterous, reliable behaviour on a real robot.

What you'll do

Own the post-training stack: how we adapt a pretrained model into a control policy using teleoperated demonstration data.

Design the recipe — demonstration collection strategy, imitation and any reinforcement or self-improvement stages, and the full sequence from foundation model to deployable policy.

Close the sim-to-real and human-to-robot gaps: make a policy learned from human demonstration survive contact with the real world.

Define what good behaviour is and build the evaluation that catches a policy that looks good offline and fails on hardware.

Work shoulder-to-shoulder with the pre-training owner; the boundary between pretraining and post-training is a conversation you'll own jointly.

You might be a fit if you

Have trained manipulation policies and put them on real hardware — you know the specific pain of a policy that demos well and dies on contact.

Know imitation learning deeply and know when reinforcement is worth the instability.

Have a point of view on learning dexterous, contact-rich skills rather than pick-and-place.

Bonus: teleoperation systems, learning from human demonstration, real-world RL.

We seek a high ownership individual up to the task of defining this vision. This is a founding research role and will be compensated accordingly with equity.

Compensation & benefits

Salary: £150–225k

Equity: 0.5-2.5%

Private healthcare with medical history disregarded

Top of the line equipment

Original posting on Adexis's site ↗

Browse similar roles