This role has closed. Adexis has taken the posting down.
hirly last saw it live on 30 September 2026. See similar open roles below, or browse all jobs in London.
Adexis
Member of Technical Staff – Post-Training
London
Similar open jobs
- Sr. Member Technical Staff - ESD and Latch-Up - HBMMicron · Folsom, CAFirst seen 6d ago
- Member Technical Staff - Applied AI Engineer (US Timing) Composio · BangaloreFirst seen 25d ago
- Member Technical StaffPirros · Los Angeles OfficeFirst seen 31d ago
- Member TechnicalBroadridge · Bengaluru-EPIP Industrial AreaFirst seen 8d ago
- Senior Member TechnicalBroadridge · Hyderabad-Hi-Tec CityFirst seen 10d ago
hirly's read of this role
- Seniority
- Lead / management
- Stated salary
- £150,000 – £225,000 per year
- Country
- GB
- Work mode
- On-site / unstated
- First seen by hirly
- 28 Sept 2026
Derived automatically from the posting.
the posting
About Adexis
Adexis is a London-based frontier research lab solving dexterous manipulation. Our mission is to bring machines the power of touch on the path to physical general intelligence, bringing about a future of abundance.
We are building a lean, talent dense organisation, inspired by history's great groups — the Manhattan Project, Apollo, Bell Labs, Skunkworks — where aggregating outlier individuals drove collective output that changed history.
We are backed by leading investors, and researchers from Anthropic, ElevenLabs, Google Deepmind and more.
Member of Technical Staff — Post-Training
As a founding hire, you will own how our foundation model becomes a policy that acts — the post-training recipe that turns a world representation into dexterous, reliable behaviour on a real robot.
What you'll do
Own the post-training stack: how we adapt a pretrained model into a control policy using teleoperated demonstration data.
Design the recipe — demonstration collection strategy, imitation and any reinforcement or self-improvement stages, and the full sequence from foundation model to deployable policy.
Close the sim-to-real and human-to-robot gaps: make a policy learned from human demonstration survive contact with the real world.
Define what good behaviour is and build the evaluation that catches a policy that looks good offline and fails on hardware.
Work shoulder-to-shoulder with the pre-training owner; the boundary between pretraining and post-training is a conversation you'll own jointly.
You might be a fit if you
Have trained manipulation policies and put them on real hardware — you know the specific pain of a policy that demos well and dies on contact.
Know imitation learning deeply and know when reinforcement is worth the instability.
Have a point of view on learning dexterous, contact-rich skills rather than pick-and-place.
Bonus: teleoperation systems, learning from human demonstration, real-world RL.
We seek a high ownership individual up to the task of defining this vision. This is a founding research role and will be compensated accordingly with equity.
Compensation & benefits
Salary: £150–225k
Equity: 0.5-2.5%
Private healthcare with medical history disregarded
Top of the line equipment