hirly

Babbel

Senior Machine Learning Engineer (all genders)

Berlin

See how you match this job — and similar ones. Free.

Upload your resume and hirly scores it against this role at Babbel first, then against similar open jobs, and shows where you fit and why.

PDF or DOCX, up to 12MB. No sign-up to see your matches.

Get past the screening software and onto a recruiter's desk

hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.

  • Keywords matched to this posting
  • Fit score before you apply
  • Cover letter included

Matched against 2.5M live jobs from 200,000+ employers in 200+ countries.

Tailor my resume for this job →

Apply from your AI assistant

Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.

Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.

hirly's read of this role

Role family
Data & ML
Seniority
Senior
Country
DE
Work mode
On-site / unstated
First seen by hirly
21 Sept 2026

Derived automatically from the posting. Upload your resume above to see how the role scores against it.

the posting

About Babbel Labs

Babbel Labs is building the future of language learning. We are an AI-first, independent company within the Babbel group, based in Berlin. Our teams bring AI, research and product together to ship experiences that set a new standard for how people learn.

The role

Babbel's learner-personalisation engine tracks what a learner has and hasn't mastered, and decides what they should practise next. We are actively pushing personalisation further: moving beyond describing a user and prescribing targeted practice, to adapting the learning journey ahead of them.

This is a hands-on senior individual-contributor role on the team behind it. You will own real subsystems end to end and make decisions about how to improve the personalisation engine. That includes everything from research and benchmarking through production, monitoring and the incidents that follow six months later.

How you'll make an impact

Work directly with the Principal Scientist to take designs from spec into production, then keep improving what you've built on your own judgement rather than waiting to be told what's next.

Help shape new features from the beginning, rather than only implementing a spec handed to you.

Take real ownership of core personalisation and mastery-tracking subsystems, operate independently, and make decisions confidently and competently.

Design the evaluation that tells you whether a change is real: a rigorous offline benchmark against a real baseline, and the online experiment that confirms or kills it.

Run what you build. Instrument it, notice when it's silently wrong rather than only when it errors, and fix it before it becomes an incident.

Deliver with coding agents as a matter of course, and verify what they produce before you rely on it.

Your skills and qualifications

Strong, hands-on ML engineering that has shipped real models to production. Recommendation, ranking, scoring, or trust-and-safety systems under real user load are the closest match. Research or competition experience is a plus.

Experience with probabilistic modelling, latent-variable modelling and Bayesian inference, or the equivalent rigour from an adjacent domain.

ML system evaluation: monitoring metrics you define, debugging output that doesn't look right, and rolling out a change to a live scoring or ranking system without breaking it.

Rigorous experimentation practice: benchmarking against a real baseline, running or reading A/B tests correctly, and the judgement to know when an offline improvement won't survive contact with production.

TypeScript/Python as your primary languages, with enough command of our surrounding stack (AWS, Terraform, CI/CD) to ship and own your own service's delivery. This is not an infrastructure role, so depth there is not the bar.

Coding agents are part of your daily workflow, and you check their output before you rely on it. You are neither dismissive of them nor careless with them.

Nice to have

Experience with psychometric models, such as Item Response Theory.

Graph ML experience at real scale, such as embeddings, graph neural networks, or relational modelling.

Public technical work, such as open-source contributions, writing, or competitive ML.

Ways of working

This is a hands-on individual-contributor role on a small team that moves fast and works agent-first, but with a disciplined approach to delivery. It is remote and Berlin-friendly.

Diversity at Babbel

As part of our ongoing journey towards building a diverse, equitable, and inclusive company, we welcome everyone to apply, especially individuals who are underrepresented in tech. We are a learning company, inside and out, and we encourage you to apply even if you do not fit all the technical requirements — all candidates are assessed based on skills, qualifications, and our business needs. Please state your pronouns in your application, and let us know if you'd like to be addressed by a name other than the one appearing on your official documents. If you have a disability or special need, feel welcome to inform us so we can provide proper assistance in the application process.

Original posting on Babbel's site ↗

Browse similar roles

Want this one?

Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.

Tailor my resume for this job