hirly

EggAI

Platform Tech Lead - Banking

Paris, Paris, France · München, Bavaria, Germany · Rome, Italy · London, England, United Kingdom

See how you match this job — and similar ones. Free.

Upload your resume and hirly scores it against this role at EggAI first, then against similar open jobs, and shows where you fit and why.

PDF or DOCX, up to 12MB. No sign-up to see your matches.

Get past the screening software and onto a recruiter's desk

hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.

  • Keywords matched to this posting
  • Fit score before you apply
  • Cover letter included

Matched against 2.3M live jobs from 200,000+ employers in 200+ countries.

Tailor my resume for this job →

hirly's read of this role

Seniority
Lead / management
Countries
FR, DE, IT, GB
Work mode
Remote-friendly
First seen by hirly
29 Sept 2026

Derived automatically from the posting. Upload your resume above to see how the role scores against it.

the posting

EggAI · Banking transformation programme · Remote — Regular travel to London, UK

About the Programme

EggAI is the transformation partner on a core banking system replacement for a client to lay the foundations of operating Agentic Generative AI in production.

EggAI's role is to specify, set the standard and programme-manage. You will define what operational readiness means and prove it at each gate, working with the client's own teams and its managed service providers, as well as EggAI's platform engineers.

The platform is a mix of classical banking infrastructure and agentic AI workloads, so the reliability problem runs across both deterministic and probabilistic systems.

The Role

Own the technical definition of operational readiness for go-live: what "ready to run in production" means, how the bank proves it at each regulatory gate, and what needs to change for that proof to hold.

This is a lead/principal position reporting to the Programme Plan and Execution Lead. It is remote, with regular London presence required for gate weeks and workshops.

Responsibilities

Reliability targets — the SLO and error-budget framework, adopted alongside the client's existing structure, and the behavioural quality measures that do the same job for agentic workloads

Observability and detection — emittances of Tier-1 services, trajectory logging for agentic workloads, an automated path from a firing alert to the affected services and their owners, and closing the logging gaps in what runs today

The incident and change path — runbook standards and its coverage, incident classification and intervention mechanisms extended to agentic failure modes, and rollback and change-failure rates brought under measurement

Capacity and performance — the headroom position through the parallel-run peak at cutover, the chaos and game-day programme for payment rails and the core ledger, and capacity modelling extended to inference workloads

Gate evidence — technical content of the Operational Acceptance Report and Operational Readiness Certificate, the reliability section of each gate pack, and reliability NFRs defended as ADRs through the client's Architecture Design Authority

Expected experience

10+ years in production engineering — SRE, platform reliability or infrastructure ownership.

Has personally introduced SLOs and error budgets into an organisation that did not have them, and can build on the experienced success factors

Observability in depth — OpenTelemetry, distributed tracing, metrics and log architecture, and the retention and cost trade-offs that come with them

Incident handling and handling processes , with post-incident reviews that led to real changes in the system

Regulated financial services , or another environment where an external body audits operational evidence.

Has worked through a managed service provider and can set standards they do not personally execute

Comfortable operating in a complex, multi-stakeholder programme where scope and ownership are still being established across parallel workstreams

Valuable, not required

LLM or agentic system operations: eval harnesses, drift detection, trajectory analysis, prompt and model versioning treated as change types

Core banking or payments exposure, particularly scheme gateway integration

What This Role Is Not

Not a platform architect. A separate role owns agentic platform design and build

Not an on-call engineer. You will not carry a client pager; out-of-hours operations are run by a third party

Not a service delivery manager. This is engineering judgement applied to readiness, not ITSM process administration

About EggAI

Agentic workforces are inevitable. Making them work is our mission.

EggAI is building the engineering methods, operating models, and reusable technology to deploy agentic workforces with the quality and control required for sustained business impact.

Working closely with enterprise clients, we design, implement and operate agentic systems that evolve from augmenting individual tasks to autonomously executing end-to-end processes alongside people—resilient, controlled, and scalable. We begin with the client problem, select the technical approach that best serves it, and remain accountable through deployment, adoption, and operations. Each deployment strengthens the next by turning production learning into reusable capability.

We are an experienced, international team working across Europe. We set high standards, take ownership of outcomes, and value clear thinking, candid feedback, and the ability to turn ideas into tangible results.

Original posting on EggAI's site ↗

Browse similar roles

Want this one?

Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.

Tailor my resume for this job