This role has closed. FactFinder has taken the posting down.
hirly last saw it live on 4 September 2026. See similar open roles below, or browse all jobs in Berlin.
FactFinder
Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)
Berlin
Similar open jobs
- Senior Site Reliability Engineer (f/m/d)Forto Logistics SE & Co. KG · Berlin, GermanyFirst seen 4d ago
- Staff Site Reliability Engineer (f/m/d)IONOS SE · Berlin, GermanyFirst seen 2d ago
- Senior Site Reliability Engineer, Vehicle SWWayve · Leonberg, GermanyFirst seen 2d agoremote
- Kiwigrid GmbH: (Senior) Site Reliability Engineer(m/f/d)Kiwigrid GmbH · Dresden, Sachsen, GermanyFirst seen 3d ago
- Kiwigrid GmbH: (Senior) Site Reliability Engineer (m/w/d)Kiwigrid GmbH · Dresden, Sachsen, GermanyFirst seen 3d ago
- 586963 Senior X-Segment Physics & Reliability Engineer (all genders)Philips Deutschland GmbH · Hamburg, GermanyFirst seen 4d ago
- Senior Reliability Engineer, Energy Products (m/f/d)Tesla Germany GmbH · Freiburg im Breisgau, Baden-Württemberg, GermanyFirst seen 4d ago
- Site Reliability Engineer - Kubernetes / DevOps (m/w/d)Workwise GmbH · Karlsruhe, Baden, Baden-Württemberg, GermanyFirst seen yesterday
- Site Reliability Engineer (w/m/d)IONOS DE · Hinterm Hauptbahnhof 3-5, 76137 KarlsruheFirst seen yesterday
- Site Reliability Engineer (w/m/d)IONOS DE · Revaler Straße 28-31, 10245 BerlinFirst seen yesterday
- Staff Site Reliability Engineer (f/m/d) IONOS EN · Revaler Straße 28-31, 10245 BerlinFirst seen yesterdayremote
- Site Reliability Engineer (f/m/d) IONOS EN · Revaler Straße 28-31, 10245 BerlinFirst seen yesterdayremote
- System Engineer/Site Reliability Engineer (m/w/d) | OMCOPAAtruvia · Aschheim, Deutschland; Karlsruhe, Deutschland; Münster, DeutschlandFirst seen 2d ago
- Site Reliability Engineer (m/f/d)Deepslate · GermanyFirst seen 2d agoremote
- Site Reliability Engineer (m/f/d)Codesphere · Remote (Germany)First seen 2d agoremote
hirly's read of this role
- Role family
- Engineering
- Seniority
- Senior
- Country
- DE
- Work mode
- On-site / unstated
- First seen by hirly
- 4 Sept 2026
Derived automatically from the posting.
the posting
Introduction
At a glance
Location &workmodel:Berlin, hybrid
Tech stack:Kubernetes on our own servers, Harvester (KubeVirt), Argo CD/Flux, Prometheus/Grafana, Longhorn/Ceph
Team:A growing SRE team – you report to our CTPO for now and to the Team Lead SREwe'rehiring next; two system administrators in Pforzheim run the physical hardware
Process:Intro call · take-home task (~2h) · 90-min tech interview with our developers · leadership conversation · meet the team
Languages:Fluent Englishrequired; German is a plus, nota must
Why this role is special
Most SRE jobs today mean clicking around a managed cloud console. This one doesn't. We run our own hardware in Frankfurt and are building a modern private cloud platform on Kubernetes and Harvester – on-prem by default, with elastic burst into the public cloud and the option to go cloud-only later. You won't inherit a finished SRE practice: you'll help define it, side by side with our Berlin development teams – and you won't do it alone, a Team Lead SRE hire is coming next.
SRE here is an enabling discipline: you build what our developers need to ship reliably, while two system administrators in Pforzheim run the physical hardware. And the impact is direct – our product discovery technology powers more than 2,000 European online shops (Intersport, SPAR, Douglas and more), handling billions of shopper queries a year. When product discovery is slow or down, our customers lose revenue in real time.
Your first 90 days
You get to know both products, join the on-call rotation with a buddy, and own your first reliability topic – SLOs for one product, alerting that actually helps at 3 a.m., or automating away a piece of toil. By day 90 you've shipped visible improvements and know where you want to take the platform next.
Your mission
Define and own SLOs, SLIs and error budgets; drive data-informed reliability decisions
Lead incident response end-to-end: fast detection, clear communication, blameless postmortems – and reduce whole classes of incidents structurally, not case by case
Eliminatetoil through automation andGitOps; evolve our observability (metrics, logs, traces, alerting, runbooks) across two different stacks
Help build our custom Kubernetes operator (CRDs) that makes stateful search clusters declarative, self-healing and safely upgradable – and roll out the auto-scaling (HPA/VPA, KEDA, clusterautoscaler) today's architecture makes hard
Plan capacity,performanceand cost across on-premises and cloud – including the large-catalogue and peak-season loads our merchants care about – and use AI tools wherever they measurably speed up diagnosis and operations
Your profile
Must-haves:
Kubernetes in production – built, not just used:you'veset up andmaintainedclusters on your own servers (e.g.kubeadm, RKE2, k3s) and know cluster lifecycle and upgrades – managed-only experienceisn'tenough for this role
Lived SRE practice: SLOs, error budgets, incident management,on-call
Hands-on experience withGitOpsor comparable infrastructure/deployment automation– experience with Argo CD or Flux is a strong plus
Solid observability skills– metrics, logs, traces, alerting that people trust
A strong automation instinct–you'drather fix a problem's cause than repeat its workaround
A collaborative, enabling mindset– you see SRE as a service to our developers: you ask what they need, discuss trade-offs openly, anddon'tfall in love with your own solution
Nice-to-haves (genuinely optional – we'll teach you the rest):
Harvester,KubeVirt, vSphere/ESXi, OpenStack or similar virtualization/HCI platforms
Container storage (Longhorn, Ceph) and datacenter networking (load balancing, ingress, VLAN)
Auto-scaling (HPA, VPA, KEDA, clusterautoscaler) and capacity/cost planning
Experience building Kubernetes operators/CRDs
- German language skills
- Certifications (CKA, CKS) are welcome but no substitute for hands-on experience – in the tech interview we'll ask about what you've actually built and operated.
You don't tick every box – or your title was never “SRE”? Apply anyway. If you've owned production systems, handled incidents and worked deeply with Kubernetes, we want to hear from you – production experience and engineering mindset matter more to us than titles or buzzwords.
THE JOY OF WORKING WITH US
Impact from day one: Your work directly influences the revenue of leading eCommerce brands across Europe.
Modern tech stack: Kubernetes, Harvester, GitOps, auto-scaling, and an exciting path toward the cloud – with room to build things right.
AI-first mindset: We use AI as a real part of our daily work, not as a buzzword.
Ownership & growth: Clear responsibility, short decision paths, and the opportunity to actively shape your role.
Flexible work: Hybrid work model three office days per week with a focus on outcomes.
Strong team: Experienced engineers, an open feedback culture, and an environment where reliability is treated as a real engineering discipline.
Job Location
Berlin, Munich, Pforzheim or Stockholm (all Hybrid)
Find more English Speaking Jobs in Germany on Arbeitnow