Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)
Fact Finder
Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d) at Fact Finder is a full time role based in Berlin, State of Berlin, Germany. Listed skills: join. It was published on 8 September 2026 and was open at last check.
| Role | Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d) |
|---|---|
| Company | Fact Finder |
| Location | Berlin, State of Berlin, Germany |
| Employment type | Full Time |
| Skills listed | join |
| Published | 8 September 2026 |
| Status | Open at last check |
At a glance
- Location & work model : Berlin, hybrid
- Tech stack: Kubernetes on our own servers, Harvester ( KubeVirt ), Argo CD/Flux, Prometheus/Grafana, Longhorn/Ceph
- Team: A growing SRE team – you report to our CTPO for now and to the Team Lead SRE we're hiring next; two system administrators in Pforzheim run the physical hardware
- Process: Intro call · take-home task (~2h) · 90-min tech interview with our developers · leadership conversation · meet the team
- Languages: Fluent English required; German is a plus, not a must
Why this role is special
Your first 90 days
Define and own SLOs, SLIs and error budgets; drive data-informed reliability decisions Lead incident response end-to-end: fast detection, clear communication, blameless postmortems – and reduce whole classes of incidents structurally, not case by case Eliminate toil through automation and GitOps; evolve our observability (metrics, logs, traces, alerting, runbooks) across two different stacks Help build our custom Kubernetes operator (CRDs) that makes stateful search clusters declarative, self-healing and safely upgradable – and roll out the auto-scaling (HPA/VPA, KEDA, cluster auto scaler) today's architecture makes hard Plan capacity, performance and cost across on-premises and cloud – including the large-catalogue and peak-season loads our merchants care about – and use AI tools wherever they measurably speed up diagnosis and operations
Must-haves:
Kubernetes in production – built, not just used : you've set up and maintained clusters on your own servers (e.g. kubeadm , RKE2, k3s) and know cluster lifecycle and upgrades – managed-only experience isn't enough for this role Lived SRE practice : SLOs, error budgets, incident management, on-call Hands-on experience with GitOps or comparable infrastructure/deployment automation – experience with Argo CD or Flux is a strong plus Solid observability skills – metrics, logs, traces, alerting that people trust A strong automation instinct – you'd rather fix a problem's cause than repeat its workaround A collaborative, enabling mindset – you see SRE as a service to our developers: you ask what they need, discuss trade-offs openly, and don't fall in love with your own solution Nice-to-haves (genuinely optional – we'll teach you the rest):
Harvester, KubeVirt , vSphere/ ESXi , OpenStack or similar virtualization/HCI platforms Container storage (Longhorn, Ceph) and datacenter networking (load balancing, ingress, VLAN) Auto-scaling (HPA, VPA, KEDA, cluster auto scaler ) and capacity/cost planning Experience building Kubernetes operators/CRDs German language skills Certifications (CKA, CKS) are welcome but no substitute for hands-on experience – in the tech interview we'll ask about what you've actually built and operated.
You don't tick every box – or your title was never “SRE”? Apply anyway. If you've owned production systems, handled incidents and worked deeply with Kubernetes, we want to hear from you – production experience and engineering mindset matter more to us than titles or buzzwords.
Impact from day one: Your work directly influences the revenue of leading eCommerce brands across Europe. Modern tech stack: Kubernetes, Harvester, GitOps, auto-scaling, and an exciting path toward the cloud – with room to build things right. AI-first mindset: We use AI as a real part of our daily work, not as a buzzword. Ownership & growth: Clear responsibility, short decision paths, and the opportunity to actively shape your role. Flexible work: Hybrid work model three office days per week with a focus on outcomes. Strong team: Experienced engineers, an open feedback culture, and an environment where reliability is treated as a real engineering discipline.
Berlin, Munich, Pforzheim or Stockholm (all Hybrid)
Create your free OnJob profile to apply — we'll take you to Fact Finder's application after sign-up. · Posted 8 Sept 2026.
Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d) at Fact Finder — questions answered
What does the Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d) role at Fact Finder pay?
Fact Finder does not publish a salary on this Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d) listing, so OnJob shows no figure for it rather than an estimate. For what this role pays across the market, the OnJob salary guides aggregate the live listings that do disclose pay.
Where is the Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d) role at Fact Finder based?
Fact Finder lists this Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d) role in Berlin, State of Berlin, Germany, advertised as full time work at that location. Larger employers sometimes cover several sites under one city name, so confirm the exact office with Fact Finder before you apply.
What skills does the Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d) role at Fact Finder require?
The Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d) at Fact Finder listing names join. Those are the skills the employer put on the posting itself, so they are the ones worth matching in your profile and covering first in an interview.
Is the Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d) role at Fact Finder still open?
The Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d) posting at Fact Finder was open at OnJob's last check of the employer's careers page, having been published on 8 September 2026. OnJob re-checks source listings on each build and marks a role closed once it disappears, but listings can close without notice, so the employer's own page is the final word.
How do you apply for the Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d) role at Fact Finder?
Apply to the Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d) at Fact Finder role through OnJob with a free profile: OnJob scores your fit against the listing, shows the skills lowering that score, and submits an ATS-ready profile to Fact Finder's own application page. Creating a profile is free and needs no card.
Explore more on OnJob
Hiring for a role like this?
Post a job on OnJob and reach AI-matched candidates.