Team Lead - Site Reliability Engineering (all genders)
FactFinder
Tailor my CV for this job, freeView job and applyYour CV rewritten for this role, from your real experience. Sign in with Google, nothing to install.
Got this interview? Our apps help you get the job.
Skills named in this job
Read from the description itself, not inferred.
This role on the market
937 open reliability roles across 175 companies are on ApplySarthi right now, most of them in Bengaluru (39), Delhi NCR (12), Hyderabad (9).
- Platform Reliability & Salesforce EngineerOdaseva
- Maintenance Supervisor, RME (Reliability Maintenance Engineering)Amazon
- Sr. Technical Program Manager - AI Manufacturing, Annapurna AI Manufacturing Quality and ReliabilityAnnapurna Labs
- Sr. Supplier Quality Engineer, IRQ (Infrastructure Reliability & Quality)Amazon Web Services
- Senior Manager, Technical Program Management (Data, Reliability & Developer Experience)Airbnb
What reliability roles keep asking for: Python (39%), Observability (36%), Kubernetes (34%), AWS (29%), Linux (24%), Terraform (23%), CI/CD (20%), System design (19%) — counted across their open postings here.
Kubernetes jobs · Observability jobs
FactFinder has 2 open roles listed here.
Counted across 14 company job boards, updated as roles open and close.
Preparing for this interview
Interviews for reliability roles keep coming back to Python, Observability, Kubernetes, AWS. Practise those questions before you sit with FactFinder.
Questions you are likely to be asked
- Why do you want to join FactFinder?
- What is your experience with Kubernetes? Tell me one thing you learned the hard way.
- How do you decide what to monitor, and what should wake someone up at night?
- How would you cut the cloud bill of a system without hurting it?
- How do you keep secrets and access safe in your infrastructure?
Prep Sarthi gives you a free mock interview: an AI interviewer asks you questions like these out loud, from your own CV and this job, and shows your score and your weakest answer.
Practise the Team Lead - Site Reliability Engineering (all genders) at FactFinder interview free →Introduction FACT-Finder builds product discovery technology for eCommerce and is trusted by leading online shops across Europe with its two products Next Generation and Infinity. We are actively modernizing our hosting toward Kubernetes on Harvester – as an on-prem hybrid with the option to scale fully into the cloud in the mid-term. As Team Lead Site Reliability Engineering (all genders), you own the reliability, scalability, and cost of our hosting environments, drive this transformation end-to-end, and lead the team that delivers it. Your mission You own the operational health of our hosting across on-premise (Frankfurt, Stockholm) and cloud – availability, performance, and incident management. You actively drive the modernization toward Kubernetes on Harvester: cluster topology, storage (Longhorn), networking (VLAN, load balancing, ingress), backup, and disaster recovery. You build a production-grade k8s platform: lifecycle, upgrades, RBAC, secrets, GitOps (Argo CD / Flux), observability, and policy guardrails. You shape the NG Search Operator (custom Kubernetes operator) and solve auto-scaling (HPA, VPA, KEDA, cluster autoscaler) for the current architecture. You concretely define our on-prem hybrid model: which workloads run where, how we burst into the cloud, how we keep latency and cost under control – while keeping the architecture portable enough for a future cloud-only move. You own capacity planning and hosting cost and turn cost into a deliberate, managed lever. You lead and develop our currently 4-person Hosting team, own performance and technical direction, and set the standards and ownership culture. You make AI a core part of our operations: diagnosis, automation, monitoring, and insight. Your profile Strong background in infrastructure or platform engineering across on-premise and cloud. Hands-on depth with Kubernetes in production: cluster lifecycle, upgrades, networking, storage, RBAC, observability, GitOps delivery. Proven people leadership experience, excellent communication and stakeholder management skills. Ideally practical experience with Harvester or comparable HCI/virtualization platforms (KubeVirt, vSphere/ESXi, OpenStack). Experience leading a real migration from bare metal / classic VMs to a k8s-based platform – including stateful workloads, storage migration, cutover, and rollback. Comfort designing or operating Kubernetes operators (custom controllers / CRDs), ideally for stateful systems like search, databases, or streaming. Solid grasp of auto-scaling primitives (HPA, VPA, cluster autoscaler, KEDA) and how they interact with capacity planning on-prem and in the cloud. Experience with on-prem hybrid architectures and owning reliability, capacity, and cost for production systems. Hands-on fluency with AI tools in day-to-day operations. Fluent English; German is a plus. THE JOY OF WORKING WITH US Impact from day one: Your work directly influences the revenue of leading eCommerce brands across Europe. Leadership with real scope: You lead an established team and shape our platform in a decisive phase of our transformation. Modern tech stack: Kubernetes, Harvester, GitOps, auto-scaling, and an exciting path toward the cloud – with room to build things right. AI-first mindset: We use AI as a real part of our daily work, not as a buzzword. Ownership & growth: Clear responsibility, short decision paths, and the opportunity to actively shape your role. Flexible work: Hybrid work model with a focus on outcomes. Strong team: Experienced engineers, an open feedback culture, and an environment where reliability is treated as a real engineering discipline. Attractive benefits: Competitive salary, modern equipment, learning budget, and regular team events. Job Location Berlin, Munich, Pforzheim or Stockholm (hybrid) Find Jobs in Germany on Arbeitnow
Match this job to your CV
ApplySarthi scores your CV against this role, shows the skills you are missing, and writes a tailored version for the application.
Check my match →Similar open roles
Need answers during your interview? Try Live Sarthi.
Live Sarthi, an Interview Sarthi app, shows answer suggestions during the call.
- Hidden from supported screen sharingThe overlay stays out of supported Windows screen captures.
- Answers start in about 1.5 secondsResponse time varies with your connection and model.
- From your own CVYour projects and your experience, not a generic script.
- 30 minutes freeThen ₹99 for a 2-day pass with unlimited calls — you pay for the days you are interviewing, not a subscription.
A Windows app, from the same team as ApplySarthi.
Listed on arbeitnow · posted 2026-10-07. ApplySarthi collects openings and links to application pages; the role is advertised by FactFinder, not by us.