Site Reliability Engineering Lead
Jobgether
Tailor my CV for this job, freeView job and applyYour CV rewritten for this role, from your real experience. Sign in with Google, nothing to install.
Got this interview? Our apps help you get the job.
Skills named in this job
Read from the description itself, not inferred.
This role on the market
943 open reliability roles across 175 companies are on ApplySarthi right now, most of them in Bengaluru (39), Delhi NCR (12), Pune (9).
- Manufacturing Engineer, Trainium Manufacturing, Quality and ReliabilityAnnapurna Labs
- Observability DevOps Engineer - RDT Digital Operations and ReliabilityRoche
- Software Engineer – Site Reliability Engineering Qube Research & Technologies
- Sr. Site Reliability Engineer (US Federal)Workday
- Intern - NAND Wafer ReliabilityMicron
What reliability roles keep asking for: Python (38%), Observability (35%), Kubernetes (34%), AWS (29%), Terraform (24%), Linux (22%), System design (19%), CI/CD (18%) — counted across their open postings here.
Azure jobs · CI/CD jobs · IAM jobs · Kubernetes jobs
Jobgether has 4,273 open roles listed here.
- (Technical Targeter- Virtual Operations) Cyber Technical Analyst Principal (TS/SCI with Poly Required)
- Advogado(a) | Bancário
- Agentic Workforce Adoption Manager
- AI Engineer
- AI Marketing Project Manager
Counted across 14 company job boards, updated as roles open and close.
Preparing for this interview
Interviews for reliability roles keep coming back to Python, Observability, Kubernetes, AWS. Practise those questions before you sit with Jobgether.
Questions you are likely to be asked
- Why do you want to join Jobgether?
- What is your experience with Observability? Tell me one thing you learned the hard way.
- How would you cut the cloud bill of a system without hurting it?
- How do you keep secrets and access safe in your infrastructure?
- Walk me through how code gets from a commit to production where you work.
Prep Sarthi gives you a free mock interview: an AI interviewer asks you questions like these out loud, from your own CV and this job, and shows your score and your weakest answer.
Practise the Site Reliability Engineering Lead at Jobgether interview free →Accountabilities: Lead, mentor, and develop a small to medium-sized team of Site Reliability Engineers through regular 1:1s, performance reviews, career planning, and ongoing coaching. Own hiring, onboarding, team capacity, resourcing, and workforce planning decisions to ensure the team can effectively support business and platform priorities. Establish team objectives, prioritize the engineering backlog, coordinate planning, and ensure projects and operational tasks remain aligned with reliability goals. Lead reliability initiatives across infrastructure and services, improving availability, scalability, resilience, security, and operational performance. Drive incident response activities and facilitate blameless post-incident reviews, ensuring timely root-cause analyses and actionable follow-up. Partner with Development, Security, Product, and other engineering teams to resolve cross-functional issues and strengthen collaboration. Champion automation and operational excellence by reducing manual work, eliminating recurring toil, and introducing self-healing systems and infrastructure automation. Support the design and evolution of scalable, secure, cloud-native environments and continuously identify opportunities to improve performance, reliability, and cost efficiency. Requirements Demonstrated experience in SRE, DevOps, infrastructure engineering, or a related discipline, including experience leading engineering teams. Expert-level knowledge of Kubernetes, including cluster architecture, upgrades, autoscaling, security hardening, and large-scale troubleshooting. Advanced experience with Terraform, including modular infrastructure-as-code design, state management, multi-environment provisioning, and policy-as-code. Deep knowledge of Azure Cloud services, including compute, networking, identity and access management, storage, and cost optimization. Experience designing and scaling CI/CD pipelines using GitHub Actions, release strategies, and automated rollback approaches. Strong knowledge of observability platforms such as Prometheus, Grafana, and OpenTelemetry, along with experience managing SLOs, SLAs, and error budgets. Strong automation capabilities and advanced proficiency in Python, Bash, and/or PowerShell for infrastructure tooling and operational automation. Deep understanding of networking fundamentals, including TCP/IP, DNS, load balancing, VPNs, and cloud-native networking. Proven experience leading incident response, conducting root-cause analysis, and implementing measurable reliability improvements. Strong people leadership, communication, prioritization, and problem-solving skills, with the ability to support engineers while coordinating effectively across technical teams. Benefits U.S. national base salary range of $118,300–$219,800 , with geographic differentials potentially applying depending on location. Eligibility for an annual incentive bonus . Country- and location-specific employee benefits designed to support overall health and well-being. Support for an accessible and inclusive hiring process, including reasonable accommodations where required. Remote/home-based opportunities available in multiple U.S. locations, including Florida, Connecticut, New Jersey, New York, and Pennsylvania. Opportunity to lead a technically sophisticated SRE function focused on cloud platforms, automation, resilience, and operational excellence. Professional development opportunities through team leadership, cross-functional collaboration, and exposure to large-scale cloud environments.
Match this job to your CV
ApplySarthi scores your CV against this role, shows the skills you are missing, and writes a tailored version for the application.
Check my match →Similar open roles
- Accounting & Regulatory ReportingJobgether
- Accounts Receivable CoordinatorJobgether
- AI Security AnalystJobgether
- Assistant Manager of Virtual Legal SupportJobgether
- Associate Experience Design ResearcherJobgether
- Associate ProducerJobgether
- AVP, Corporate PartnershipsJobgether
- AWS ManagerJobgether
Need answers during your interview? Try Live Sarthi.
Live Sarthi, an Interview Sarthi app, shows answer suggestions during the call.
- Hidden from supported screen sharingThe overlay stays out of supported Windows screen captures.
- Answers start in about 1.5 secondsResponse time varies with your connection and model.
- From your own CVYour projects and your experience, not a generic script.
- 30 minutes freeThen ₹99 for a 2-day pass with unlimited calls — you pay for the days you are interviewing, not a subscription.
A Windows app, from the same team as ApplySarthi.
Listed on lever · posted 2026-10-02. ApplySarthi collects openings and links to application pages; the role is advertised by Jobgether, not by us.