Senior Site Reliability Engineer
Garner Health
Make my CV for this job, freeView job and applyYour CV, rewritten for this role using only your real experience. Sign in with Google and upload your CV. Nothing to install.
Skills named in this job
Read from the description itself, not inferred.
This role on the market
925 open reliability roles across 208 companies are on ApplySarthi right now, most of them in Bengaluru (37), Delhi NCR (13), Pune (9).
- RME Operator with Admin skills, RME (Reliability Maintenance Engineering) Team in ErfurtAmazon Erfurt GmbH - O80
- Senior Site Reliability EngineerCamunda
- Senior Site Reliability Engineer - Hybrid CloudGeniussports
- Site Reliability Engineer / SRE (all genders)Lio
- Site Reliability EngineerDeepJudge
What reliability roles keep asking for: Python (34%), Kubernetes (33%), Observability (31%), AWS (25%), Terraform (22%), Linux (21%), System design (19%), CI/CD (16%) — counted across their open postings here.
Remote Site Reliability Engineer jobs · AWS jobs · Go jobs · Kubernetes jobs · Observability jobs
Garner Health has 81 open roles listed here.
- Senior Commercial Counsel
- FP&A Manager, GTM
- Senior Security Engineer
- Senior Security Engineer
- Staff Security Engineer
Counted across 14 company job boards, updated as roles open and close.
Preparing for this interview
Interviews for reliability roles keep coming back to Python, Kubernetes, Observability, AWS. Practise those questions before you sit with Garner Health.
Questions you are likely to be asked
- Why do you want to join Garner Health?
- What is your experience with Observability? Tell me one thing you learned the hard way.
- Walk me through how code gets from a commit to production where you work.
- Tell me about an outage you handled. What did you learn from it?
- How do you decide what to monitor, and what should wake someone up at night?
Prep Sarthi gives you a free mock interview: an AI interviewer asks you questions like these out loud, from your own CV and this job, and shows your score and your weakest answer.
Practise the Senior Site Reliability Engineer at Garner Health interview free →What you’ll be part of
Garner is on a mission to transform the U.S. healthcare system — and we’re the only proven player doing exactly that. We partner with employers to redesign how healthcare works: applying 550+ proprietary clinical metrics across 80+ specialties to a dataset of 320M+ patients to identify the best-performing doctors, then using compelling incentives to steer members to the care that helps them get healthier, faster.
The result is a rare “win win” — better care and lower costs for both members and employers. In just five years, our work has helped over 2.5 million people access higher-quality care and saved $1B in healthcare costs. We recently raised our Series E and have doubled five years running. If you've ever wanted your work to solve a problem that touches every person in this country, this is the opportunity to do exactly that. You'd be joining a team fundamentally reimagining healthcare in the U.S. — and using AI to scale that impact further and faster than anyone else can.
About the role:
We are seeking a Senior Site Reliability Engineer to own the reliability, performance, and resilience of the cloud infrastructure powering Garner’s products and AI/ML workloads. This role sits on our Platform Engineering team. You will run the machine: defining and upholding SLOs, leading incident response, and driving the automation and standards that let every Garner engineer ship faster and more reliably. Because our systems directly influence health outcomes for millions of patients, maintaining the highest standards of production quality is imperative. This is an automation-first role: you will use AI tools to continuously convert manual operational work into monitored, hands-free processes, so the role gets more leveraged as you build.
Where you will work:
Garner is headquartered in NYC, but this position is available for individuals who are comfortable with remote work and occasional travel to HQ.
What you will do:
- Run the Machine: Own the end-to-end reliability, performance, and resilience of Garner’s cloud environments (AWS, Kubernetes), including those powering AI/ML workloads; define, measure, and uphold SLOs across our critical services
- Lead Incident Response: Serve in the on-call rotation, lead incident response, and drive deep-dive root cause analysis, seeing corrective actions through to resolution and rigorously reviewing infrastructure changes
- Own Observability: Build and maintain the monitoring, alerting, and observability systems that let us detect and resolve issues before users feel them
- Scale & Optimize: Translate ambiguous, high-performance scaling requirements into well-defined, automated, and composable infrastructure-as-code deliverables (Terraform); proactively identify and implement cost-efficiency and performance gains across the stack to maximize cloud ROI
- Automate Away Toil: Pay down impactful tech debt and reduce operational toil, using AI tools and automation to convert repetitive operational work into hands-free, monitored processes, and holding our internal platform to the same rigorous standards as our customer-facing products
- Enable Engineering: Build and maintain the deployment and observability standards that empower the broader engineering team to ship AI features faster and more reliably; communicate complex cloud and reliability concepts clearly to technical and non-technical stakeholders
- Uphold Security & Compliance: Ensure our infrastructure and operations meet Garner’s security and HIPAA compliance obligations
The ideal candidate has:
- 4+ years of hands-on experience operating production cloud infrastructure at scale in an SRE, DevOps, or platform engineering role
- Deep expertise with Kubernetes and Terraform in a cloud-first environment (AWS preferred)
- A strong track record with production observability: defining SLOs, building monitoring and alerting, and leading incident response and blameless post-incident reviews
- Strong software engineering fundamentals in Python or Go, applied to infrastructure automation (experience with Kubernetes APIs a plus)
- Experience driving cloud cost-efficiency and performance optimization across compute, storage, and networking
- Experience supporting AI/ML or data-intensive workloads in production is a plus
- Experience operating in a security-conscious or regulated environment (HIPAA, SOC 2) is a plus
- Fluency with AI tools (e.g., Claude) applied to real engineering and operations workflows, or strong motivation to build it fast
- A desire to be a part of a high-performing, mission-driven team that operates with intense urgency, a strong sense of individual accountability, and a commitment to authentic feedback
Technologies we use:
- AWS, Kubernetes, Terraform, Istio, Python, Go, TypeScript, Postgres, NATS, Datadog, GitLab
This is a unique opportunity to join a fast-growing company in a transformative role, helping shape the future of healthcare.
We are unable to sponsor or take over sponsorship of an employment visa at this time.
Compensation Transparency:
The target range for this position is $191,000 - $226,000. Individual compensation for this role will depend on various factors, including qualifications, skills, and applicable laws. In addition to base compensation, this role is eligible to participate in our equity incentive and competitive benefits plans, including but not limited to: flexible PTO, Medical/Dental/Vision plan options, 401(k) with company match, flexible spending accounts, Teladoc Health and more.
Fraud and Security Notice:
Please be aware of recent job scam attempts. Our recruiters use getgarner.com and garnerhealth.com email domains exclusively. If you have been contacted by someone claiming to be a Garner recruiter or a hiring manager from a different domain about a potential job, please report it to law enforcement here and to candidateprotection@garnerhealth.com.
Equal Employment Opportunity:
Garner Health is proud to be an Equal Employment Opportunity employer and values diversity in the workplace. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, reproductive health decisions, or related medical conditions), sexual orientation, gender identity, gender expression, age, status as a protected veteran, status as an individual with a disability, genetic information, political views or activity, or other applicable legally protected characteristics.Garner Health is committed to providing accommodations for qualified individuals with disabilities in our recruiting process. If you need assistance or an accommodation due to a disability, you may contact us at talent@garnerhealth.com.
Match this job to your CV
ApplySarthi scores your CV against this role, shows the skills you are missing, and writes a tailored version for the application.
Check my match →Similar open roles
- Account Executive - CentralGarner Health
- Account Executive - Mid-CentralGarner Health
- Account Executive - NortheastGarner Health
- Account Executive - SouthwestGarner Health
- Account Executive - WestGarner Health
- Business Development RepresentativeGarner Health
- Client Success Manager (Small-Market)Garner Health
- Consultant Relations VPGarner Health
Need answers during your interview? Try Live Sarthi.
Live Sarthi, an Interview Sarthi app, shows answer suggestions during the call.
- Hidden from supported screen sharingThe overlay stays out of supported Windows screen captures.
- Answers start in about 1.5 secondsResponse time varies with your connection and model.
- From your own CVYour projects and your experience, not a generic script.
- 30 minutes freeThen ₹99 for a 2-day pass with unlimited calls — you pay for the days you are interviewing, not a subscription.
A Windows app, from the same team as ApplySarthi.
Listed on greenhouse · posted 2026-09-02. ApplySarthi collects openings and links to application pages; the role is advertised by Garner Health, not by us.