Site Reliability Engineer - Vice President
iCapital
Make my CV for this job, freeView job and applyYour CV, rewritten for this role using only your real experience. Sign in with Google and upload your CV. Nothing to install.
Skills named in this job
Read from the description itself, not inferred.
This role on the market
925 open reliability roles across 208 companies are on ApplySarthi right now, most of them in Bengaluru (37), Delhi NCR (13), Pune (9).
- RME Operator with Admin skills, RME (Reliability Maintenance Engineering) Team in ErfurtAmazon Erfurt GmbH - O80
- Senior Site Reliability EngineerCamunda
- Senior Site Reliability Engineer - Hybrid CloudGeniussports
- Site Reliability Engineer / SRE (all genders)Lio
- Site Reliability EngineerDeepJudge
What reliability roles keep asking for: Python (34%), Kubernetes (33%), Observability (31%), AWS (25%), Terraform (22%), Linux (21%), System design (19%), CI/CD (16%) — counted across their open postings here.
Site Reliability Engineer jobs in the United States · Remote Site Reliability Engineer jobs · AWS jobs · Kubernetes jobs · MongoDB jobs · Observability jobs
iCapital has 207 open roles listed here.
- Python Developer - Associatejaipur
- Salesforce Administrator - Assistant Vice President
- Salesforce Administrator - Assistant Vice President
- Business Solutions Consultant II - Associate
- Director, Client Delivery - Vice President
Counted across 14 company job boards, updated as roles open and close.
Preparing for this interview
Interviews for reliability roles keep coming back to Python, Kubernetes, Observability, AWS. Practise those questions before you sit with iCapital.
Questions you are likely to be asked
- Why do you want to join iCapital?
- What is your experience with Observability? Tell me one thing you learned the hard way.
- Tell me about an outage you handled. What did you learn from it?
- How do you decide what to monitor, and what should wake someone up at night?
- How would you cut the cloud bill of a system without hurting it?
Prep Sarthi gives you a free mock interview: an AI interviewer asks you questions like these out loud, from your own CV and this job, and shows your score and your weakest answer.
Practise the Site Reliability Engineer - Vice President at iCapital interview free →- Define, implement, and iterate service level objectives (SLOs) and service level indicators (SLIs) that reflect customer and business expectations.
- Lead monitoring and alerting standardization through “monitors as code” (Terraform preferred), including quality gates such as severity, ownership, and runbook links.
- Develop observability standards across metrics, logs, and traces, including instrumentation and dependency mapping patterns (OpenTelemetry where applicable).
- Lead technical evaluations and PoCs for observability platforms and integrations; define success criteria and migration approach for adoption.
- Define and implement reliability and operability standards for Kubernetes-based services, including scaling patterns, resource constraints, rollout safety, and baseline dashboards and alerts as part of service onboarding.
- Drive automation to eliminate toil, improve repeatability, and accelerate recovery (incident workflows, runbooks, and remediation where appropriate).
- Serve as Incident Commander for high-severity incidents, lead postmortems, and drive systemic improvements through action items and measurable follow-through using established tooling workflows.
- Participate in on-call rotations with a focus on improving reliability, reducing alert noise, and increasing signal quality over time.
- 7+ years in SRE or related roles, with evidence of technical seniority across multiple services and teams.
- Strong experience with AWS and container orchestration (Kubernetes) in production environments.
- Demonstrated experience defining SLOs/SLIs and using them to drive operational and engineering decisions.
- Proven ability to design and implement observability solutions that produce actionable insights while reducing alert fatigue and operational noise.
- Strong IaC skills (Terraform preferred) and the ability to build reusable automation and standards (monitoring as code, configuration patterns).
- Familiarity with common data stores and managed services (e.g., Postgres, MongoDB, DynamoDB) and how they fail in distributed systems.
- Experience with at least two observability stacks (Prometheus/Grafana, New Relic, Splunk, CloudWatch, ELK, etc.) and driving standardization across them.
- Strong incident response skills, including leading retrospectives/postmortems and improving reliability through systematic follow-up.
- Strong debugging skills across distributed systems and production environments, including performance and reliability investigations.
- Clear written and verbal communication skills with the ability to influence engineering teams through standards, tooling, and practical guidance.
Benefits
The base salary range for this role is $130,000 to $160,000 depending on level. iCapital offers a compensation package which includes salary, equity for all full-time employees, and an annual performance bonus. Employees also receive a comprehensive benefits package that includes an employer matched retirement plan, generously subsidized healthcare with 100% employer paid dental, vision, telemedicine, and virtual mental health counseling, parental leave, and unlimited paid time off (PTO).
We believe the best ideas and innovation happen when we are together. Employees in this role will work in the office Monday-Thursday, with the flexibility to work remotely on Friday.
For additional information on iCapital, please visit https://www.icapitalnetwork.com/about-us Twitter: @icapitalnetwork | LinkedIn: https://www.linkedin.com/company/icapital-network-inc | Awards Disclaimer: https://www.icapitalnetwork.com/about-us/recognition/
iCapital is proud to be an Equal Employment Opportunity and Affirmative Action employer. We do not discriminate based upon race, religion, color, national origin, gender, sexual orientation, gender identity, age, status as a protected veteran, status as an individual with a disability, or other applicable legally protected characteristics.
Match this job to your CV
ApplySarthi scores your CV against this role, shows the skills you are missing, and writes a tailored version for the application.
Check my match →Similar open roles
- Actuarial Software Engineer II - AnalystiCapital · jaipur
- Agentic AI Engineer - Assistant Vice President / Vice President iCapital · jaipur
- Alternative Distribution Relationship Manager, Illinois, Minnesota, and Wisconsin - Vice PresidentiCapital
- Alternative Distribution Relationship Manager, Texas and Oklahoma - Vice PresidentiCapital
- Canadian Alternatives Distribution, Eastern Canada - Vice PresidentiCapital
- Configuration Developer - AnalystiCapital · jaipur
- Lead Python Developer iCapital · jaipur
- Python Developer iCapital · jaipur
Need answers during your interview? Try Live Sarthi.
Live Sarthi, an Interview Sarthi app, shows answer suggestions during the call.
- Hidden from supported screen sharingThe overlay stays out of supported Windows screen captures.
- Answers start in about 1.5 secondsResponse time varies with your connection and model.
- From your own CVYour projects and your experience, not a generic script.
- 30 minutes freeThen ₹99 for a 2-day pass with unlimited calls — you pay for the days you are interviewing, not a subscription.
A Windows app, from the same team as ApplySarthi.
Listed on greenhouse · posted 2026-07-07. ApplySarthi collects openings and links to application pages; the role is advertised by iCapital, not by us.