ApplySarthi

SDE 2 Infra

Fampay

Tailor my CV for this job, freeView job and applyYour CV rewritten for this role, from your real experience. Sign in with Google, nothing to install.

Got this interview? Our apps help you get the job.

Skills named in this job

Read from the description itself, not inferred.

This role on the market

372 open infra roles across 46 companies are on ApplySarthi right now, most of them in Bengaluru (24), Hyderabad (10), Mumbai (5).

What infra roles keep asking for: AWS (27%), Observability (15%), System design (13%) — counted across their open postings here.

CI/CD jobs · Kubernetes jobs · Linux jobs · Observability jobs

Fampay has 17 open roles listed here.

Counted across 14 company job boards, updated as roles open and close.

Preparing for this interview

Interviews for infra roles keep coming back to AWS, Observability, System design. Practise those questions before you sit with Fampay.

Questions you are likely to be asked

  1. Why do you want to join Fampay?
  2. What is your experience with Kubernetes? Tell me one thing you learned the hard way.
  3. Describe a time a deadline forced a trade-off in quality. What did you choose and why?
  4. How would you design an API for a feature you have worked on?
  5. What do you do when a production issue happens on your code?

Prep Sarthi gives you a free mock interview: an AI interviewer asks you questions like these out loud, from your own CV and this job, and shows your score and your weakest answer.

Practise the SDE 2 Infra at Fampay interview free →

About Fam (previously FamPay) Fam is India’s first payments app for everyone above 11. FamApp helps make online and offline payments through UPI and FamCard. We are on a mission to raise a new, financially aware generation, and drive 250 million+ young users in India to kickstart their financial journey super early in their life. We’re reimagining how the next generation experiences fintech—going beyond payments to build a lifestyle brand that blends money, identity, and everyday experiences into one seamless, intuitive journey. Founded in 2019 by IIT Roorkee alumni, Fam is backed by some of the most respected investors around the world like Elevation Capital, Y-Combinator, Peak XV (Sequoia Capital) India, Venture Highway, Global Founder’s Capital and the likes of Kunal Shah, Amrish Rao as angel investors. About the team The Core Infrastructure team builds the platforms and systems that power everything at Fam, from real-time app experiences to high-volume payments. It operates as a small, high-leverage team with deep ownership. Every service at Fam runs on the platform this team builds, so it cares obsessively about system fundamentals, operational excellence, and mechanical sympathy. What our systems process 1 Billion+ API requests per day powering ~10M daily transactions for 10M+ users at 20,000+ RPS. 470+ node Kubernetes cluster (~2,500 vCPUs and ~8 TiB RAM) under active orchestration. Dynamic, multi-dimensional auto-scaling powered by 195+ HPAs paired with Karpenter. 12+ TB of active transactional data (and counting) across our high-throughput database clusters. Event Streaming - 400+ topics, and 3,800+ partitions at ~46,700 messages/sec Our Next Chapter While our core systems already handle massive scale, our next vision as a team is to turn infrastructure into a frictionless platform discipline. We are building towards a world where internal engineering teams consume platform primitives compute, storage, and event streams via self-serve contracts without waiting on us. We're also laying the groundwork for AI-native platform operations, developing agent plane and agent skills so agents can safely interact & operate infrastructure primitives alongside our team. We look to the CNCF landscape for best-in-class tools our systems, with a long-term commitment to contributing back to open source. What You'll Do: Build and operate compute, storage, and networking at scale that powers FamApp. Design platform contracts for other teams: Modules, golden paths, and abstractions with clear inputs, outputs, and guarantees. Take medium-to-large infrastructure projects from design to rollout, manage dependencies across engineering teams, and work through ambiguity without waiting to be unblocked. Build idempotent, retry-safe automation with its own observability, integrated with CI/CD, ticketing, compliance, and incident workflows. Build reusable infrastructure with Infrastructure as Code and Configuration as Code with clear interfaces, validation and versioning backed by safe provisioning ci/cd workflows. Operate Kubernetes as the Operating system - Own the cluster lifecycle (worker plane, node groups, add-ons) . Ability to extend it with controllers, operators, and CRDs so teams consume platform capabilities as Kubernetes-native APIs. Define user-facing SLIs, SLO targets, burn-rate alerts, and error-budget actions for what you own. Carry the pager, cut alert noise, and lead post-incident follow-up with blameless postmortem. Track cloud spend for your services and look for savings without compromising reliability or compliance. Maintain golden pipeline templates, GitOps release patterns, and self-service onboarding, with security gates and least-privilege pipeline identities built in, which enables developers to ship their code to different environments. Turn requirements into technical design docs that cover the full life of a system: how teams adopt it, how it's operated and maintained, and how it stays reliable (monitoring, SLIs, backup, DR, incident response), so anyone on the team can follow them. Must Haves: 3-6 years in DevOps, SRE, or platform engineering, with ownership of production systems. Cloud compute and Linux fundamentals - Instances, images, and block storage on any major cloud: diagnosing throughput/IOPS limits, memory pressure, inode and file-descriptor exhaustion, and systemd issues, and correlating OS-level evidence with cloud metrics and dashboards. Cloud networking - Network and subnet design, CIDR planning, routing, firewall and security-group rules, and DNS. You can design cross-account or cross-project connectivity, weigh peering against transit hubs and private endpoints, and debug asymmetric routing and overlapping CIDRs. Hands-on Kubernetes in production managed services like EKS, GKE, or AKS. Terraform or equivalent IaC with reusable, versioned modules and safe change workflows. Programming ability in Python or Go to build automation that is idempotent, retry-safe, and observable. CI/CD and GitOps experience - pipeline templates, release and rollback patterns, security gates. Reliability practice: SLIs/SLOs, actionable alerting, on-call, and blameless post-mortems. Everyday use of AI tools, with the habit of validating their output before it ships. Good to Have: Karpenter in production - consolidation, spot handling, multi-NodePool strategy. Observability at scale : Prometheus/VictoriaMetrics, Loki, Jaeger, OpenTelemetry, high-cardinality metrics. CI/CD supply-chain security (SAST, secret detection, image scanning, artifact promotion). ClickHouse Operator on kubernetes or another analytical store solutions managed at scale in production cluster. PCI, RBI, SOC 2, or similar compliance-heavy environments. Experience in building scaling and managing infrastructure for Customer focused products.

Match this job to your CV

ApplySarthi scores your CV against this role, shows the skills you are missing, and writes a tailored version for the application.

Check my match →

Similar open roles

Need answers during your interview? Try Live Sarthi.

Live Sarthi, an Interview Sarthi app, shows answer suggestions during the call.

Try Live Sarthi free →

A Windows app, from the same team as ApplySarthi.

Listed on lever · posted 2026-10-01. ApplySarthi collects openings and links to application pages; the role is advertised by Fampay, not by us.