Senior Platform Engineer, Network Infrastructure
Nvidia
Make my CV for this job, freeView job and applyYour CV, rewritten for this role using only your real experience. Sign in with Google and upload your CV. Nothing to install.
Skills named in this job
Read from the description itself, not inferred.
This role on the market
2,030 open infrastructure roles across 290 companies are on ApplySarthi right now, most of them in Bengaluru (68), Hyderabad (28), Delhi NCR (12).
- Member of technical staff (Infrastructure) - ParisHcompany
- Platforms and Infrastructure EngineerEnactintelligence
- Engineering Manager - InfrastructureCamunda
- Senior Infrastructure EngineerRivia
- Senior Project Manager - Ground Systems Infrastructure (m/f/d)isaraerospace
What infrastructure roles keep asking for: AWS (30%), System design (22%), Kubernetes (21%), Python (21%), Observability (19%), Terraform (15%), CI/CD (13%), Linux (12%) — counted across their open postings here.
Platform Engineer jobs in Bengaluru · Platform Engineer jobs in India · Remote Platform Engineer jobs · CI/CD jobs · Deep learning jobs · Kubernetes jobs · Observability jobs
Nvidia has 2,064 open roles listed here.
- Engineering Manager - OpenBMC Platform
- Senior Systems Software Engineer - GPU Performance at Scale
- HPC Operations Engineer
- Senior Systems Software Engineer, Data Center Platform Enablement
- Senior Software Architect - Data Center Systems
Counted across 14 company job boards, updated as roles open and close.
Preparing for this interview
Interviews for infrastructure roles keep coming back to AWS, System design, Kubernetes, Python. Practise those questions before you sit with Nvidia.
Questions you are likely to be asked
- Why do you want to join Nvidia?
- What is your experience with Kubernetes? Tell me one thing you learned the hard way.
- How do you decide what to monitor, and what should wake someone up at night?
- How would you cut the cloud bill of a system without hurting it?
- How do you keep secrets and access safe in your infrastructure?
Prep Sarthi gives you a free mock interview: an AI interviewer asks you questions like these out loud, from your own CV and this job, and shows your score and your weakest answer.
Practise the Senior Platform Engineer, Network Infrastructure at Nvidia interview free →Cloud Foundations Reliability (CFR) is part of NVIDIA’s Global Network Infrastructure (GNI) organization. We deploy, integrate, and operate the Kubernetes-based platform and shared services used to provision, monitor, and operate NVIDIA’s global network across data centers, colocation facilities, and cloud environments. The team owns the architecture and lifecycle of this platform, including cluster provisioning and upgrades, GitOps delivery, observability, capacity, and service enablement. We build software and automation to standardize how network platforms and services are deployed, scaled, and managed across environments. We are looking for a hands-on senior engineer to own the lifecycle and automation of the Kubernetes platform supporting GNI network systems. You will also provide production support for network services running on the platform, partnering with their engineering owners when issues or changes cross the platform boundary. You will take complex problems from design through production and remain accountable for the outcome. You will bring deep Kubernetes expertise and help establish consistent engineering practices across the US and Bangalore teams. This is a senior individual contributor role with end-to-end ownership and production responsibility. What You’ll Be Doing: Design, build, and operate the Kubernetes platform that powers GNI network automation, telemetry, and operations across data center, colocation, and cloud environments. Own the lifecycle management for GNI Kubernetes environments, including cluster onboarding, upgrades, capacity, availability, and recovery. Develop production-quality software and automation for cluster provisioning, validation, upgrades, remediation, and safe multi-cluster delivery through GitOps. Provide production support for network services hosted on the platform, working with Network Automation and service teams that retain ownership of application architecture, code, and features. Diagnose complex Kubernetes platform and hosted-service failures involving control-plane health, cluster networking, storage, scheduling, workload placement, and multi-cluster dependencies. Drive issues from initial signal through verified resolution. Define production-readiness and observability standards for the platform and hosted network services, including health signals, capacity, alerts, runbooks, and recovery. Participate in CFR’s production on-call rotation, including scheduled after-hours and weekend coverage. Lead incident response and recovery, then drive corrective actions to completion. What We Need to See: Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent experience. 8+ years of experience building or operating production Kubernetes platforms, network infrastructure, or distributed systems. Deep experience with Kubernetes at scale, including cluster lifecycle, upgrades, networking, storage, and recovery. Proficiency in at least one general-purpose programming language, such as Go or Python. Experience with GitOps, infrastructure as code, CI/CD, and automated production delivery. Experience deploying and supporting network automation or telemetry services on Kubernetes. Experience with production on-call, incident response, root-cause analysis, and driving corrective actions to completion. Ways to Stand Out From the Crowd: Strong knowledge of IP routing, data center fabrics, and cloud networking is a great plus. Experience designing and operating large, multi-region Kubernetes fleets, including fleet-wide upgrades and recovery. Hands-on experience with Cluster API (CAPI) and Metal3 for bare-metal provisioning, cluster lifecycle, machine remediation, and upgrades. Experience building Kubernetes controllers or operators in Go using custom resources and reconciliation patterns. Experience designing or operating network automation and telemetry services on Kubernetes at global scale. Contributions to Cluster API, Metal3, or other open-source Kubernetes infrastructure projects. NVIDIA’s deep learning platforms have made major impact to various fields is broadly used across leading academic institutions, start-ups, and industry, including the world’s largest Internet companies. We need passionate, hard-working and creative people to help us take on more of these unique opportunities in deep learning cloud solutions. NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hard-working people in the world working for us. Are you creative and autonomous? Do you love a challenge? If so, we want to hear from you.
Match this job to your CV
ApplySarthi scores your CV against this role, shows the skills you are missing, and writes a tailored version for the application.
Check my match →Similar open roles
- Head of Startups - India and South AsiaNvidia · bengaluru
- Manager, AI/HPC Infrastructure Technical Delivery — IndiaNvidia · pune
- Architect – AI-Powered Performance Verification AutomationNvidia · bengaluru
- Data Center Infrastructure SpecialistNvidia · bengaluru
- Senior Developer Relations ManagerNvidia · bengaluru
- Server Performance Architect - HardwareNvidia · bengaluru
- Senior Solutions Architect, Infiniband and Networking Ethernet - NVISNvidia · bengaluru
- Senior Software Engineer, Fabric Networking - GPUNvidia · bengaluru
Need answers during your interview? Try Live Sarthi.
Live Sarthi, an Interview Sarthi app, shows answer suggestions during the call.
- Hidden from supported screen sharingThe overlay stays out of supported Windows screen captures.
- Answers start in about 1.5 secondsResponse time varies with your connection and model.
- From your own CVYour projects and your experience, not a generic script.
- 30 minutes freeThen ₹99 for a 2-day pass with unlimited calls — you pay for the days you are interviewing, not a subscription.
A Windows app, from the same team as ApplySarthi.
Listed on workday · posted 2026-08-27. ApplySarthi collects openings and links to application pages; the role is advertised by Nvidia, not by us.