Senior System GPU Performance Engineer
Nvidia
Tailor my CV for this job, freeView job and applyYour CV rewritten for this role, from your real experience. Sign in with Google, nothing to install.
Got this interview? Our apps help you get the job.
Skills named in this job
Read from the description itself, not inferred.
This role on the market
698 open performance roles across 180 companies are on ApplySarthi right now, most of them in Bengaluru (42), Hyderabad (23), Pune (10).
- Automated Driving Performance HIL Infrastructure EngineerGeneralmotors
- Business Performance Mgr - OperationsAmgen
- Managing Consultant, Advisors & Consulting Services, Performance AnalyticsMastercard
- Investment Performance Analysis - Team LeadStatestreet · bengaluru
- Manager, Performance MarketingPaypal
What performance roles keep asking for: Python (15%), C++ (15%), Stakeholder management (13%) — counted across their open postings here.
C++ jobs · Product management jobs · Python jobs · SQL jobs
Nvidia has 2,295 open roles listed here.
- Senior Compiler Optimization Engineer – LLVMbengaluru
- Engineering Manager - OpenBMC Platform
- Principal Firmware Engineer - Data Center Server Management
- Senior Systems Software Engineer- EDA Infrastructure
- Senior Data Backend Engineer
Counted across 14 company job boards, updated as roles open and close.
Preparing for this interview
Interviews for performance roles keep coming back to Python, C++, Stakeholder management. Practise those questions before you sit with Nvidia.
Questions you are likely to be asked
- Why do you want to join Nvidia?
- What is your experience with C++? Tell me one thing you learned the hard way.
- How would you design an API for a feature you have worked on?
- What do you do when a production issue happens on your code?
- Walk me through a system you built. How was it designed, and what would you change now?
Prep Sarthi gives you a free mock interview: an AI interviewer asks you questions like these out loud, from your own CV and this job, and shows your score and your weakest answer.
Practise the Senior System GPU Performance Engineer at Nvidia interview free →NVIDIA has transformed computer graphics, PC gaming, and accelerated computing for more than 25 years. Today, our GPUs power advances in AI, Datacenter, Gaming, Robotics, Automotive, and scientific discovery. NVIDIA's Silicon Co-Design Group (SCG) takes GPU, SoC, and CPU programs from first power-on to high-volume production. We sit at the crossroads of architecture, design, marketing, operations, and productization across Datacenter, Gaming, Robotics, Automotive, and Embedded markets. We are hiring a Senior System GPU Performance Engineer to maximize the performance and power efficiency of production GPU systems. You will connect workload behavior, silicon capability, software policy, and platform constraints to identify bottlenecks and productize improvements. This is not a benchmark-execution or validation-only role—you will own analysis from hypothesis through root-cause closure, plan-of-record integration, and confirmed product impact. Great work turns complex system data into faster, more efficient, and more predictable products. What you'll be doing: Own system-level GPU performance and power characterization from first silicon through production across representative applications, benchmarks, and product configurations. Drive performance and power feature productization, translating measured behavior into firmware, driver, BIOS, platform, and silicon recommendations that meet product targets and speed-of-light schedules. Design experiments, execute test plans, and build models that isolate bottlenecks across GPU compute, memory, interconnect, CPU interaction, power delivery, and thermal limits. Analyze production-silicon data across process, voltage, temperature, workloads, and bins to quantify performance-per-watt trade-offs and identify causal optimization opportunities. Lead multi-functional root-cause closure across architecture, design, validation, software/firmware, power and thermal, reliability, ATE, product management, manufacturing, and operations; own fixes through confirmation. Establish reusable automation, visualization, and closed-loop methodologies that improve experiment coverage, analysis accuracy, debug velocity, and learning across future GPU programs. Translate complex system signals into decision-ready options for executive leadership on feature readiness, product configuration, targets, and program risks. What we need to see: BS or MS in Electrical Engineering, Computer Engineering, Computer Science, Systems Engineering, or related field (or equivalent experience). 8+ overall years of experience in GPU or system performance engineering, post-silicon characterization, silicon productization, or hardware-software performance optimization. Hands-on experience with silicon bring-up, frequency and power characterization, product binning, and performance-per-watt optimization across process, voltage, temperature, workloads, and system configurations. Strong understanding of GPU and system architecture, including compute pipelines, memory hierarchy, interconnects, CPU-GPU interactions, scheduling, telemetry, and sustained-performance limits. Proven ability to design controlled experiments, develop performance or power models, analyze large datasets, and use statistics to separate bottlenecks and causal effects from noise. Strong programming and analysis skills using Python and one or more of C, C++, SQL, JMP, or equivalent, with experience automating tests, data processing, and visualization. Demonstrated ability to structure ambiguous system-level problems and drive them to root-cause closure across globally distributed, multi-functional hardware and software teams. Strong written and verbal communication; able to translate complex technical issues into crisp, decision-ready options for executive leadership. Ways to stand out from the crowd: Track record of shipping GPU performance or power features that measurably improved application performance, performance per watt, product segmentation, or time to market. Experience optimizing large GPU, CPU, AI accelerator, or other complex SoC platforms for Datacenter, Gaming, Automotive, Robotics, or Embedded products. Deep experience with GPU profiling, workload characterization, production telemetry, performance counters, or simulation-to-silicon correlation. Experience building reusable performance models, test frameworks, or analysis methodologies adopted across multiple silicon programs or advanced process nodes. Applied AI tools to accelerate experiment design, anomaly detection, debug, analysis, or reporting workflows and can describe the measurable outcome and the guardrails used to protect correctness. NVIDIA is the world leader in accelerated computing, powering AI, gaming, robotics, autonomous systems, and scientific discovery. We invest in our people with competitive benefits, continuous learning, and a team where everyone can do their best work. NVIDIA is committed to fostering a diverse work environment and is proud to be an equal opportunity employer. We do not discriminate on the basis of race, color, national origin, gender, gender identity, sexual orientation, religion, age, marital status, veteran status, disability, or any other legally protected status.
Match this job to your CV
ApplySarthi scores your CV against this role, shows the skills you are missing, and writes a tailored version for the application.
Check my match →Similar open roles
- Head of Startups - India and South AsiaNvidia · bengaluru
- Manager, AI/HPC Infrastructure Technical Delivery — IndiaNvidia · pune
- Architect – AI-Powered Performance Verification AutomationNvidia · bengaluru
- Data Center Infrastructure SpecialistNvidia · bengaluru
- Senior Developer Relations ManagerNvidia · bengaluru
- Server Performance Architect - HardwareNvidia · bengaluru
- Senior Solutions Architect, Infiniband and Networking Ethernet - NVISNvidia · bengaluru
- Senior Software Engineer, Fabric Networking - GPUNvidia · bengaluru
Need answers during your interview? Try Live Sarthi.
Live Sarthi, an Interview Sarthi app, shows answer suggestions during the call.
- Hidden from supported screen sharingThe overlay stays out of supported Windows screen captures.
- Answers start in about 1.5 secondsResponse time varies with your connection and model.
- From your own CVYour projects and your experience, not a generic script.
- 30 minutes freeThen ₹99 for a 2-day pass with unlimited calls — you pay for the days you are interviewing, not a subscription.
A Windows app, from the same team as ApplySarthi.
Listed on workday · posted 2026-10-01. ApplySarthi collects openings and links to application pages; the role is advertised by Nvidia, not by us.