Senior System Software Engineer - Datacenter Power Management
at Nvidia
USD 224,000-431,200 per year
Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI
Communication @ 4
Debugging @ 7
GPU @ 6
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
NVIDIA is seeking a Senior System Software Engineer to join its Datacenter Performance and Power Management Software team. The role focuses on architecting production software for datacenter power-management controllers, cluster-level dynamic provisioning, and power-management control systems for next-generation datacenters. The work spans hardware architecture, embedded firmware, device drivers, operating systems, and AI workloads across the full product lifecycle, from architectural design and pre-silicon validation through silicon bring-up, productization, and deployment in large-scale datacenter infrastructures.
Responsibilities
- Architect, build, implement, and debug datacenter power- and performance-management software, with an emphasis on end-to-end power management across racks and data halls.
- Develop real-time controllers, policies, and mechanisms that dynamically optimize performance, power, thermals, and energy efficiency under demanding datacenter workloads.
- Drive features across requirements, architecture, implementation, pre-silicon validation, silicon bring-up, productization, and production support.
- Collaborate with GPU architects and hardware designers to define hardware-software interfaces and influence power-management capabilities in next-generation processors.
- Analyze interactions among workloads, clocks, voltages, power limits, thermals, telemetry, and system-level policies.
- Investigate and resolve complex power, performance, stability, and reliability issues across firmware, drivers, hardware, and platform software.
- Develop validation strategies and automation for functional correctness, transition latency, performance-per-watt, and robustness across operating conditions.
- Create architecture, interface, and component specifications for new power-management features.
- Partner with geographically distributed architecture, ASIC, firmware, driver, validation, platform, and datacenter systems teams.
Requirements
- BS or MS degree in Computer Science, Computer Engineering, Electrical Engineering, or a related field, or equivalent experience.
- At least 12 years of industry experience developing system software, embedded firmware, device drivers, or other hardware-related production software.
- Strong C programming skills, including experience developing, debugging, and maintaining complex low-level software.
- Deep understanding of operating-system fundamentals, device-driver architecture, embedded or real-time software, concurrency, interrupt handling, and hardware programming.
- Experience interpreting hardware specifications and developing software interfaces for registers, telemetry, interrupts, and control mechanisms.
- Strong understanding of computer architecture and hardware-software interactions.
- Ability to independently debug complex problems across multiple software and hardware layers.
- Excellent written and verbal communication skills, including experience working across organizational and geographic boundaries.
Preferred Qualifications
- Experience architecting, designing, or implementing datacenter provisioning, monitoring, and power management at a cloud service provider or network/cloud provider.
- Hands-on experience with GPU, CPU, accelerator, or SoC power and performance management.
- Experience implementing P-state selection, DVFS, clock control, voltage control, power capping, thermal management, throttling, or workload-aware control policies.
- Experience developing real-time feedback controllers, including stability, responsiveness, hysteresis, latency, and noisy telemetry considerations.
- Understanding of voltage-frequency relationships, transient power, thermal limits, power delivery, and system-level power budgets.
Compensation and Benefits
- Base salary range: $224,000–$356,500 USD for Level 5.
- Base salary range: $272,000–$431,250 USD for Level 6.
- Base salary is determined by location, experience, and pay for employees in similar positions.
- Eligible for equity and benefits.
- Applications will be accepted at least until September 26, 2026.
- NVIDIA is an equal opportunity employer committed to an inclusive work environment.
More jobs at Nvidia
Senior Full Stack Software Engineer, Data Platform
Nvidia · Santa Clara, United States
USD 140,000-270,200 per year
Senior System Software Engineer - Embedded Controller
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior System Software Engineer - GPU Power Management
Nvidia · Santa Clara, United States
USD 224,000-431,200 per year
Technical Program Manager - LPU Software
Nvidia · Santa Clara, United States
USD 108,000-212,800 per year
Data Application Engineer, Enterprise Data Management
Nvidia · Santa Clara, United States
USD 168,000-310,500 per year
Similar jobs
Platform Security Engineer, OpenBMC
Anthropic · New York City, United States, San Francisco, United States, Seattle, United States
USD 320,000-405,000 per year
Software Engineer - Network (C++)
SpaceXAI · Palo Alto, United States, Seattle, United States
USD 180,000-440,000 per year
Senior Storage Software Engineer - DGX Cloud
Nvidia · Santa Clara, United States
USD 224,000-431,200 per year
Senior System Software Engineer - Dynamo-Triton Inference Server
Nvidia · Santa Clara, United States
USD 224,000-356,500 per year
Technical Program Manager, Multimodal
OpenAI · San Francisco, United States
USD 207,000-445,000 per year
Senior Graphics Driver Engineer
Nvidia · Santa Clara, United States
USD 152,000-241,500 per year
Manager, Software Engineering - Networking
Nvidia · Santa Clara, United States
USD 224,000-431,200 per year
Senior Software Engineer, GNN
Nvidia · Austin, United States
USD 152,000-287,500 per year