Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI
Communication @ 7
GPU
HPC
InfiniBand
Mentoring @ 7
NVLink
Networking
People Management @ 7
System Architecture @ 7
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
NVIDIA is seeking a strong technology leader to manage its Server Software Technical Program Management (TPM) team. This role operates at the intersection of execution and strategy, leading a team of Senior TPMs responsible for firmware and system software for NVIDIA's next-generation server platforms, including DGX, MGX, and HGX. These platforms combine NVIDIA GPUs, NVLink, InfiniBand networking, Grace CPUs, and an optimized AI/HPC software stack. The role focuses on software development processes that bring new server hardware to life.
Responsibilities
- Lead a team of TPMs driving technical software and firmware execution for NVIDIA's New Product Introduction (NPI) and sustaining engineering teams.
- Drive the end-to-end software development lifecycle (SDLC) for low-level server components, including BMC, UEFI/BIOS firmware, drivers, and system management software, ensuring alignment with hardware schedules.
- Collaborate with NVIDIA product management and hardware engineering teams to define release plans and program objectives.
- Build a strong connection and feedback loop between sustaining and NPI engineering teams to improve product quality and development velocity.
- Lead process improvement initiatives and help propagate SDLC standards across multiple engineering and TPM organizations.
- Interact with diverse technical groups across all organizational levels.
Requirements
- Bachelor of Science degree, equivalent experience, or Master of Science degree in Computer Science, Electrical Engineering, or a related field.
- At least 12 years of overall experience developing and leading complex low-level or system software projects.
- At least 7 years of experience in people management.
- Deep understanding of system architecture, including server topologies, out-of-band management, UEFI/BIOS, interconnects such as PCIe and CXL, memory management, and RAS architecture.
- Experience working with complex system software for accelerators such as GPUs, DPUs, or FPGAs.
- Strong interpersonal, verbal, and written communication skills, with the ability to achieve objectives under fast-paced timelines.
- Proven ability to lead multiple projects with competing priorities.
- Strong people management and mentoring skills, with a consistent record of building cohesive teams.
Preferred Qualifications
- Prior Senior Manager experience leading engineering or program management teams.
- Deep understanding of software engineering principles and large-scale enterprise system architecture.
- Experience coordinating activities between hardware, firmware, and software application organizations.
Compensation and Benefits
- Base salary range: $272,000–$425,500 USD, determined by location, experience, and compensation for similar positions.
- Eligible for equity and benefits.
- Applications will be accepted at least until August 3, 2026.
- NVIDIA is an equal opportunity employer committed to fostering an inclusive work environment.
More jobs at Nvidia
Senior Solution Engineer, Networking
Nvidia · Santa Clara, United States
USD 168,000-322,000 per year
Senior AI and ML Software Engineer
Nvidia · Santa Clara, United States
USD 140,000-270,200 per year
Enterprise AV Design and Collaboration Engineer
Nvidia · Santa Clara, United States
USD 144,000-230,000 per year
Senior Staff Site Reliability Operations
Nvidia · Seattle, United States
USD 184,000-264,500 per year
Senior Systems Software Engineer, Observability and Telemetry Platform
Nvidia · Santa Clara, United States
USD 152,000-241,500 per year
Similar jobs
Senior Software Engineer - Manufacturing and Factory
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Software Engineer, DGX Cloud AI Infrastructure
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Director, Rack-Scale Software Architecture
Nvidia · Santa Clara, United States
USD 320,000-488,800 per year
Distinguished Engineer - Rack Scale Architecture
Nvidia · Santa Clara, United States
USD 320,000-488,800 per year
Senior Software Architect - Deep Learning and HPC Communications
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior System Software Engineer - GPU Performance
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Performance Modeling Lead
OpenAI · San Francisco, United States, Seattle, United States
USD 293,000-385,000 per year
Principal Software Engineer, Rack-Scale System Software — CSP Engagements
Nvidia · Santa Clara, United States
USD 272,000-431,200 per year