Senior Systems Software Engineer, Data Center Platform Enablement
at Nvidia
USD 184,000-356,500 per year
Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 4
Bash @ 4
Communication @ 6
Debugging @ 4
GPU
HPC
Leadership @ 6
Linux @ 4
NVLink
Networking
Python @ 4
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
NVIDIA is seeking a Senior Systems Software Engineer to work on bring-up, integration, validation, and troubleshooting for compute tray platforms in GPU racks. The role focuses on ensuring servers are fully functional and validated before mass deployment in data centers. NVIDIA data center systems combine GPUs, NVLink, networking, data center CPUs, and an optimized AI and HPC software stack.
Responsibilities
- Collaborate with global architects and developers on NVIDIA AI server system designs, including CPU, GPU, memory, NIC, PCIe, NVMe SSD, and cooling components.
- Work closely with hardware teams to influence system design and review schematics and board designs.
- Use simulation and emulation to shift design validation earlier and develop new concepts.
- Debug and support early server prototype bring-up and power-on testing in data center system labs.
- Support manufacturing flows, firmware updates, and diagnostic procedures.
- Ensure BOM change signoff and optimize manufacturing processes.
- Drive root-cause analysis and resolution of bring-up failures.
- Collaborate with partners, ODMs, and customers to provide technical support.
- Work with design architects to develop diagnostic test tools and automation for qualifying early system software and firmware stacks.
Requirements
- Strong experience with computer system and firmware architecture and design for server products.
- Experience with CPLD or FPGA design and RTL.
- Experience delivering high-end enterprise server products end to end, from definition through customer deployment.
- Solid understanding of low-level interfaces between SBIOS, BMC, and operating systems, including I2C, SPI, PCIe, and JTAG.
- Knowledge of system boot and initialization, PCIe enumeration, and high-speed I/O at the enterprise platform level.
- Experience working with hardware teams, ODMs, and vendors to introduce and support server platforms.
- Experience with C/C++ development, Bash and Python scripting, and hands-on debugging in embedded Linux environments.
- Experience using QEMU and emulation tools to validate early design work.
- Experience using AI tools to accelerate design and debugging.
- Excellent written and oral communication, teamwork, work ethic, and commitment to quality.
- Self-starter with creative problem-solving skills.
- Bachelor's degree or higher in Electrical Engineering or Computer Science, or equivalent experience.
- 8+ years of experience with demonstrated ability as an individual contributor.
Preferred Qualifications
- Experience with early design and power-on debugging.
- Proven leadership and ability to address challenging issues in a fast-paced environment.
Compensation And Benefits
- Level 4 base salary: USD 184,000–287,500 per year.
- Level 5 base salary: USD 224,000–356,500 per year.
- Eligible employees may also receive equity and benefits.
More jobs at Nvidia
Senior System Software Engineer, Platform - OpenBMC
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Software QA Test Development Engineer - Diagnostics
Nvidia · Santa Clara, United States
USD 140,000-270,200 per year
Senior Security Engineer, RTOS and Virtualization
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Software Engineer - NVIDIA Warp
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Software And System Architect
Nvidia · Santa Clara, United States
USD 124,000-241,500 per year
Similar jobs
Senior System Software Engineer – Data Center Compute Diagnostics
Nvidia · Durham, United States
USD 224,000-356,500 per year
Senior Storage Production Engineer - DGX Cloud
Nvidia · Santa Clara, United States
USD 176,000-333,500 per year
Senior Storage Production Engineer - DGX Cloud
Nvidia · Santa Clara, United States
USD 176,000-333,500 per year
Senior Software Engineer, DGX Cloud AI Infrastructure
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Software Engineer - NVLink Rack Scale Stability and Reliability
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior HPC Cluster Engineer
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior Solution Engineer, Networking
Nvidia · Santa Clara, United States
USD 168,000-322,000 per year
System Software Engineer – Data Center Compute Diagnostics
Nvidia · Durham, United States
USD 152,000-241,500 per year