Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI
Bash @ 4
Communication @ 7
Debugging @ 4
GPU
Linux @ 4
Python @ 4
Security @ 6
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
NVIDIA is seeking a Senior Software Engineer to design and deliver server manageability solutions for GPU-based AI servers. The role focuses on out-of-band management, BMC firmware development, server architecture, enterprise systems, and end-to-end platform delivery.
Responsibilities
- Design, implement, and deliver innovations for managing GPU-based AI servers, with a focus on out-of-band management, firmware development, server architecture, and enterprise systems.
- Lead BMC firmware design with a global team of engineers.
- Design and develop performance-optimized active-monitoring BMC solutions using DMTF standards, including MCTP, Redfish, SPDM, and PLDM.
- Instrument code to ensure maximum code coverage, write and automate unit tests for each implemented module, and maintain detailed unit test reports.
- Provide software quality reports based on static analysis, code coverage, and CPU load.
- Work with security teams to align developed code with product security goals.
- Collaborate closely with hardware teams to influence hardware design and review hardware architecture and schematics.
- Drive the definition and end-to-end delivery of platforms by collaborating with internal teams, ODMs, OEMs, and industry partners for AI servers.
- Work with QA and test architects to develop test tools and automation for qualifying the complete system software and firmware stack.
Requirements
- Domain expertise in BMC firmware development on x86 or ARM platforms, including BMC-BIOS communication, thermal management, power management, firmware updates, device monitoring, and firmware security.
- Experience delivering high-end enterprise servers end to end, from definition through customer deployment.
- Solid understanding of low-level interfaces between SBIOS, BMC, and operating systems, including I2C, SPI, PCIe, and JTAG.
- Experience with PCIe enumeration and platform-level I/O for enterprise systems.
- Experience working with hardware teams, ODMs, and vendors to introduce and support server platforms.
- Experience with C/C++ development, Bash or Python scripting, and debugging in embedded Linux operating environments.
- Excellent written and oral communication skills, strong teamwork, high-quality work standards, and the ability to work independently and find creative solutions.
- Bachelor's, master's, or doctoral degree in Electrical Engineering, Computer Science, or an equivalent qualification.
- At least 5 years of experience with demonstrated individual-contributor capabilities.
Preferred Qualifications
- Contribution to industry standards or projects such as Open Compute, IPMI, DMTF standards, or OpenBMC open source.
- Proven experience delivering BMC solutions for enterprise servers using the OpenBMC firmware stack.
Compensation and Benefits
- Base salary range of USD 152,000–241,500 for Level 3.
- Base salary range of USD 184,000–287,500 for Level 4.
- Additional eligibility for equity and benefits.
More jobs at Nvidia
Engineering Manager, Data Labeling Platform
Nvidia · Santa Clara, United States
USD 200,000-391,000 per year
Engineering Manager, Local AI Agents
Nvidia · Santa Clara, United States
USD 224,000-431,200 per year
Senior Deep Learning Software Engineer, DLSim
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior Software Engineer, Fleet Intelligence Backend
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior Staff Business Systems Analyst
Nvidia · Santa Clara, United States
USD 144,000-270,200 per year
Similar jobs
Senior Storage Production Engineer - DGX Cloud
Nvidia · Santa Clara, United States
USD 176,000-333,500 per year
Senior Storage Production Engineer - DGX Cloud
Nvidia · Santa Clara, United States
USD 176,000-333,500 per year
Platform Security Engineer, OpenBMC
Anthropic · New York City, United States, San Francisco, United States, Seattle, United States
USD 320,000-405,000 per year
Senior Embedded System Software Engineer – Platform Execution Lead
Nvidia · Santa Clara, United States
USD 224,000-356,500 per year
Senior Systems Software Engineer, Data Center Platform Enablement
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Systems Software Engineer, Kubernetes Node Lifecycle - DGX Cloud
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Software Engineer, DGX Cloud Production Engineering
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Distinguished Engineer, Storage – AI Cloud
Nvidia · Santa Clara, United States
USD 320,000-488,800 per year