Senior Software Engineer - Server Manageability

at Nvidia
USD 152,000-287,500 per year
SENIOR
✅ Remote

Tech Stack

AI Bash @ 4 Communication @ 7 Debugging @ 4 GPU Linux @ 4 Python @ 4 Security @ 6

Details

NVIDIA is seeking a Senior Software Engineer to design and deliver server manageability solutions for GPU-based AI servers. The role focuses on out-of-band management, BMC firmware development, server architecture, enterprise systems, and end-to-end platform delivery.

Responsibilities

  • Design, implement, and deliver innovations for managing GPU-based AI servers, with a focus on out-of-band management, firmware development, server architecture, and enterprise systems.
  • Lead BMC firmware design with a global team of engineers.
  • Design and develop performance-optimized active-monitoring BMC solutions using DMTF standards, including MCTP, Redfish, SPDM, and PLDM.
  • Instrument code to ensure maximum code coverage, write and automate unit tests for each implemented module, and maintain detailed unit test reports.
  • Provide software quality reports based on static analysis, code coverage, and CPU load.
  • Work with security teams to align developed code with product security goals.
  • Collaborate closely with hardware teams to influence hardware design and review hardware architecture and schematics.
  • Drive the definition and end-to-end delivery of platforms by collaborating with internal teams, ODMs, OEMs, and industry partners for AI servers.
  • Work with QA and test architects to develop test tools and automation for qualifying the complete system software and firmware stack.

Requirements

  • Domain expertise in BMC firmware development on x86 or ARM platforms, including BMC-BIOS communication, thermal management, power management, firmware updates, device monitoring, and firmware security.
  • Experience delivering high-end enterprise servers end to end, from definition through customer deployment.
  • Solid understanding of low-level interfaces between SBIOS, BMC, and operating systems, including I2C, SPI, PCIe, and JTAG.
  • Experience with PCIe enumeration and platform-level I/O for enterprise systems.
  • Experience working with hardware teams, ODMs, and vendors to introduce and support server platforms.
  • Experience with C/C++ development, Bash or Python scripting, and debugging in embedded Linux operating environments.
  • Excellent written and oral communication skills, strong teamwork, high-quality work standards, and the ability to work independently and find creative solutions.
  • Bachelor's, master's, or doctoral degree in Electrical Engineering, Computer Science, or an equivalent qualification.
  • At least 5 years of experience with demonstrated individual-contributor capabilities.

Preferred Qualifications

  • Contribution to industry standards or projects such as Open Compute, IPMI, DMTF standards, or OpenBMC open source.
  • Proven experience delivering BMC solutions for enterprise servers using the OpenBMC firmware stack.

Compensation and Benefits

  • Base salary range of USD 152,000–241,500 for Level 3.
  • Base salary range of USD 184,000–287,500 for Level 4.
  • Additional eligibility for equity and benefits.

More jobs at Nvidia

Similar jobs