Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI
Bash @ 4
Communication @ 6
Debugging @ 4
GPU
Linux @ 4
Python @ 4
Security @ 6
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
NVIDIA is looking for a Senior Software Engineer - Server Manageability to design, implement, and deliver innovations for managing GPU-based AI servers, with a focus on OOB management, firmware development, server architecture, and building systems for enterprise.
Responsibilities
- Designing, implementing, and delivering innovations for managing GPU based AI servers with focus on OOB management, firmware development, server architecture and building systems for enterprise.
- Leading BMC firmware design with a global team of engineers.
- Designing and developing performance optimized active monitoring BMC solutions using DMTF Standards including MCTP, Redfish, SPDM and PLDM specifications.
- Instrumenting code to ensure maximum code coverage, writing and automating unit tests for each implemented module and maintain detailed unit test case reports.
- Providing software quality reports based on static analysis, code coverage, CPU load.
- Working with security team to ensure developed code is in line with product security goals. Working closely with hardware teams to influence hardware design and review HW architecture & schematics.
- Driving definition and end to end delivery of all platforms by collaborating with internal teams, ODMs/OEMs and industry partners for AI servers.
- Working with QA/Test architects to come up with proper test tools and automation for qualifying the whole system software and firmware stack.
Requirements
- Domain expertise in BMC Firmware development on X86 or ARM Platforms including BMC-BIOS communication, thermal management, power management, firmware update, device monitoring, firmware security, etc.
- Solid experience of end-to-end delivery of high-end enterprise servers from definition to customer deployment.
- Solid understanding of low-level interfaces between SBIOS, BMC and OS like I2C/SPI/PCIe/JTAG etc. PCIe enumeration, IO at platform level for enterprise systems.
- Experience working closely with HW teams, ODMs and vendors to introduce and support server platforms.
- Experience with C/C++ development, bash/python for scripting, and debugging skills in embedded Linux operating environments.
- Excellent written and oral communication skills, good work ethics, high sense of team-work, love to produce quality work and commitment to finish tasks every single day. A self-starter who loves to find creative solutions to exciting problems.
- Bachelor’s degree, Master’s Degree, or a PhD; in Electrical Engineering or Computer Science (or equivalent experience) and 5+ years of experience, with demonstrated strong ability as individual contributor.
Benefits
- Eligible for equity and benefits.
More jobs at Nvidia
Ncx Senior Engineer
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
System Test Engineer
Nvidia · Santa Clara, United States
USD 132,000-253,000 per year
Senior Software Engineer, DGX Cloud Orchestration
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Technical Program Manager, Deep Learning Frameworks
Nvidia · Santa Clara, United States
USD 168,000-322,000 per year
Senior Software Engineer, CUDA Core Libraries
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Similar jobs
Developer Tools DevOps Engineer
Nvidia · Santa Clara, United States
USD 148,000-276,000 per year
Senior System Firmware Engineer - Bios Uefi
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Platform Security Engineer, OpenBMC
Anthropic · San Francisco, United States, New York City, United States, Seattle, United States
USD 320,000-405,000 per year
Senior Systems Software Engineer, Data Center Platform Enablement
Nvidia · Santa Clara, United States
USD 184,000-287,500 per year
Senior Storage Production Engineer - DGX Cloud
Nvidia · Santa Clara, United States
USD 176,000-333,500 per year
Senior Software Engineer, DGX Cloud Production Engineering
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Distinguished Engineer, Storage – AI Cloud
Nvidia · Santa Clara, United States
USD 320,000-488,800 per year
Senior Hpc Cluster Engineer
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year