Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI
Communication @ 7
Debugging @ 4
GPU @ 4
Python @ 7
System Architecture @ 4
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
NVIDIA is looking for a highly motivated, creative Factory System Software and Diagnostics Integration Engineer to join the Datacenter Platform Software team. This role will coordinate factory projects, guide cross-functional teams, and collaborate with stakeholders to deliver firmware, software, and diagnostics aligned with business objectives. The position supports NVIDIA GPU- or DPU-based products and includes analyzing, evaluating, and improving factory processes.
Responsibilities
- Lead factory validation projects from software, firmware, and diagnostics perspectives, ensuring successful implementation, integration, and standardization across multiple locations.
- Lead the integration and handover of software, firmware, and diagnostics for all products.
- Act as a primary consultant, providing strategic direction and leadership to engineering teams.
- Engage with ODMs and cross-functional teams to collect and analyze factory requirements and deliver solutions that meet business and operational needs.
- Analyze factory processes, systems, and workflows to identify opportunities for improvement and optimization.
- Develop and maintain detailed documentation, including system requirements, processes, and standard operating procedures (SOPs).
- Collaborate across teams to design, configure, and implement factory changes with seamless integration and minimal disruption.
- Perform system testing, validation, and fixes to ensure functionality and alignment with factory requirements.
- Monitor, collect, and analyze data to identify and address issues and propose solutions to improve system reliability and efficiency.
- Collaborate with vendors and external partners to evaluate, select, and implement new factory systems or software and firmware upgrades.
Requirements
- At least 5 years of relevant experience.
- BS, MS, or PhD in Electrical Engineering, Computer Science, or a related field, or equivalent experience.
- 2–5 years of experience working as an Integration Engineer in factory settings.
- Strong knowledge of server manageability, bring-up, and deployment in data centers.
- Proven understanding of firmware and diagnostics for x86 and ARM servers.
- Experience working with ODMs and OEMs to deliver quality servers.
- Strong, demonstrable programming skills in Python, C/C++, and shell scripting.
- Experience programming and debugging GPU platforms.
- Excellent written and oral communication skills, strong work ethic, teamwork, commitment to quality, and the ability to complete tasks consistently.
- Self-starter with hands-on coding experience and the ability to find creative solutions to complex problems.
Preferred Qualifications
- Experience with factory integration or hardware bring-up projects.
- Hands-on experience with x86 or ARM system architecture.
Compensation and Benefits
- Base salary range: USD 152,000–241,500 for Level 3 and USD 184,000–287,500 for Level 4.
- Eligible for equity and benefits.
- Applications will be accepted at least until August 1, 2026.
- This posting is for an existing vacancy.
- NVIDIA uses AI tools in its recruiting processes and is an equal opportunity employer.
More jobs at Nvidia
Engineering Manager, Data Labeling Platform
Nvidia · Santa Clara, United States
USD 200,000-391,000 per year
Engineering Manager, Local AI Agents
Nvidia · Santa Clara, United States
USD 224,000-431,200 per year
Senior Deep Learning Software Engineer, DLSim
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior Software Engineer, Fleet Intelligence Backend
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior Staff Business Systems Analyst
Nvidia · Santa Clara, United States
USD 144,000-270,200 per year
Similar jobs
Senior Power Architect, Power and Performance Analysis Tools
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Systems Software Engineer - Fleet Debuggability
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Data Center Power Test Architect
Nvidia · Santa Clara, United States
USD 168,000-322,000 per year
Principal Architect, System Software - Orbital Data Center
Nvidia · Santa Clara, United States
USD 272,000-431,200 per year
Principal System Software Engineer - AV Platform
Nvidia · Santa Clara, United States
USD 272,000-431,200 per year
Senior Software Engineer - NVLink Rack Scale Stability and Reliability
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior System Software Engineer - GPU Performance
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior HPC Performance Engineer
Nvidia · Germany
PLN 221,200-507,000 per year