Distinguished Software Engineer - NVLink Fusion Software

at Nvidia
USD 320,000-488,800 per year
SENIOR
✅ On-site

Tech Stack

AI Communication @ 6 GPU HPC InfiniBand NVLink Networking @ 4 System Architecture @ 4

Details

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. Today, NVIDIA is using AI to define the next era of computing, with GPUs powering computers, robots, and self-driving cars.

NVIDIA has a rapidly expanding ecosystem of data center platforms and node designs, ranging from single-node HGX/DGX systems to large multi-node NVLink domain rack architectures. These systems combine NVIDIA GPUs, NVIDIA NVLink, NVIDIA InfiniBand networking, NVIDIA Grace CPUs, and an optimized NVIDIA AI and HPC software stack.

NVIDIA NVLink Fusion will enable AI scale-up and scale-out performance using NVIDIA technology together with semi-custom ASICs or CPUs. The role will champion collaboration across NVIDIA's software, architecture, networking, and systems engineering teams to define the NVLink Fusion architecture and ensure seamless integration of partner ASICs and CPUs into NVIDIA's rack-scale architecture.

Responsibilities

  • Define the NVLink Fusion architecture, leveraging NVIDIA scale-up and scale-out technologies as a foundation.
  • Establish software abstraction layers and reference software for NVLink Fusion partners to extend NVIDIA's rack-scale architecture.
  • Work directly with major customers to understand their requirements and align their roadmaps with NVIDIA's roadmap.
  • Work with business partners and vendors to shape their products to meet NVIDIA's needs.
  • Mentor architects and engineering teams to develop future leaders.
  • Make key technical decisions in ambiguous situations and mitigate execution risks by applying a left-shift strategy to accelerate time to market.

Requirements

  • Bachelor's or master's degree in Computer Engineering, Computer Science, or a related field, or equivalent experience.
  • 16 or more years of experience in system architecture and design.
  • Deep experience designing architectures for scalable and performant server systems, particularly at the software/hardware interface.
  • Experience working with complex system software for accelerators such as GPUs, DPUs, or FPGAs.
  • Expertise in out-of-band and in-band management architectures.
  • Knowledge of device management protocols such as MCTP, PLDM, and RDE.
  • Knowledge of system management protocols such as Redfish and IPMI.
  • Demonstrable experience implementing a left-shift strategy to de-risk program execution.
  • Excellent written and verbal communication skills.

Preferred Qualifications

  • Knowledge of cloud and cluster-level deployment and management systems.
  • Participation in and contributions to standards bodies such as OCP and DMTF.
  • Familiarity with CXL, UCIE, and other C2C technology architectures.
  • Knowledge of storage and networking technologies.

Benefits

  • Equity and benefits are provided.
  • NVIDIA is committed to fostering an inclusive work environment and is an equal opportunity employer.
  • Applications will be accepted at least until July 30, 2026.
  • This posting is for an existing vacancy.
  • NVIDIA uses AI tools in its recruiting processes.

More jobs at Nvidia

Similar jobs