Software Devops Engineer, Networking

at Nvidia
USD 148,000-276,000 per year
MIDDLE SENIOR
✅ On-site

Tech Stack

AI Bash @ 3 CI/CD DevOps @ 5 Docker @ 6 GPU @ 3 InfiniBand @ 2 JSON @ 3 Linux @ 6 Networking @ 6 Python @ 3 SRE @ 5 Software Development @ 6 gRPC @ 3

Details

NVIDIA is looking for an outstanding candidate to solve SW integration challenges for next-generation data center platforms. You will be at the heart of the latest GPU architectures and advanced AI infrastructure projects, ensuring the seamless integration of technologies in high-speed communication and virtualization. You will support products that leverage Ethernet and InfiniBand protocols, delivering advanced compute and networking technologies for demanding AI workloads. In this role, you will provide first-tier support to R&D teams, acting as the bridge between hardware and stable software deployments.

Responsibilities

  • Fixing and prioritising complex systems during high-stakes bringups and Proof of Concepts (PoCs) for next-generation computing architectures.
  • Managing the integration of large-scale products involving GPUs, complex Network Stacks, Firmware, and Drivers.
  • Creating, recreating, and redeploying software artifacts, including fixing code, updating builds, or providing workarounds to unblock development.
  • Serving as the primary technical point of contact for R&D teams to resolve immediate infrastructure and integration blockers.
  • Working closely with R&D, Verification, and DevOps teams to streamline the CI/CD pipeline for specialized high-speed interconnect and system management projects.

Requirements

  • Bachelor of Science degree in Computer Science or similar academic degree, or equivalent experience.
  • Proven software engineering background with deep understanding of standard methodologies in software development, modern Linux-based operating systems, and computer networking.
  • 5+ years of overall experience in DevOps, SRE, or Systems Integration roles.
  • Deep knowledge of Linux distributions (Ubuntu/RHEL) and containerization using Docker.
  • Coding skills in C/C++, Python, and Bash for automation and system-level fixes.
  • Experience with GitLab and GitLab CI for managing complex build pipelines.
  • Ability to multi-task, self-manage in a fast-paced environment, and lead technically during critical system failures.
  • Excellent problem-solving and critical thinking abilities.

Ways to stand out from the crowd

  • In-depth knowledge and familiarity with high-performance networking (InfiniBand, Ethernet).
  • Practical experience with gRPC, gNMI, REST, and JSON for system management and telemetry.
  • Experience working on large-scale HW+SW converged systems (e.g., rack-scale computing or GPU clusters).

Compensation

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions.

  • Level 3 base salary range: 148,000 USD - 235,750 USD
  • Level 4 base salary range: 176,000 USD - 276,000 USD

You will also be eligible for equity and benefits.

More jobs at Nvidia

Similar jobs