Software DevOps Engineer, Networking

at Nvidia
USD 148,000-276,000 per year
MIDDLE
✅ On-site

Tech Stack

AI Bash @ 3 CI/CD DevOps @ 5 Docker @ 3 GPU @ 3 InfiniBand @ 3 JSON @ 3 Linux @ 6 Networking @ 6 Python @ 3 SRE @ 5 Software Development @ 6 gRPC @ 3

Details

NVIDIA is seeking an engineer to solve software integration challenges for next-generation data center platforms. The role supports GPU architectures and AI infrastructure projects involving high-speed communication, virtualization, Ethernet, and InfiniBand. You will provide first-tier support to R&D teams and help bridge hardware development with stable software deployments.

Responsibilities

  • Fix and prioritize complex system issues during high-stakes bring-ups and proof-of-concept activities for next-generation computing architectures.
  • Manage the integration of large-scale products involving GPUs, network stacks, firmware, and drivers.
  • Create, recreate, and redeploy software artifacts.
  • Fix code, update builds, and develop workarounds to unblock development.
  • Serve as the primary technical point of contact for R&D teams addressing infrastructure and integration blockers.
  • Work with R&D, verification, and DevOps teams to streamline CI/CD pipelines for specialized high-speed interconnect and system management projects.
  • Lead technically during critical system failures in a fast-paced environment.

Requirements

  • Bachelor's degree in Computer Science or a similar discipline, or equivalent experience.
  • Software engineering experience and a strong understanding of software development methodologies, modern Linux-based operating systems, and computer networking.
  • At least 5 years of experience in DevOps, SRE, or systems integration roles.
  • In-depth knowledge of Linux distributions, including Ubuntu and RHEL.
  • Experience with Docker containerization.
  • Coding skills in C/C++, Python, and Bash for automation and system-level fixes.
  • Experience with GitLab and GitLab CI for managing complex build pipelines.
  • Ability to multitask and self-manage.
  • Excellent problem-solving and critical-thinking skills.

Preferred Qualifications

  • In-depth knowledge of high-performance networking, including InfiniBand and Ethernet.
  • Practical experience with gRPC, gNMI, REST, and JSON for system management and telemetry.
  • Experience working on large-scale hardware and software converged systems, such as rack-scale computing or GPU clusters.

Benefits

The role includes competitive compensation, equity, and benefits. NVIDIA is an equal opportunity employer committed to fostering an inclusive work environment.

More jobs at Nvidia

Similar jobs