Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI
API
BGP @ 4
Data Structures
GPU
Go @ 7
HPC
Kubernetes @ 3
Networking @ 6
Python @ 7
gRPC @ 3
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
NVIDIA is seeking an experienced software engineer with infrastructure experience to join the Cloud Foundations Automation Development Team. The team builds and manages the automation ecosystem supporting NVIDIA's GPU Cloud and NVIDIA SuperPod deployments.
Responsibilities
- Develop software for efficient network design, deployment, and day-two management.
- Build product-focused software solutions for internal and external customers.
- Help transform workflows and organizational processes into a centrally orchestrated configuration management framework operating at scale across geographies.
- Own and drive integrations with service APIs, including cloud service providers, to automate environment creation and populate data sources.
- Build on open-source software and design and implement data structures and user interfaces to automate processes from equipment purchase through device configuration, deployment, and operations.
- Streamline deployment mechanisms and lifecycle operations.
- Develop modern service architectures around streaming data and event pipelines.
- Work with infrastructure domain experts on zero-touch deployment solutions and use high-performance computing management solutions.
- Identify opportunities to improve services and customer experience while collaborating across the organization.
Requirements
- Bachelor's degree or equivalent experience, with 12 or more years of relevant industry experience.
- Background in networking and network automation, including datacenters, points of presence, and routing.
- Strong proficiency in Python and Go web frameworks.
- Experience building and shipping production-quality software products.
- Experience with DCIM tools such as NetBox or Nautobot.
- Expertise designing and implementing network configuration management systems.
- Familiarity with containerization using Kubernetes, including on-premises Kubernetes and EKS.
- Familiarity with streaming telemetry protocols such as gRPC and gNMI.
Preferred Experience
- Architected, built, and deployed large-scale networks supporting thousands of machines and hundreds of engineers.
- Hands-on experience building tooling and automation for provisioning, monitoring, and managing network infrastructure.
- Understanding of VRFs, VXLAN, EVPN, BGP, OSPF, ISIS, and datacenter design.
Compensation and Benefits
The base salary range is USD 200,000–322,000 for Level 5 and USD 248,000–391,000 for Level 6. Compensation is determined by location, experience, and the pay of employees in similar positions. The role also includes eligibility for equity and benefits.
Applications will be accepted at least until August 3, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes and is an equal opportunity employer.
More jobs at Nvidia
Senior Software Engineer, Compute Sanitizer - GPU
Nvidia · Canada
CAD 170,000-275,000 per year
Senior Full Stack Software Engineer - DGX Cloud
Nvidia · United States
USD 184,000-356,500 per year
Senior Customer Program Manager – AI Platform
Nvidia · Santa Clara, United States
USD 200,000-322,000 per year
Senior System Software Engineer – Data Center Compute Diagnostics
Nvidia · Durham, United States
USD 224,000-356,500 per year
Engineering Manager, AI Compiler Analysis
Nvidia · Santa Clara, United States
USD 168,000-322,000 per year
Similar jobs
Principal Network Automation Engineer
Nvidia · Santa Clara, United States
USD 248,000-396,800 per year
Forward Deployed Engineer - Physical AI Cloud Platform
Nebius · United States, Austin, United States
USD 179,500-224,300 per year
Senior Storage Production Engineer - DGX Cloud
Nvidia · Santa Clara, United States
USD 176,000-333,500 per year
Senior Storage Production Engineer - DGX Cloud
Nvidia · Santa Clara, United States
USD 176,000-333,500 per year
Senior Staff+ Software Engineer, Kubernetes Platform
Anthropic · San Francisco, United States, New York City, United States, Seattle, United States
USD 405,000-485,000 per year
Member of Technical Staff (AI Infrastructure Engineer)
Perplexity AI · San Francisco, United States, Palo Alto, United States
USD 220,000-405,000 per year
Senior Customer Success Engineer - DGX Cloud
Nvidia · Santa Clara, United States
USD 200,000-322,000 per year
Senior Storage Software Engineer - DGX Cloud
Nvidia · Santa Clara, United States
USD 224,000-431,200 per year