Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 7
API @ 6
Communication @ 6
DevOps @ 4
Docker @ 4
GPU
Go
Java @ 7
Kubernetes @ 4
Leadership @ 6
Linux @ 4
Mentoring @ 6
Networking @ 4
Rust @ 7
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
NVIDIA is seeking a seasoned leader to manage a software engineering team within the DGX Cloud Kubernetes (DGXC-K8s) organization, supporting internal and external clients. The team is building a GPU cloud using NVIDIA technologies on top of Kubernetes and operating a platform focused on emerging AI needs.
Responsibilities
- Partner with multiple internal teams to gather requirements, provide feedback to engineering teams, and develop solutions to support their success.
- Develop, guide, and supervise a team of engineers with varying talents across organizations.
- Educate partner teams on standard methodologies for cloud-based solutions.
- Build a deep understanding of how NVIDIA develops across multiple verticals, including autonomous vehicles, chip manufacturing, hardware design, and AI.
- Present roadmaps, vision, and demonstrations to internal partners and NVIDIA leadership.
- Participate in open-source communities related to software used and built by the team.
Requirements
- Bachelor's or master's degree in Computer Science or a related field, or equivalent experience.
- At least 10 years of experience designing and building distributed software systems.
- At least 4 years of experience leading teams of engineers across organizations.
- Strong ability to write code in mainstream systems programming languages such as Golang, Java, C, C++, or Rust.
- Demonstrated ability to design and implement maintainable APIs for consumers.
- Experience with Kubernetes and/or distributed task scheduling.
- Understanding of Linux kernel scheduling, memory management, and networking subsystems.
- Understanding of infrastructure, networking, storage, and DevOps scripting and tooling.
- Experience with container orchestration technologies such as containerd, CRI-O, or Docker.
- Familiarity with Identity and Access Management approaches.
- Ability to reach cross-functional consensus while dealing with ambiguity.
Preferred Qualifications
- Experience at a hyperscale cloud service provider, whether public-facing or internal.
- Experience measuring and improving operational excellence.
- Persuasive and effective written and verbal presentation skills.
- Leadership, communication, mentoring, analysis, problem-solving, and short- and long-term planning skills.
- Excellent interpersonal and organizational skills, with the ability to handle diverse situations.
Compensation and Benefits
The base salary depends on location, experience, and the pay of employees in similar positions. The base salary ranges are $224,000-$356,500 USD for Level 3 and $272,000-$431,250 USD for Level 4. Employees are also eligible for equity and benefits.
Applications will be accepted at least until September 6, 2026. This posting is for an existing vacancy. NVIDIA is an equal opportunity employer.
More jobs at Nvidia
Senior NPI Program Manager
Nvidia · Santa Clara, United States
USD 168,000-258,800 per year
GPU PCIe and Boot Architect - New College Grad 2026
Nvidia · Santa Clara, United States
USD 124,000-241,500 per year
Senior AI Engineer, High Performance AI
Nvidia · Santa Clara, United States
USD 152,000-241,500 per year
Senior Salesforce CPQ Developer
Nvidia · Santa Clara, United States
USD 176,000-276,000 per year
Senior Technical Program Manager - LLM Safety
Nvidia · Santa Clara, United States
USD 168,000-322,000 per year
Similar jobs
Senior Software Engineer
SentinelOne · United States
USD 132,000-182,000 per year
Staff+ Software Engineer, Infrastructure (Distributed Systems)
Anthropic · New York City, United States, San Francisco, United States, Seattle, United States
USD 320,000-485,000 per year
Senior Staff+ Software Engineer, Kubernetes Platform
Anthropic · New York City, United States, San Francisco, United States, Seattle, United States
USD 405,000-485,000 per year
Principal Software Engineer – Infrastructure
Nvidia · Santa Clara, United States
USD 248,000-391,000 per year
Senior Software Engineer, Fleet Intelligence Agent Systems
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Principal Software Engineer, DGX Cloud Production Engineering
Nvidia · Santa Clara, United States
USD 272,000-431,200 per year
NCX Senior Engineer
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Staff Software Engineer, Identity & Access Management
Reddit · United States
USD 217,000-303,900 per year