Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI
Communication @ 7
Debugging @ 4
Deep Learning
Distributed Systems @ 7
GPU @ 6
Go
Kubernetes @ 7
Machine Learning
Software Development @ 6
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
NVIDIA is seeking a Senior Software Engineer to develop distributed storage services for AI/ML. The role focuses on designing and building reliable, scalable, and efficient storage-as-a-service solutions tailored to AI applications. These services must be deployable anywhere and scale without limitations, supporting NVIDIA's business across graphics drivers, autonomous vehicles, and deep learning frameworks.
Responsibilities
- Lead the overall architecture and design of a distributed storage service optimized for AI/ML.
- Develop and maintain robust, scalable distributed Go programs deployed to open-source ecosystems, including Kubernetes.
- Develop and maintain user-space applications, containers, Go bindings, and CLI tools.
- Build features that improve availability and reliability for large-scale distributed storage deployments.
- Collaborate with NVIDIA Research, Computing, Product teams, cross-functional teams, and external customers to deliver cloud services.
- Automate the distributed storage service end to end, including deployment, management, and monitoring.
Requirements
- Bachelor's degree in Computer Science or a related field, or equivalent experience.
- At least 8 years of industry experience.
- Strong background developing distributed systems with Golang, Kubernetes, and cloud service provider integrations.
- Proven track record delivering distributed services in a variety of distributed computing environments.
- Experience implementing storage services and interfaces that provide scalable, high-performance, and reliable solutions.
- Experience owning product delivery from inception through support.
- Experience developing and maintaining enterprise software.
- Experience deploying, managing, and debugging applications in Kubernetes environments.
- Strong communication and presentation skills.
Preferred Qualifications
- Experience architecting, building, and deploying distributed services running on large-scale clusters ranging from multi-petabyte to exabyte scale and supporting millions of users.
- Ownership of all software development and delivery lifecycle stages.
- Interest in accelerated computing environments and technologies such as GPU Direct Storage, DPU, and RDMA.
- Experience building and delivering cloud services, particularly distributed systems.
Benefits
- Equity and benefits are available.
- NVIDIA is committed to fostering an inclusive work environment and is an equal opportunity employer.
- Applications will be accepted at least until August 30, 2026.
- NVIDIA uses AI tools in its recruiting processes.
More jobs at Nvidia
Senior Systems Software Engineer, Low Latency Streaming Technology - Automotive
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Deep Reinforcement Learning Engineer - Autonomous Driving
Nvidia · Santa Clara, United States
USD 224,000-356,500 per year
Senior Manager, Storage Engineering
Nvidia · Santa Clara, United States
USD 248,000-396,800 per year
Senior Software and System Architect
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior Customer Technical Program Manager
Nvidia · Santa Clara, United States
USD 200,000-322,000 per year
Similar jobs
Senior Software Engineer, Cloud-Native Stack – CSP Engagements
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior AI Infrastructure Software Engineer - DGX Cloud
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Full-Stack Lead Engineer
Nvidia · Santa Clara, United States
USD 224,000-356,500 per year
Senior Software Engineer, AI Inference Systems
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Software Engineer, AI Inference Systems
Nvidia · Germany
PLN 292,500-650,000 per year
Senior Software Engineer, AI Inference Systems
Nvidia · Toronto, Canada
CAD 170,000-275,000 per year
Senior Software Engineer, DGX Cloud Orchestration
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
NVIDIA 2027 Internships: Software Engineering
Nvidia · Santa Clara, United States
USD 20-71 per hour