Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 4
Communication @ 7
Data Modeling @ 3
Distributed Systems @ 7
Go
MongoDB @ 3
Observability
PostgreSQL @ 3
Security @ 4
Software Development @ 4
TypeScript
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
Responsibilities
- Leading cloud VM platform initiatives from architecture and systems design through implementation, testing, and deployment.
- Designing, building, and operating scalable services that support secure, isolated agent workloads.
- Developing automation for infrastructure provisioning, VM lifecycle management, application delivery, configuration, monitoring, and remediation.
- Collaborating with engineering teams across NVIDIA to translate business and AI workflow requirements into dependable platform capabilities.
- Improving the platform’s reliability, scalability, observability, performance, and operational efficiency.
- Continuously raising standards for code quality, testing, documentation, infrastructure security, and production readiness.
Requirements
- Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field—or equivalent experience.
- 8+ years of experience in software, platform, infrastructure, or cloud engineering, with a focus on backend or distributed systems.
- Strong understanding of distributed-systems concepts, including service communication, asynchronous processing, failure handling, retries, idempotency, and eventual consistency.
- Familiarity with PostgreSQL and MongoDB, including data modeling, migrations, and performance.
- Understanding of secure software development, cloud security, identity and access management, secrets management, and network security.
- Experience with AI-native development and coding agents to multiply developer efficiency.
- Strong communication skills and a collaborative, team-oriented approach.
Ways to stand out from the crowd
- Deep expertise in Go and TypeScript.
- Production experience with Temporal for durable execution, workflow orchestration, and long-running distributed processes.
- Comprehensive understanding of infrastructure security, workload isolation, container security, and zero-trust principles.
- Experience building platforms for AI agents or sandbox execution.
- Excited to contribute to a high-trust, collaborative culture while bringing a unique perspective to the team!
Additional information
- Your base salary will be determined based on your location, experience, and the pay of employees in similar positions.
- The base salary range is 168,000 USD - 270,250 USD.
- You will also be eligible for equity and benefits.
More jobs at Nvidia
Ncx Senior Engineer
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
System Test Engineer
Nvidia · Santa Clara, United States
USD 132,000-253,000 per year
Senior Software Engineer, DGX Cloud Orchestration
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Technical Program Manager, Deep Learning Frameworks
Nvidia · Santa Clara, United States
USD 168,000-322,000 per year
Senior Software Engineer, CUDA Core Libraries
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Similar jobs
Staff Forward Deployed Engineer
GitLab · United States
USD 254,000-297,000 per year
Product Engineer, Enterprise AI Platform
OpenAI · San Francisco, United States
USD 230,000-385,000 per year
Staff Backend Engineer - Grafana Enterprise
Grafana Labs · United States
USD 175,000-210,000 per year
Staff Backend Engineer - Grafana Enterprise
Grafana Labs · Canada
CAD 186,400-223,600 per year
Member Of Technical Staff (Software Engineer, Enterprise Experience)
Perplexity AI · San Francisco, United States, New York City, United States
USD 220,000-405,000 per year
Resident Solutions Architect
Glean · United States
USD 170,000-240,000 per year
Backend Software Engineer, API Enterprise Controls
OpenAI · San Francisco, United States
USD 293,000-385,000 per year
Software Engineer, Product Security Data Platforms
Stripe · Seattle, United States
USD 156,800-235,200 per year