Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI
API
AWS @ 3
CI/CD @ 4
CUDA @ 4
Communication @ 7
Debugging
DevOps @ 4
GPU @ 3
IaC
Kubernetes @ 3
Linux @ 3
Microservices
Profiling @ 3
Python @ 6
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
We are seeking a Senior Software Engineer to help build and improve AI developer tools connected through web APIs, SDKs, CLIs, and agents. The role focuses on delivering user-focused cloud products that enable rapid prototyping and highly automated AI-assisted CUDA development, from coding to profiling and performance fine-tuning.
You will architect cloud solutions leveraging NVIDIA microservices and frameworks, as well as AI applications and research, to integrate deeply with developer workflows and accelerate the full product lifecycle of planning, coding, testing, debugging, profiling, and fine-tuning.
Responsibilities
- Work closely with Product and Design teams to define feature specifications and build the next generation of AI-assisted coding and profiling services and skills.
- Architect, design, and develop high-performance, responsive SaaS products that improve developer workflows with AI.
- Support a large number of concurrent developers with high scalability, reliability, and cost efficiency.
- Work with other engineering teams to align on corporate infrastructure strategies and improve or enhance existing services.
- Mentor engineers and review code and design.
Requirements
- Bachelor's degree or equivalent experience in Computer Science.
- 8 years of industry experience.
- Experience developing large-scale, user-facing applications using web and cloud services.
- Familiarity with Kubernetes, Infrastructure as Code, and AWS.
- Experience with DevOps, including CI/CD, monitoring, and alerts.
- Proficiency in Python.
- Strong communication skills for collaboration with cross-functional partners.
- Familiarity with either Linux virtualization for running isolated GPU workloads in the cloud or GPU profiling and benchmarking performance in a consistent, repeatable way.
Preferred Qualifications
- Experience writing CUDA code and optimizing CUDA kernels.
- Experience with Kata Containers or another VM-based container runtime.
Benefits
- Competitive salary and benefits.
- Eligibility for equity and NVIDIA benefits.
- NVIDIA is committed to fostering an inclusive work environment and is an equal opportunity employer.
Applications will be accepted at least until September 5, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.