Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 4
API @ 4
Agentic AI
CUDA @ 4
Communication @ 6
Debugging @ 4
GPU @ 6
GenAI @ 4
LLM @ 4
Linux @ 3
OSS @ 4
Profiling @ 6
Python @ 7
Technical Leadership
TensorRT @ 4
vLLM
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
At NVIDIA, we’re tapping into the unlimited potential of AI to define the next era of computing. As a Developer Technology Engineer, you will be at the forefront of innovation, working with leading industry partners and pioneering open-source projects to enable professional agentic AI workflows at the edge powered by NVIDIA's RTX and DGX platforms.
Responsibilities
- Work closely across internal engineering and product teams as well as external app developers and enterprise ISVs on solving local end-to-end agentic AI GPU deployment challenges on NVIDIA RTX & DGX.
- Apply powerful profiling and debugging tools for analyzing most demanding accelerated end-to-end agentic AI workflows to detect insufficient system utilization resulting in suboptimal runtime performance.
- Conduct hands-on trainings, develop sample code and host presentations to give good guidance on efficient end-to-end agentic AI deployment targeting optimal runtime performance.
- Improve LLM & GenAI user experience by working on feature and performance enhancements of OSS software, including but not limited to projects like GGML, Llama.cpp, Ollama, vLLM, ONNX Runtime.
- Collaborate with GPU driver and architecture teams as well as NVIDIA research to influence next generation GPU features by providing real-world workflows and giving feedback on partner and customer needs.
- Provide technical leadership and mentorship to junior engineers, encouraging an inclusive and high-performing team environment.
Requirements
- A proven track record 5+ years of professional experience in local GPU deployment, profiling and optimization.
- A Bachelor’s or Master’s degree or equivalent experience in Computer Science, Engineering, or a related field.
- Strong proficiency in C/C++, Python, software design, programming techniques.
- Familiarity with and development experience on Windows and Linux.
- Experience with CUDA and NVIDIA's Nsight GPU profiling and debugging suite.
- Some travel is required for conferences and for on-site visits with external partners.
- Strong problem-solving skills and the ability to work both independently and collaboratively in a fast-paced environment.
- Excellent interpersonal and communication skills and a passion for keeping track with the latest advancements in AI technology.
Ways to stand out from the crowd
- Experience with GPU-accelerated AI inference driven by NVIDIA APIs and SDKs, specifically TensorRT-RTX, cuDNN, NVIDIA Model Optimizer.
- Expertise with professional agentic AI use cases, i.e., digital content creation and productivity workflows.
- Experience working with open-source LLM and GenAI software.
- Detailed knowledge of the latest generation GPU architectures.
- Experience with AI deployment on NPUs and ARM architectures.
Benefits
- Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package.
- You will also be eligible for equity and benefits.
More jobs at Nvidia
Senior System Software Engineer - Halos Core And Robotics Platform
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Director, Autonomous Vehicles Platform
Nvidia · Santa Clara, United States
USD 320,000-488,800 per year
Senior System Software Engineer - Halos
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Engineering Manager, Drive Os Communication Infrastructure
Nvidia · Santa Clara, United States
USD 224,000-356,500 per year
ECAD PCB Tools Developer
Nvidia · Santa Clara, United States
USD 168,000-310,500 per year
Similar jobs
Principal ML Solutions Architect - Token Factory
Nebius · United States
USD 208,000-261,000 per year
Senior Software Engineer, DGX Cloud AI Infrastructure
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Developer Technology Engineer - Windows Ai Platform
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Software Engineer - Autonomous Driving
Nvidia · Santa Clara, United States
USD 224,000-356,500 per year
AI Inference Performance Engineer - New College Grad 2026
Nvidia · Santa Clara, United States
USD 124,000-241,500 per year
Ai Inference Performance Engineer
Nvidia · Santa Clara, United States
USD 152,000-241,500 per year
Systems Generalist, GPT Infrastructure
OpenAI · San Francisco, United States, Seattle, United States
USD 293,000-445,000 per year
System Software Engineer, Dynamo-Triton Inference Server - New College Grad 2026
Nvidia · Santa Clara, United States
USD 124,000-241,500 per year