Senior Developer Technology Engineer - Edge Agentic AI

at Nvidia
USD 152,000-287,500 per year
SENIOR
✅ On-site

Tech Stack

AI @ 4 API @ 4 Agentic AI @ 6 CUDA @ 4 Communication @ 6 Debugging @ 4 GPU @ 4 GenAI @ 4 LLM @ 4 Linux @ 3 Profiling @ 6 Python @ 7 Technical Leadership TensorRT @ 4 vLLM

Details

At NVIDIA, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work.

As a Developer Technology Engineer, you will be at the forefront of innovation, working with leading industry partners and pioneering open-source projects to enable professional agentic AI workflows at the edge powered by NVIDIA’s RTX and DGX platforms. This role offers an opportunity to collaborate with world-class talent and contribute to the evolving landscape of enterprise and consumer agentic AI.

Responsibilities

  • Work closely with internal engineering and product teams, external app developers, and enterprise ISVs to solve local, end-to-end agentic AI GPU deployment challenges on NVIDIA RTX and DGX.
  • Apply profiling and debugging tools to analyze demanding, accelerated end-to-end agentic AI workflows and detect insufficient system utilization that results in suboptimal runtime performance.
  • Conduct hands-on training, develop sample code, and host presentations to provide guidance on efficient end-to-end agentic AI deployment and optimal runtime performance.
  • Improve LLM and GenAI user experience through feature and performance enhancements to open-source software, including GGML, Llama.cpp, Ollama, vLLM, and ONNX Runtime.
  • Collaborate with GPU driver and architecture teams and NVIDIA Research to influence next-generation GPU features by providing real-world workflow insights and feedback on partner and customer needs.
  • Provide technical leadership and mentorship to junior engineers while encouraging an inclusive, high-performing team environment.

Requirements

  • At least 5 years of professional experience in local GPU deployment, profiling, and optimization.
  • Bachelor’s or master’s degree, or equivalent experience, in Computer Science, Engineering, or a related field.
  • Strong proficiency in C/C++, Python, software design, and programming techniques.
  • Familiarity with and development experience on Windows and Linux.
  • Experience with CUDA and NVIDIA’s Nsight GPU profiling and debugging suite.
  • Strong problem-solving skills and the ability to work independently and collaboratively in a fast-paced environment.
  • Excellent interpersonal and communication skills and a passion for keeping up with the latest advancements in AI technology.
  • Some travel is required for conferences and on-site visits with external partners.

Preferred Qualifications

  • Experience with GPU-accelerated AI inference driven by NVIDIA APIs and SDKs, specifically TensorRT-RTX, cuDNN, and NVIDIA Model Optimizer.
  • Expertise with professional agentic AI use cases, such as digital content creation and productivity workflows.
  • Experience working with open-source LLM and GenAI software.
  • Detailed knowledge of the latest-generation GPU architectures.
  • Experience with AI deployment on NPUs and ARM architectures.

Benefits

NVIDIA offers equity, competitive salaries, and a comprehensive benefits package. NVIDIA is committed to fostering an inclusive work environment and is an equal opportunity employer.

Applications for this job will be accepted at least until July 26, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.

More jobs at Nvidia

Similar jobs