Senior Software Engineer, Deep Learning Inference – TensorRT

at Nvidia
USD 152,000-287,500 per year
SENIOR
✅ Hybrid

Tech Stack

AI Algorithms @ 4 C @ 7 C++ @ 7 CUDA @ 4 Communication @ 6 Data Structures @ 4 Deep Learning DevOps GPU @ 4 LLM Machine Learning @ 7 OpenCL @ 4 Profiling @ 4 PyTorch @ 4 Python @ 6 Software Development @ 6 TensorFlow @ 4 TensorRT @ 4

Details

Help build a state-of-the-art inference framework for accelerating deep learning models, especially large language models, on NVIDIA GPUs. The role is part of NVIDIA’s TensorRT software team and involves developing high-performance inference software for multiple platforms.

Responsibilities

  • Craft and develop robust inference software that can be scaled across multiple platforms for functionality and performance.
  • Develop components of TensorRT, NVIDIA’s SDK for high-performance deep learning inference.
  • Follow academic developments in artificial intelligence and update TensorRT features accordingly.
  • Use C++ and Python to build graph parsers, optimizers, and tools for effective deployment of trained deep learning models.
  • Collaborate with deep learning experts, GPU architects, and DevOps engineers across diverse teams.

Requirements

  • Bachelor’s, Master’s, PhD, or equivalent experience in Computer Science, Computer Engineering, Electrical Engineering, or a related field.
  • 3+ years of software development experience.
  • Strong experience with modern C++ standards, including C++11, C++14, C++17, and C++20.
  • Strong grasp of machine learning concepts.
  • Experience and knowledge of computer architecture, data structures, and algorithms.
  • Excellent communication skills and an aptitude for collaboration and teamwork.

Preferred Qualifications

  • Experience developing system software.
  • Proficiency in Python.
  • Experience with GPU kernel programming using CUDA or OpenCL.
  • Experience with software performance benchmarking, profiling, and optimization.
  • Background in compiler development.
  • Experience working with TensorRT, PyTorch, TensorFlow, ONNX Runtime, or other machine learning frameworks.

Benefits

  • Eligibility for equity and benefits.
  • NVIDIA is committed to fostering an inclusive work environment and is an equal opportunity employer.

The position is full time. Applications will be accepted at least until June 13, 2026. NVIDIA uses AI tools in its recruiting processes.

More jobs at Nvidia

Similar jobs