Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI
Algorithms @ 4
C @ 7
C++ @ 7
CUDA @ 4
Communication @ 6
Data Structures @ 4
Deep Learning
DevOps
GPU @ 4
LLM
Machine Learning @ 7
OpenCL @ 4
Profiling @ 4
PyTorch @ 4
Python @ 6
Software Development @ 6
TensorFlow @ 4
TensorRT @ 4
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
Help build a state-of-the-art inference framework for accelerating deep learning models, especially large language models, on NVIDIA GPUs. The role is part of NVIDIA’s TensorRT software team and involves developing high-performance inference software for multiple platforms.
Responsibilities
- Craft and develop robust inference software that can be scaled across multiple platforms for functionality and performance.
- Develop components of TensorRT, NVIDIA’s SDK for high-performance deep learning inference.
- Follow academic developments in artificial intelligence and update TensorRT features accordingly.
- Use C++ and Python to build graph parsers, optimizers, and tools for effective deployment of trained deep learning models.
- Collaborate with deep learning experts, GPU architects, and DevOps engineers across diverse teams.
Requirements
- Bachelor’s, Master’s, PhD, or equivalent experience in Computer Science, Computer Engineering, Electrical Engineering, or a related field.
- 3+ years of software development experience.
- Strong experience with modern C++ standards, including C++11, C++14, C++17, and C++20.
- Strong grasp of machine learning concepts.
- Experience and knowledge of computer architecture, data structures, and algorithms.
- Excellent communication skills and an aptitude for collaboration and teamwork.
Preferred Qualifications
- Experience developing system software.
- Proficiency in Python.
- Experience with GPU kernel programming using CUDA or OpenCL.
- Experience with software performance benchmarking, profiling, and optimization.
- Background in compiler development.
- Experience working with TensorRT, PyTorch, TensorFlow, ONNX Runtime, or other machine learning frameworks.
Benefits
- Eligibility for equity and benefits.
- NVIDIA is committed to fostering an inclusive work environment and is an equal opportunity employer.
The position is full time. Applications will be accepted at least until June 13, 2026. NVIDIA uses AI tools in its recruiting processes.
More jobs at Nvidia
Senior Staff Network Automation Engineer
Nvidia · Santa Clara, United States
USD 208,000-333,500 per year
Senior MLOps Engineer
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Technical Program Manager - Autonomous Vehicles
Nvidia · Santa Clara, United States
USD 168,000-322,000 per year
Technical Product Marketing Engineer, Metropolis - New College Grad 2026
Nvidia · Santa Clara, United States
USD 92,000-184,000 per year
Senior Data Analyst - Automotive
Nvidia · Santa Clara, United States
USD 168,000-322,000 per year
Similar jobs
Senior Software Engineer, CUDA Deep Learning Systems
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Software Engineer, CUDA Deep Learning Systems
Nvidia · Santa Clara, United States
USD 124,000-195,500 per year
Senior AI-Native Systems Software Engineer, TensorRT
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior Software Engineer, Machine Learning Inference
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior Software Engineer, AI Inference Systems
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Deep Learning Frameworks Sustaining Engineer
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
DL Performance Software Engineer - LLM Inference
Nvidia · Toronto, Canada
CAD 135,000-220,000 per year
Senior Software Engineer, Metropolis Vision AI
Nvidia · Santa Clara, United States
USD 224,000-356,500 per year