Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI
API
Algorithms @ 4
CUDA @ 4
Debugging @ 6
Deep Learning @ 4
GPU @ 4
GenAI
Generative AI
LLM
LLVM @ 4
OpenCL @ 4
Performance Analysis @ 4
Performance Optimization
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
NVIDIA is hiring software engineers for the CUDA Tile team. NVIDIA GPUs are at the center of the deep learning revolution, enabling breakthroughs in generative AI, large language models, recommendation systems, speech recognition, image classification, and other areas.
Responsibilities
- Work on CUDA Tile, a tile-based programming model for NVIDIA GPUs that shipped with CUDA 13.1.
- Design and implement compiler transformations.
- Develop MLIR-based dialects and lowering passes.
- Optimize the performance of tile-based kernels across multiple generations of NVIDIA GPU architectures.
- Define public APIs and implement compiler and optimization techniques.
- Perform performance optimization and general software engineering work.
- Independently define project goals and scope and lead development efforts.
Requirements
- Bachelor's, Master's, or Ph.D. in Computer Science, Computer Engineering, or a related field, or equivalent experience.
- Three or more years of relevant work or research experience in compiler optimization, performance analysis, and IR design.
- Excellent C/C++ programming and software design skills, including debugging, performance analysis, and test design.
- Ability to work independently and in a dynamic, product-oriented team.
- Strong interpersonal skills.
Preferred Qualifications
- Knowledge of CPU and/or GPU architecture.
- CUDA or OpenCL programming experience.
- Experience with MLIR, LLVM, XLA, TVM, and deep learning models and algorithms.
Benefits
- Equity and benefits are provided.
- NVIDIA is committed to fostering an inclusive work environment and is an equal opportunity employer.
Applications for this job will be accepted at least until June 20, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.
The base salary range is USD 152,000–241,500, determined by location, experience, and the pay of employees in similar positions.
More jobs at Nvidia
Senior System Software Engineer - CPU SoC Boot Firmware
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
System Software Engineer – Data Center Compute Diagnostics
Nvidia · Durham, United States
USD 152,000-241,500 per year
Senior Research Scientist - Generative World Models for Autonomous Driving and Physical AI
Nvidia · Santa Clara, United States
USD 192,000-356,500 per year
Senior Technical Marketing Engineer - CAE Performance
Nvidia · Santa Clara, United States
USD 136,000-253,000 per year
Senior Software Technical Program Driver - OEM and NCP Escalations
Nvidia · Santa Clara, United States
USD 168,000-258,800 per year
Similar jobs
Senior DL Compiler Engineer – CUDA Tile
Nvidia · Santa Clara, United States
USD 152,000-241,500 per year
Senior Compiler Engineer, AI Inference Platforms
Nvidia · Santa Clara, United States
USD 152,000-241,500 per year
Senior Deep Learning Software Engineer, Inference and Model Optimization
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Software Engineer - Local AI
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior AI Compiler Engineer, Algorithms and Code Generation
Nvidia · Santa Clara, United States
USD 152,000-241,500 per year
Senior Deep Learning Software Engineer, Inference
Nvidia · United States
USD 152,000-287,500 per year
Senior Performance Architect, Nemotron
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior Performance Compiler Engineer - Triton
Nvidia · Redmond, United States
USD 184,000-287,500 per year