Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI
CUDA @ 4
Debugging @ 7
GPU @ 6
LLVM @ 4
Parallel Programming @ 4
Profiling @ 7
Rust @ 6
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
NVIDIA is hiring a Senior Compiler Engineer to drive the next generation of GPU systems programming. The team is bringing the safety, expressiveness, and modern tooling of Rust to native GPU and CUDA development. The role focuses on building compiler pipelines, custom intermediate representation frameworks, and JIT compilation systems that bridge host and device execution, enabling memory-safe, high-performance GPU kernels in idiomatic Rust.
Responsibilities
- Design, implement, and maintain Rust-to-GPU compiler backends, including
rustccodegen backends and procedural macros. - Build Rust-native intermediate representation frameworks to compile standard Rust to high-performance CUDA PTX and machine code.
- Lower Rust AST and MIR into MLIR, PTX, LLVM, and custom IRs, including GPU-specific optimizations.
- Develop tooling for ahead-of-time, just-in-time, and link-time optimization workflows targeting NVIDIA GPUs, host platforms, and feature sets.
- Architect compiler-enforced safety models that extend Rust ownership, borrowing, and lifetime disciplines across the GPU launch boundary.
- Implement type-safe, device-side Rust abstractions for GPU hardware primitives, including shared memory, barriers, scoped atomics, Tensor Memory Accelerator (TMA), and warp- and cluster-level operations.
- Build composable systems-programming foundations for future accelerated-computing applications.
Requirements
- Bachelor's, Master's, or Ph.D. in Computer Science, Computer Engineering, a related field, or equivalent experience.
- 5+ years of relevant work or research experience in compiler development, language design, or GPU code generation.
- Deep expertise in Rust, including
rustcinternals, Rust MIR, procedural macros, and the borrow-checker and lifetime model. - Hands-on experience with compiler infrastructures, intermediate representations, and code generation, such as LLVM IR, MLIR, or custom IR systems.
- Solid understanding of parallel programming models, GPU architectures, and CUDA programming.
- Strong software design skills, including debugging, profiling, and benchmarking compilers and GPU kernels.
- Ability to orchestrate agents for product requirement design, architecture, code development, testing, code review, and issue triage.
- Ability to work independently, define project goals and scope, and drive complex compiler-engineering efforts from research to production.
Preferred Qualifications
- Contributions to the Rust compiler (
rustc), Cargo tooling, or open-source Rust-to-GPU projects. - Experience building custom compiler front ends, AST translators, or JIT engines.
- Familiarity with MLIR or other extensible compiler frameworks.
- Deep proficiency in low-level GPU programming, including Tensor Cores, warp-level shuffles, and asynchronous transfer pipelines.
- Experience designing domain-specific languages (DSLs) or tile-based programming abstractions for tensor processing.
Benefits
- Competitive salary with a base salary range of USD 152,000–241,500.
- Equity and benefits.
- Inclusive work environment and equal-opportunity employment.
Applications will be accepted at least until September 6, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.
More jobs at Nvidia
Senior NPI Program Manager
Nvidia · Santa Clara, United States
USD 168,000-258,800 per year
GPU PCIe and Boot Architect - New College Grad 2026
Nvidia · Santa Clara, United States
USD 124,000-241,500 per year
Senior AI Engineer, High Performance AI
Nvidia · Santa Clara, United States
USD 152,000-241,500 per year
Senior Salesforce CPQ Developer
Nvidia · Santa Clara, United States
USD 176,000-276,000 per year
Senior Technical Program Manager - LLM Safety
Nvidia · Santa Clara, United States
USD 168,000-322,000 per year
Similar jobs
Senior Software Engineer, AI Inference Systems
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
DL Performance Software Engineer - LLM Inference
Nvidia · Toronto, Canada
CAD 135,000-220,000 per year
Senior Software Engineer, AI Inference Systems
Nvidia · Toronto, Canada
CAD 170,000-275,000 per year
Systems Generalist, GPT Infrastructure
OpenAI · San Francisco, United States, Seattle, United States
USD 293,000-445,000 per year
Senior Compiler Engineer – Rust GPU
Nvidia · Santa Clara, United States
USD 152,000-241,500 per year
Staff+ Software Engineer, Inference Runtime
Anthropic · New York City, United States, San Francisco, United States, Seattle, United States
USD 405,000-485,000 per year
Senior Developer Technology Engineer - Agentic SoC Performance
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior Software Engineer, CUDA Rust Core Libraries
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year