Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 4
Agentic Systems @ 4
CUDA @ 3
GPU @ 3
LLM @ 6
Machine Learning
Profiling @ 3
Python @ 7
RAG @ 4
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
NVIDIA's AI Developer Tools organization is seeking a Senior Research Engineer to join its Research team. The team builds AI coding agents, models, datasets, and evaluations supporting NVIDIA's strategy to bring AI-powered coding tools to CUDA developers.
The team prototypes and ships coding agents, fine-tunes and evaluates code LLMs, publishes benchmarks such as ComputeEval, and contributes datasets to NVIDIA's Nemotron foundation models. It also develops MCP servers and Agent Skills that integrate NVIDIA developer tools, including Nsight Compute and Nsight Systems, with AI agents. This is a small, collaborative, in-person, high-velocity team focused on applied AI research and shipping production features.
Responsibilities
- Build and improve coding agents that help NVIDIA developers write, optimize, and maintain CUDA code and work alongside other AI agents.
- Design and ship evaluations, including extensions of the public ComputeEval benchmark, for AI-powered CUDA development.
- Fine-tune and specialize code LLMs and partner with the Nemotron team on datasets and evaluations for NVIDIA's foundation models.
- Develop Agent Skills, MCP servers, and other tool-use interfaces for NVIDIA developer tools such as Nsight Compute and Nsight Systems.
- Generate, curate, and validate synthetic training and evaluation data for CUDA programming.
- Deliver new knowledge to frontier LLMs through retrieval-augmented generation (RAG) and skill-based systems that keep models current with NVIDIA's software stack.
- Collaborate with product teams to turn research prototypes into shipping features used by NVIDIA and external customers.
Requirements
- Bachelor's degree in Computer Science or a related technical field, or equivalent experience. A master's degree or Ph.D. is a plus.
- 12 or more years of industry experience in applied AI/ML, including meaningful recent work in AI for code, coding agents, code LLMs, AI developer tools, or adjacent systems.
- Strong proficiency in Python and sound software engineering practices.
- Hands-on experience fine-tuning or evaluating LLMs.
- Fluency with systems aspects of LLM-powered agents, including context management, prompt caching, tool-use design, MCP, and Agent Skills.
- Experience designing or contributing to rigorous evaluations for code generation or agentic systems.
- A track record of taking work beyond the prototype stage and shipping it to real users.
- Comfort working in a small, collaborative, in-person team with rapidly changing priorities and limited process overhead.
Preferred Qualifications
- Public contributions to the AI-for-code space, such as open-source agents or tools, widely used benchmarks, influential papers, or impactful technical blog posts.
- Experience building coding agents or code LLMs used regularly by real users.
- Familiarity with CUDA or other GPU programming, NVIDIA profiling tools such as Nsight Compute and Nsight Systems, or libraries including cuDNN, cuBLAS, Thrust, and CUB.
- Experience with synthetic data generation and quality validation for code.
- Experience with zero-to-one product development or work in a recently rebooted organization with strong customer demand.
Compensation and Benefits
The base salary range is $224,000-$356,500 USD for Level 5 and $272,000-$431,250 USD for Level 6. Compensation is determined based on location, experience, and the pay of employees in similar positions. The role also includes eligibility for equity and benefits.
Applications will be accepted at least until May 4, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes and is an equal opportunity employer.