Senior Research Engineer - AI Coding Tools

at Nvidia
USD 224,000-431,200 per year
SENIOR
✅ On-site

Tech Stack

AI @ 4 Agentic Systems @ 4 CUDA @ 3 GPU @ 3 LLM @ 6 Machine Learning Profiling @ 3 Python @ 7 RAG @ 4

Details

NVIDIA's AI Developer Tools organization is seeking a Senior Research Engineer to join its Research team. The team builds AI coding agents, models, datasets, and evaluations supporting NVIDIA's strategy to bring AI-powered coding tools to CUDA developers.

The team prototypes and ships coding agents, fine-tunes and evaluates code LLMs, publishes benchmarks such as ComputeEval, and contributes datasets to NVIDIA's Nemotron foundation models. It also develops MCP servers and Agent Skills that integrate NVIDIA developer tools, including Nsight Compute and Nsight Systems, with AI agents. This is a small, collaborative, in-person, high-velocity team focused on applied AI research and shipping production features.

Responsibilities

  • Build and improve coding agents that help NVIDIA developers write, optimize, and maintain CUDA code and work alongside other AI agents.
  • Design and ship evaluations, including extensions of the public ComputeEval benchmark, for AI-powered CUDA development.
  • Fine-tune and specialize code LLMs and partner with the Nemotron team on datasets and evaluations for NVIDIA's foundation models.
  • Develop Agent Skills, MCP servers, and other tool-use interfaces for NVIDIA developer tools such as Nsight Compute and Nsight Systems.
  • Generate, curate, and validate synthetic training and evaluation data for CUDA programming.
  • Deliver new knowledge to frontier LLMs through retrieval-augmented generation (RAG) and skill-based systems that keep models current with NVIDIA's software stack.
  • Collaborate with product teams to turn research prototypes into shipping features used by NVIDIA and external customers.

Requirements

  • Bachelor's degree in Computer Science or a related technical field, or equivalent experience. A master's degree or Ph.D. is a plus.
  • 12 or more years of industry experience in applied AI/ML, including meaningful recent work in AI for code, coding agents, code LLMs, AI developer tools, or adjacent systems.
  • Strong proficiency in Python and sound software engineering practices.
  • Hands-on experience fine-tuning or evaluating LLMs.
  • Fluency with systems aspects of LLM-powered agents, including context management, prompt caching, tool-use design, MCP, and Agent Skills.
  • Experience designing or contributing to rigorous evaluations for code generation or agentic systems.
  • A track record of taking work beyond the prototype stage and shipping it to real users.
  • Comfort working in a small, collaborative, in-person team with rapidly changing priorities and limited process overhead.

Preferred Qualifications

  • Public contributions to the AI-for-code space, such as open-source agents or tools, widely used benchmarks, influential papers, or impactful technical blog posts.
  • Experience building coding agents or code LLMs used regularly by real users.
  • Familiarity with CUDA or other GPU programming, NVIDIA profiling tools such as Nsight Compute and Nsight Systems, or libraries including cuDNN, cuBLAS, Thrust, and CUB.
  • Experience with synthetic data generation and quality validation for code.
  • Experience with zero-to-one product development or work in a recently rebooted organization with strong customer demand.

Compensation and Benefits

The base salary range is $224,000-$356,500 USD for Level 5 and $272,000-$431,250 USD for Level 6. Compensation is determined based on location, experience, and the pay of employees in similar positions. The role also includes eligibility for equity and benefits.

Applications will be accepted at least until May 4, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes and is an equal opportunity employer.

More jobs at Nvidia

Similar jobs