Senior AI Engineer, High Performance AI

at Nvidia
USD 152,000-241,500 per year
SENIOR
✅ On-site

Tech Stack

AI @ 6 Agentic AI CUDA @ 4 Deep Learning @ 4 GPU @ 4 Leadership @ 6 Performance Optimization @ 4 Python @ 7 Reinforcement Learning @ 6

Details

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. Today, NVIDIA is using AI to define the next era of computing, with GPUs powering computers, robots, and self-driving cars.

The team is looking for an outstanding High Performance AI Engineer to build groundbreaking multi-agent systems for the CUDA ecosystem. The team develops innovative agentic runtimes and compiler-integrated orchestration that work with NVIDIA's software stack to accelerate modern agent workloads powered by foundational models. The role involves developing agent abstractions, GPU-centric runtimes, and compiler- or runtime-driven system solutions to accelerate agent planning, tool use, code generation, and other high-impact AI workloads. You will collaborate with internal NVIDIA software and hardware teams to bring the latest developments into NVIDIA products.

Responsibilities

  • Design, build, and optimize agentic AI systems for the CUDA ecosystem.
  • Co-design agentic system solutions with software, hardware, and algorithm teams.
  • Influence and adopt new capabilities as they become available.
  • Develop reproducible, high-fidelity evaluation frameworks covering performance, quality, and developer productivity.
  • Collaborate across the AI stack, including hardware, compilers and toolchains, kernels and libraries, frameworks, distributed training, inference and serving, and model and agent teams.

Requirements

  • Bachelor's degree in Computer Science, Electrical Engineering, or a related field, or equivalent experience. A master's degree or PhD is preferred.
  • At least 3 years of industry or academic experience in AI systems development.
  • Exposure to building foundational models, agents, or orchestration frameworks.
  • Hands-on experience with deep learning frameworks and modern inference stacks.
  • Strong C/C++ and Python programming skills.
  • Solid software engineering fundamentals.
  • Experience with GPU programming and performance optimization using CUDA or an equivalent technology.

Preferred Qualifications

  • Strong experience building and evaluating deep learning models, coding agents, and developer tooling.
  • Ability to optimize and deploy high-performance models, including on resource-constrained platforms.
  • Demonstrated GPU performance optimization experience, evidenced by benchmark wins or published results.
  • Publications or open-source leadership in deep learning, multi-agent systems, reinforcement learning, or AI systems.
  • Contributions to widely used repositories or standards.

Compensation And Benefits

The base salary range is USD 152,000–241,500 per year. Base salary is determined by location, experience, and the pay of employees in similar positions. The role also includes eligibility for equity and benefits.

Applications will be accepted at least until September 1, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes and is an equal opportunity employer.

More jobs at Nvidia

Similar jobs