AI Software Engineer, Lightspeed Studios

at Nvidia
USD 224,000-431,200 per year
SENIOR
✅ Remote

Tech Stack

AI @ 4 CUDA @ 4 Deep Learning @ 8 GPU @ 4 GenAI Generative AI @ 4 LLM @ 4 Machine Learning @ 8 PyTorch @ 4 Python @ 6 SGLang @ 6 TensorRT @ 4 vLLM @ 6

Details

At NVIDIA Lightspeed Studios, we are passionate about pushing the limits of technology. The team combines NVIDIA's AI and graphics technology with advanced games and tools to shape the future of gaming. This role focuses on bringing generative AI to games and taking new, unannounced projects powered by state-of-the-art AI models from research to real-time, interactive gaming experiences.

Existing projects include RTX Remix, Zorah, Project R2X, and NVIDIA AI for Gaming.

Responsibilities

  • Build and ship software for upcoming, not-yet-announced projects that bring world models and video diffusion models to real-time gaming, from prototype to release.
  • Train, fine-tune, and evaluate models, including data curation and adapting models to game-specific content and controls.
  • Optimize models and inference for latency, token throughput, and quality using techniques such as distillation, quantization, and reduced-step sampling on RTX devices and in the cloud.
  • Profile and remove bottlenecks across the stack, from model architecture to GPU kernels.
  • Integrate models into Unreal Engine, Unity, and custom engines so they run efficiently alongside rendering.
  • Lead technical decisions, mentor other engineers, and collaborate with NVIDIA Research, rendering, and art teams.

Requirements

  • BS, MS, or PhD in Computer Science, Electrical Engineering, or a related field, or equivalent experience.
  • 12+ years of software engineering experience, including substantial hands-on experience shipping AI, deep learning, or machine learning systems.
  • Expert-level C++ and Python.
  • Hands-on experience training and fine-tuning deep learning models with PyTorch or a similar framework, including distributed training.
  • Experience in one or more of the following areas: video generation, world models, diffusion or flow-matching models, transformers, or neural rendering.
  • Experience optimizing GPU inference with CUDA, TensorRT, TensorRT-LLM, or Triton.
  • A track record of turning complex prototypes into shipped products and communicating clearly across research, engineering, and art.

Preferred Qualifications

  • Hands-on experience with interactive, action-conditioned, or real-time world models or video generation.
  • Experience developing custom CUDA or Triton kernels.
  • Experience with game engine development or real-time rendering pipelines.
  • Open-source contributions such as Diffusers, FastVideo, vLLM, SGLang, or TensorRT-LLM; publications; or patents.

Benefits

  • Equity and benefits are available.
  • NVIDIA is committed to fostering an inclusive work environment and is an equal opportunity employer.

Applications will be accepted at least until October 2, 2026. This posting is for an existing vacancy.

More jobs at Nvidia

Similar jobs