Senior Technical Program Manager, GenAI and Models

at Nvidia
USD 168,000-322,000 per year
SENIOR
✅ On-site

Tech Stack

AI @ 7 Agentic AI @ 4 CI/CD @ 4 Deep Learning Distributed Systems @ 7 Experimentation @ 7 GPU @ 7 Git @ 6 GitHub @ 6 Jira @ 6 Kubernetes @ 4 Performance Analysis @ 4 QA @ 4 Reinforcement Learning @ 7

Details

NVIDIA’s Deep Learning Software team is looking for a Senior Technical Program Manager to lead programs across model pre-training, production reinforcement learning runs, evaluation, and agentic AI infrastructure. The team builds software foundations that help research and engineering teams train, evaluate, and deliver sophisticated AI models, partnering across research, platform engineering, distributed computing, evaluation, and open-source development.

Responsibilities

  • Lead multifunctional programs across training frameworks, evaluation environments, agent and model runtimes, datasets, verifiers, and distributed training infrastructure.
  • Partner with AI researchers, engineering leaders, product, infrastructure, and QA teams to define roadmaps, achievements, release plans, and measurable success criteria.
  • Coordinate large-scale reinforcement learning training and evaluation experiments, including GPU resources, dependency tracking, run scheduling, results reporting, release readiness, technical decisions, integration plans, and program updates.

Requirements

  • Bachelor’s degree in computer science, engineering, or a related technical field, or equivalent experience.
  • 10+ years of technical program management, engineering program management, or related experience delivering sophisticated software platforms.
  • Experience leading global, matrixed programs across research, software engineering, infrastructure, QA, release teams, and partner groups.
  • Strong understanding of the AI model lifecycle, including training, post-training, evaluation, experimentation, production readiness, reinforcement learning concepts, and GPU-accelerated distributed systems.
  • Experience running software releases across repositories, dependencies, test configurations, quality gates, collaborator approvals, open-source workflows, and CI/CD systems.
  • Experience with tools such as GitHub, Git, Jira, Linear, Aha!, or Confluence.

Preferred Qualifications

  • Experience supporting reinforcement learning, post-training, agentic AI, or large-scale model evaluation programs.
  • Familiarity with PPO, GRPO, asynchronous reinforcement learning, distributed inference, rollout generation, policy optimization, evaluation harnesses, verifiers, benchmark development, or reproducible experimentation.
  • Knowledge of GPU infrastructure, distributed training, Kubernetes, workload schedulers, cluster capacity management, performance analysis, open-source contributor workflows, release readiness, or operational metrics.

Benefits

  • Equity and benefits are provided.
  • NVIDIA is committed to fostering an inclusive work environment and is an equal opportunity employer.

Applications will be accepted at least until August 1, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.

More jobs at Nvidia

Similar jobs