Senior Software Engineer, Spatial Intelligence and Foundation Models

at Nvidia
USD 184,000-356,500 per year
SENIOR
✅ On-site

Tech Stack

AI @ 4 Algorithms @ 4 Claude Code @ 6 Codex @ 6 Communication @ 6 Computer Vision @ 7 Data Pipelines Deep Learning @ 7 Experimentation @ 6 GPU @ 4 JAX @ 4 PyTorch @ 4 Python @ 7 Robotics @ 4 TensorFlow @ 4

Details

NVIDIA is seeking a Software Engineer to join the Isaac Spatial Intelligence team within the Isaac Engineering organization. The team focuses on geometric and semantic understanding and reasoning for robots, building perception systems that transform raw sensor data into actionable world understanding and advance physical AI.

Responsibilities

  • Design, implement, and deploy algorithms for spatial understanding, including SLAM, structure-from-motion, optical flow, scene flow, object reconstruction, and training vision-language models on spatial reasoning skills.
  • Develop robust perception, mapping, and reasoning systems that run on robots, in simulation, and at scale in data pipelines for training foundation models.
  • Advance geometric computer vision by combining classical multi-view geometry and optimization with modern deep learning approaches.
  • Train and evaluate vision-language models on skills relevant to robotics.
  • Collaborate with the Cosmos and GR00T teams on perception and spatial understanding for foundation models.
  • Work with the Isaac Sim/Lab, Platform, and SQA teams to deploy, validate, extend, and release capabilities and features on physical robots and in large-scale simulations.
  • Integrate, validate, and release applied research in collaboration with research teams and on NVIDIA's robotics platforms.
  • Support prototypes, open-source software contributions, patents, and publications.
  • Collaborate with product, hardware, and software teams to translate engineering work into products.

Requirements

  • PhD or master's degree in Computer Science, Robotics, or a related field, or equivalent experience.
  • At least 8 years of experience working with computer vision, robotics, or deep learning technologies.
  • Strong foundation in 3D geometric computer vision, including multi-view geometry, visual odometry/SLAM, structure-from-motion, or dense correspondence such as optical flow, scene flow, or stereo.
  • Strong hands-on programming skills in Python and/or C++.
  • Experience with deep learning frameworks such as PyTorch, JAX, or TensorFlow.
  • Experience training and evaluating deep learning models for perception tasks; familiarity with vision-language models is a strong plus.
  • Proficiency with AI coding agents and agentic workflows, such as Claude Code, Cursor, or Codex, to accelerate development, testing, and experimentation.
  • Excellent communication, organizational, and interpersonal skills.

Preferred Qualifications

  • Contributions to widely used SLAM, structure-from-motion, or 3D reconstruction systems through open-source projects or shipped products.
  • Experience with large-scale model training on GPU clusters, including vision-language models or other foundation models.
  • Publications at computer vision or robotics venues such as CVPR, ICCV, ECCV, ICRA, IROS, RSS, or CoRL.
  • Hands-on experience with simulation-based training and evaluation, sim-to-real, and real-to-sim transfer.

Compensation and Benefits

The base salary range is USD 184,000–287,500 for Level 4 and USD 224,000–356,500 for Level 5. Compensation is determined by location, experience, and the pay of employees in similar positions. The role also includes eligibility for equity and benefits.

Applications will be accepted at least until July 20, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes and is an equal opportunity employer.

More jobs at Nvidia

Similar jobs