Senior Software Engineer, Spatial Intelligence and Foundation Models
Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 4
Algorithms @ 4
Claude Code @ 6
Codex @ 6
Communication @ 6
Computer Vision @ 7
Data Pipelines
Deep Learning @ 7
Experimentation @ 6
GPU @ 4
JAX @ 4
PyTorch @ 4
Python @ 7
Robotics @ 4
TensorFlow @ 4
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
NVIDIA is seeking a Software Engineer to join the Isaac Spatial Intelligence team within the Isaac Engineering organization. The team focuses on geometric and semantic understanding and reasoning for robots, building perception systems that transform raw sensor data into actionable world understanding and advance physical AI.
Responsibilities
- Design, implement, and deploy algorithms for spatial understanding, including SLAM, structure-from-motion, optical flow, scene flow, object reconstruction, and training vision-language models on spatial reasoning skills.
- Develop robust perception, mapping, and reasoning systems that run on robots, in simulation, and at scale in data pipelines for training foundation models.
- Advance geometric computer vision by combining classical multi-view geometry and optimization with modern deep learning approaches.
- Train and evaluate vision-language models on skills relevant to robotics.
- Collaborate with the Cosmos and GR00T teams on perception and spatial understanding for foundation models.
- Work with the Isaac Sim/Lab, Platform, and SQA teams to deploy, validate, extend, and release capabilities and features on physical robots and in large-scale simulations.
- Integrate, validate, and release applied research in collaboration with research teams and on NVIDIA's robotics platforms.
- Support prototypes, open-source software contributions, patents, and publications.
- Collaborate with product, hardware, and software teams to translate engineering work into products.
Requirements
- PhD or master's degree in Computer Science, Robotics, or a related field, or equivalent experience.
- At least 8 years of experience working with computer vision, robotics, or deep learning technologies.
- Strong foundation in 3D geometric computer vision, including multi-view geometry, visual odometry/SLAM, structure-from-motion, or dense correspondence such as optical flow, scene flow, or stereo.
- Strong hands-on programming skills in Python and/or C++.
- Experience with deep learning frameworks such as PyTorch, JAX, or TensorFlow.
- Experience training and evaluating deep learning models for perception tasks; familiarity with vision-language models is a strong plus.
- Proficiency with AI coding agents and agentic workflows, such as Claude Code, Cursor, or Codex, to accelerate development, testing, and experimentation.
- Excellent communication, organizational, and interpersonal skills.
Preferred Qualifications
- Contributions to widely used SLAM, structure-from-motion, or 3D reconstruction systems through open-source projects or shipped products.
- Experience with large-scale model training on GPU clusters, including vision-language models or other foundation models.
- Publications at computer vision or robotics venues such as CVPR, ICCV, ECCV, ICRA, IROS, RSS, or CoRL.
- Hands-on experience with simulation-based training and evaluation, sim-to-real, and real-to-sim transfer.
Compensation and Benefits
The base salary range is USD 184,000–287,500 for Level 4 and USD 224,000–356,500 for Level 5. Compensation is determined by location, experience, and the pay of employees in similar positions. The role also includes eligibility for equity and benefits.
Applications will be accepted at least until July 20, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes and is an equal opportunity employer.