Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 6
Algorithms @ 7
CUDA @ 4
Communication @ 4
Computer Vision @ 7
Data Structures @ 7
Debugging @ 4
Deep Learning @ 7
Distributed Systems @ 7
GPU @ 4
Linux @ 8
Microservices @ 6
Performance Optimization @ 7
PyTorch @ 4
Python @ 8
Robotics
Software Development @ 8
TensorRT @ 4
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
NVIDIA's technology is at the heart of the AI revolution, powering self-driving cars, robotics, copilots, and more. Metropolis is transforming how the physical world is perceived and understood using advanced computer vision and deep learning. The team builds large-scale distributed Vision AI platforms for intelligent spaces, smart cities, retail analytics, and digital twins.
As a System Software Engineer for Vision AI, you will develop and optimize high-performance vision systems that turn massive streams of video, image, and 3D data into actionable insights. You will collaborate with specialists in perception, simulation, and large models to bring research into production at scale.
Responsibilities
- Craft and implement high-performance Vision AI pipelines for real-time and streaming scenarios using computer vision and deep learning models.
- Develop and refine large-scale distributed services for processing video, image, and 3D data in edge and cloud environments.
- Develop multimodal perception capabilities combining 2D, 3D, and temporal information to understand complex real-world scenes.
- Use simulation and synthetic data tools to build, test, and validate perception algorithms at scale.
- Profile and tune GPU-accelerated inference pipelines to meet strict latency, efficiency, and reliability targets.
- Collaborate with product, research, and platform teams to translate requirements into technical designs and robust implementations.
- Drive technical build reviews, promote code quality and testing guidelines, and mentor engineers in Vision AI systems development.
Requirements
- BS, MS, or PhD in Computer Science, Electrical or Computer Engineering, a related field, or equivalent experience.
- 12+ years of professional software development experience using modern C++ (14/17/20) and Python on Linux.
- Strong computer science fundamentals, including algorithms, data structures, concurrency, and distributed systems.
- Demonstrated expertise in computer vision and deep learning, with experience deploying production systems in these fields.
- Experience building and debugging high-performance concurrent systems, including multithreading, asynchronous I/O, and efficient memory management.
- Proficiency working in Linux-based environments with containers and microservices, integrating AI components into scalable back-end services.
- Ability to rapidly prototype vision models and pipelines and evolve them into production-quality services.
- Practical experience with PyTorch for training, fine-tuning, and deploying models for vision tasks.
- Strong analytical and problem-solving skills, with a data-driven approach to performance optimization and system design.
- Excellent written and verbal communication skills, with experience collaborating across time zones and functions.
Preferred Qualifications
- Experience delivering end-to-end computer vision applications in production, such as video analytics, smart cities, autonomous systems, retail analytics, industrial inspection, or digital twins.
- Experience with GPU acceleration, including CUDA, TensorRT, or comparable technologies, and low-level optimization for inference and pre- and post-processing.
- Experience with simulation and synthetic data creation using tools such as Omniverse, Unreal Engine, Unity, or similar digital-twin platforms.
- Background in vision-language models or related multimodal AI, including integrating these models into real products.
- Background in multimedia, including video-centric processing and delivery, codecs, video pipelines, or media frameworks, and integrating vision models into multimedia workflows.
Benefits
NVIDIA offers competitive salaries, equity, and a generous benefits package.
The base salary range is $224,000–$356,500 USD. The salary will be determined based on location, experience, and the pay of employees in similar positions.
Applications will be accepted at least until July 12, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes and is an equal opportunity employer.