Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI
CUDA @ 4
Distributed Systems @ 4
Docker @ 6
GPU @ 4
Git @ 6
GitHub @ 6
Kubernetes @ 4
LLM @ 4
Linux @ 6
Observability
Profiling @ 4
PyTorch
Python @ 7
Slurm @ 4
Software Development @ 4
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
NVIDIA is seeking a Software Engineer - Scientific Evaluation to own a shared platform for classical testing, scientific benchmarking, and agentic evaluation. The portfolio spans CUDA and C++ libraries, Python packages, PyTorch integrations, scientific models, and AI agents. This hands-on role combines production software engineering, rigorous measurement, distributed systems, and large GPU fleets.
Responsibilities
- Own the architecture and roadmap for evaluation and benchmarking of scientific software, supporting agentic and numerical evaluations.
- Design and implement infrastructure, benchmarks, and test levels for classical software and scientific agents, balancing rapid turnaround with thorough scientific coverage.
- Establish trusted references, numerical tolerances, calibrated scorers, regression thresholds, and human-review hooks for nondeterministic workloads.
- Operate heterogeneous GPU capacity using multiple control-plane technologies, self-hosted runners, schedulers, queues, containers, caching, observability, and automated recovery.
- Partner with applied scientists, kernel and framework engineers, product teams, and release owners to create reproducible quality signals and production-readiness gates.
Requirements
- Bachelor's or master's degree, or equivalent experience, in Computer Science, Computer Engineering, or a related field.
- More than 5 years of relevant industry experience.
- Strong computer science fundamentals and production C++ and Python skills.
- Fluency with Linux, Git, GitHub/GitLab pipeline orchestration, CMake, Docker, Python packaging, and containers.
- Experience creating test, benchmark, evaluation, or distributed execution systems that deliver versioned, reproducible results across repositories.
- Hands-on experience operating shared GPU compute with Slurm, Kubernetes, or a similar scheduler, including monitoring, isolation, and failure recovery.
- Sound measurement practices covering correctness, variance, flakiness, scorer calibration, and regression detection.
Preferred Qualifications
- Experience with agentic or LLM evaluation, particularly tool-use tasks, sandboxed execution, trace analysis, model-assisted scoring, and calibration.
- Experience with multi-GPU or multi-node systems, CUDA software development, statistical benchmarking, or Nsight profiling.
- Creative and collaborative approach with a focus on reliable scientific results.
Compensation and Benefits
- Base salary range: $152,000–$241,500 for Level 3, or $184,000–$287,500 for Level 4, depending on location, experience, and pay for similar positions.
- Eligible for equity and benefits.
- Full-time position.
- Applications accepted at least until August 13, 2026.
- NVIDIA is an equal opportunity employer and is committed to an inclusive work environment.
More jobs at Nvidia
Senior Site Reliability Engineer - Storage
Nvidia · Santa Clara, United States
USD 168,000-322,000 per year
Senior Director, Global Risk and Compliance
Nvidia · Santa Clara, United States
USD 332,000-500,200 per year
Senior Software Engineer, DGX Cloud Orchestration
Nvidia · Santa Clara, United States
USD 184,000-287,500 per year
Senior System Software Engineer - CPU SoC Boot Firmware
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior Staff Forward-Deployed Engineer, Enterprise AI and Automation
Nvidia · Santa Clara, United States
USD 224,000-356,500 per year
Similar jobs
Senior Software Engineer, AI Inference Systems
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Software QA Test Development Engineer - Diagnostics
Nvidia · Santa Clara, United States
USD 140,000-270,200 per year
Senior Software Development Engineer in Test - Datacenter Server OS
Nvidia · Santa Clara, United States
USD 140,000-270,200 per year
Senior Software Engineer, Golang - DSX MaxQ
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Member of Technical Staff (AI Infrastructure Engineer)
Perplexity AI · San Francisco, United States, Palo Alto, United States
USD 220,000-405,000 per year
NCX Senior Engineer
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior System Software Engineer - AI Performance And Efficiency Tools
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Deep Learning Systems Engineer, Datacenters
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year