Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI
Algorithms
CI/CD @ 4
CUDA @ 4
Communication @ 7
Computer Vision
Debugging @ 6
GPU @ 4
Jira @ 4
LLM
MPI @ 4
Mathematics @ 4
Parallel Programming @ 6
Product Management @ 4
Project Management @ 4
Software Development @ 4
Technical Leadership
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
We are looking for software engineers to join development efforts in dense linear algebra kernels for high-performance libraries such as cuSOLVER. The team develops GPU-accelerated libraries and SDKs supporting AI, data analytics, scientific and engineering simulations, computer-aided engineering, electronic design automation, quantum chemistry, autonomous vehicles, large language models, computer vision, encryption, and other applications.
In this role, you will work with other developers to design, develop, and optimize kernels for algorithms including triangular factorizations, eigenvalue decompositions, and singular value decompositions.
Responsibilities
- Design, implement, and optimize scalable, high-performance numerical dense linear algebra software on GPUs.
- Provide technical leadership and guidance to library engineers, QA engineers, and interns working on projects.
- Work closely with product management and internal and external partners to understand feature and performance requirements and contribute to technical library roadmaps.
- Identify and realize opportunities to improve library quality, performance, and maintainability through re-architecting and establishing innovative software development practices.
Requirements
- PhD or MSc degree in Computational Science and Engineering, Computer Science, Applied Mathematics, or a related science or engineering field, or equivalent experience.
- 5+ years of experience developing, debugging, and optimizing high-performance numerical linear algebra software using C++ and parallel programming. Experience with CUDA, MPI, OpenMP, OpenACC, or pthreads is preferred.
- Strong fundamentals in numerical methods, including computational linear algebra, linear system solvers, and methods for eigenvalue, singular value, and other decompositions.
- Experience developing dense linear algebra libraries such as BLAS and LAPACK, as well as parallel counterparts such as PBLAS and SCALAPACK.
- Strong collaboration, communication, and documentation skills.
Preferred Qualifications
- Knowledge of CPU and/or GPU hardware architecture.
- Experience adopting and advancing software development practices such as CI/CD systems.
- Experience with project management tools such as JIRA.
- Experience working in a globally distributed organization.
- Strong background in large-scale computing technologies such as PDE solvers, eigenvalue solvers, and time-domain simulation methods, including CFD and FEA.
Benefits
- Equity and benefits are provided.
- NVIDIA is committed to fostering a diverse work environment and is an equal opportunity employer.
- Applications will be accepted at least until January 13, 2026.
- This posting is for an existing vacancy.
More jobs at Nvidia
Senior Systems Software Engineer - Infrastructure
Nvidia · Santa Clara, United States
USD 224,000-356,500 per year
Principal Software Architect, Networking AI
Nvidia · Santa Clara, United States
USD 272,000-431,200 per year
Senior Software Engineer, Agentic Robotics Infrastructure
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
AI and ML Power Methodology Engineer
Nvidia · Santa Clara, United States
USD 136,000-264,500 per year
Systems Software Engineer - Infrastructure
Nvidia · Santa Clara, United States
USD 152,000-241,500 per year
Similar jobs
Senior Math Libraries Engineer - Sparsity in AI
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Math Libraries Engineer - Direct Sparse Solvers
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Math Libraries Engineer - Sparse Linear Algebra
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior Libraries Engineer – AI and HPC
Nvidia · Poland
PLN 221,200-507,000 per year
Senior Math Libraries Engineer – AI and HPC
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Developer Technology Engineer - Agentic SoC Performance
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior Performance Compiler Engineer - Triton
Nvidia · Redmond, United States
USD 184,000-287,500 per year
Senior Software SDET Test Development Engineer
Nvidia · Santa Clara, United States
USD 140,000-270,200 per year