Senior Math Libraries Engineer – Emulation in AI and HPC
Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI
CUDA @ 6
Communication @ 7
Deep Learning
GPU @ 4
HPC
Mathematics @ 4
NLP
Performance Optimization @ 4
Product Management @ 4
Python @ 4
Technical Leadership
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
We are looking for software engineers to join our math libraries teams for AI and HPC kernel generation, specifically targeting emulation of math operations across different precisions. Our GPU-accelerated math libraries support applications in healthcare, natural language processing, virtual reality, deep learning, autonomous vehicles, scientific simulations, engineering, and data analytics.
Responsibilities
- Scope, design, and implement high-quality, high-performance numerical dense linear algebra software on GPUs.
- Provide technical leadership and feedback to library engineers working on projects.
- Mentor interns when appropriate.
- Work closely with product management and internal and external customers to understand feature and performance requirements.
- Help define the technical roadmaps of libraries.
- Identify opportunities to improve library performance and reduce code maintenance overhead through re-architecting.
Requirements
- PhD or Master’s degree in Computer Science, Applied Mathematics, or a related science or engineering field, or equivalent experience.
- 5+ years of experience designing, developing, testing, maintaining, and performance-optimizing production software using CUDA and C++.
- Good knowledge of GPU or CPU hardware architecture; GPU knowledge is preferred.
- Strong fundamentals in finite-precision arithmetic and numerical methods for linear algebra.
- Strong teamwork, communication, and documentation skills.
Preferred Qualifications
- Experience with CUTLASS.
- Low-level programming experience, such as assembly, for performance optimization.
- Experience with a scripting language, preferably Python.
- Experience working in a globally distributed team.
Compensation
For Poland, the base salary range is 292,500 PLN–507,000 PLN for Level 4 and 375,000 PLN–650,000 PLN for Level 5. Base salary is determined based on location, experience, and the pay of employees in similar positions.
Company Overview
NVIDIA’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. GPU deep learning later helped ignite modern AI, with GPUs powering computers, robots, and self-driving cars. NVIDIA is growing its teams at the forefront of technological advancement and offers competitive salaries and a generous benefits package.