Senior Solution Architect, HPC and AI - NVIS

at Nvidia

📍 Santa Clara, United States

$148,000-276,000 per year

SENIOR
✅ On-site

SCRAPED

Used Tools & Technologies

Not specified

Required Skills & Competences ?

Kubernetes @ 4 Python @ 6 Data Science @ 7 Communication @ 4 Mathematics @ 7 Rust @ 6

Details

Do you want to be part of the team that brings Artificial Intelligence (AI) emerging technology to the field? We are looking for a hardworking Solution Architect (SA) to join the NVIDIA AI Enterprise (NVAIE) SA Segment Team. The mission of the NVAIE Segment team is to guide and enable the successful adoption at scale of NVIDIA AI Enterprise Software in production.

In our Solutions Architecture team, we work with NVIDIA's pioneering hardware and software, driving the latest breakthroughs in artificial intelligence. We need people who enable customer adoption of NVIDIA technology and develop lasting relationships with our technology partners, making NVIDIA a key design choice for end-user solutions. On this team, you will support full stack deployment including architectural designs, workload orchestration and application optimization. At NVIDIA, you will be immersed in a diverse, encouraging environment where everyone is inspired to do their life's work. Come join the team and see how you can make a lasting impact on the world!

Responsibilities

  • Primary responsibilities will include building and enabling robust AI/HPC infrastructure for customers.
  • Support operational and reliability aspects of large-scale AI clusters, focusing on performance at scale, training stability, real-time monitoring, logging, and alerting.
  • Engage in and improve services from inception and design through deployment, operation, and optimization.
  • Co-design telemetry of AI workloads to help engineering build solutions for more robust workloads at scale.
  • Communicate across internal teams to support the continuous improvement of NVIDIA's offerings and software designs.

Requirements

  • Strong foundational expertise, from a BS, MS, or Ph.D. degree in Engineering, Mathematics, Physics, Computer Science, Data Science, or similar (or equivalent experience).
  • 8+ years of experience and knowledge of neural networks including good understanding of transformer architectures. Experience designing large scale AI workloads with SLURM and/or Kubernetes.
  • Proficiency with Python / C++ / Rust or other popular software languages.
  • Excellent verbal, written communication, and technical presentation skills in English.
  • You are motivated to work with multiple levels and teams across organizations.
  • Strong analytical and problem-solving skills.
  • Strong time-management and organization skills for coordinating multiple initiatives, priorities and implementations of new technology and products into very sophisticated projects.
  • You are a curious self-starter with a desire for continuous learning and sharing knowledge across the team.

Benefits

  • You will also be eligible for equity and benefits. NVIDIA accepts applications on an ongoing basis.