Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 6
CUDA @ 6
Debugging @ 4
Distributed Systems @ 7
InfiniBand @ 7
Leadership @ 7
NCCL @ 6
Networking @ 6
Technical Leadership @ 7
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
NVIDIA is seeking an outstanding Software Engineer to join its US-based networking software team. As a technical leader, you will help transform AI networking systems, manage complex customer engagements, and influence product and architecture direction.
Responsibilities
- Establish yourself as a technical specialist in AI networking products, specifically the BlueField DPU and ConnectX product lines.
- Architect, design, and develop innovative, scalable, and high-performance hardware-accelerated software solutions.
- Lead deep technical engagements with hyperscalers, including design-in, coding, bring-up, performance tuning, failure analysis, and production hardening.
- Partner with internal engineering, product, and architecture teams to transform customer needs into product features, reference architectures, tooling, and guidelines.
- Drive performance, reliability, and debuggability improvements across customer stacks.
- Translate findings into actionable product, firmware, and software roadmap items.
Requirements
- Bachelor's, Master's, or PhD in Software Engineering, Computer Science, Computer Engineering, Electrical Engineering, or a related science degree, or equivalent experience.
- 8+ years of relevant industry experience, including technical leadership across complex systems.
- Deep knowledge of networking protocols and distributed systems, including RoCE/InfiniBand, L1–L4 fundamentals, and performance and latency tradeoffs.
- Proven low-level software expertise with proficiency in C/C++ and the ability to debug across firmware, drivers, operating systems, and applications.
- Experience in high-performance networking and system-level debugging, including packet drops, retransmissions, congestion, QoS, ordering, and buffer management.
- Excellent interpersonal skills and the ability to explain complex topics to engineers, product managers, and customer collaborators.
- Ability to align cross-organizational teams toward decisions and manage shifting priorities and requirements.
Preferred Qualifications
- Customer-facing technical leadership experience at hyperscalers or cloud service providers.
- Hands-on expertise with RDMA verbs, DPDK, DOCA, NCCL, CUDA-aware networking, congestion control, and performance tuning at scale.
- Experience building internal tools, telemetry, and automation to improve triage speed and operational excellence.
- Experience leading multi-team initiatives across geographic regions and time zones, influencing without authority.
- Experience using AI-powered tools to accelerate debugging, documentation, and day-to-day engineering efficiency while maintaining strong engineering judgment.
Benefits
- Equity and benefits are provided.
- NVIDIA offers a competitive salary and benefits package.
- NVIDIA is an equal opportunity employer committed to a diverse work environment.
The base salary range is USD 184,000–287,500 for Level 4 and USD 224,000–356,500 for Level 5. The salary is determined based on location, experience, and the pay of employees in similar positions. Applications will be accepted at least until March 26, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.
More jobs at Nvidia
Engineering Manager, Data Labeling Platform
Nvidia · Santa Clara, United States
USD 200,000-391,000 per year
Engineering Manager, Local AI Agents
Nvidia · Santa Clara, United States
USD 224,000-431,200 per year
Senior Deep Learning Software Engineer, DLSim
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior Software Engineer, Fleet Intelligence Backend
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior Staff Business Systems Analyst
Nvidia · Santa Clara, United States
USD 144,000-270,200 per year
Similar jobs
Senior Software Engineer, AI Networking
Nvidia · Seattle, United States
USD 184,000-356,500 per year
Senior Software Engineer, DGX Cloud AI Infrastructure
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Software Engineer, AI Networking
Nvidia · Austin, United States
USD 184,000-356,500 per year
Principal Software Engineer, AI Networking
Nvidia · Santa Clara, United States
USD 272,000-431,200 per year
Senior System Software Engineer – Data Center Compute Diagnostics
Nvidia · Durham, United States
USD 224,000-356,500 per year
Senior Machine Learning Engineer, Model Training and Reinforcement Learning
Nebius · Palo Alto, United States
USD 195,200-262,200 per year
Senior HPC Cluster Engineer
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Distinguished Software Architect - Deep Learning and HPC Communications
Nvidia · Santa Clara, United States
USD 320,000-488,800 per year