Senior Software Engineer, Infrastructure and Tooling Lead - Automation
at Nvidia
USD 184,000-356,500 per year
Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 4
API @ 4
Bash @ 7
CI/CD @ 4
Distributed Systems @ 7
Kubernetes @ 7
LLM @ 4
Leadership @ 7
Linux @ 7
Mentoring @ 7
Observability @ 4
Python @ 7
RAG @ 4
Software Development @ 7
Technical Leadership @ 7
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
NVIDIA is looking for a senior technical lead to drive infrastructure and tooling development for our Automation team. This role will focus on building scalable internal platforms, automation frameworks, developer productivity tools, and LLM-powered workflows that improve engineering efficiency across complex software development and validation environments.
Responsibilities
- Lead the design and development of infrastructure, automation frameworks, and internal engineering tools.
- Build scalable services, APIs, dashboards, workflow engines, and integrations that improve developer efficiency and operational visibility.
- Develop LLM-based workflows for triage, summarization, code and log analysis, test workflow assistance, report generation, and knowledge retrieval.
- Integrate tooling with CI/CD systems, source control, issue tracking, test infrastructure, dashboards, and internal engineering services.
- Define architecture, coding standards, evaluation methods, and reliability practices for automation and LLM-enabled systems.
- Mentor engineers, review designs, and provide technical leadership across infrastructure and tooling projects.
Requirements
- BS or MS in Computer Science, Computer Engineering, Electrical Engineering, or equivalent experience.
- 8+ years of software engineering experience, with strong hands-on development skills.
- Proven experience building infrastructure, automation systems, developer tools, workflow platforms, or internal engineering services.
- Strong programming experience in Python, Bash, C, and C++, with experience building infrastructure, automation, and systems-level tooling in Linux-based environments.
- Experience designing systems that integrate with CI/CD pipelines, source control systems, issue trackers, databases, APIs, and distributed services.
- Hands-on experience developing LLM-based workflows, agents, RAG systems, timely pipelines, or AI-assisted automation tools.
- Practical understanding of LLM workflow reliability, including evaluation, guardrails, error handling, observability, and human-in-the-loop review.
- Strong technical leadership, architecture ownership, mentoring, and cross-team collaboration skills.
Ways To Stand Out From The Crowd
- Experience building engineering efficiency platforms or automation infrastructure for large-scale software organizations.
- Experience with test automation, validation infrastructure, build systems, release workflows, or developer experience tooling.
- Familiarity with embeddings, vector search, RAG, model evaluation, agent orchestration, or LLM workflow frameworks.
- Strong background in Linux, containers, Kubernetes, cloud or on-prem infrastructure, and distributed systems.
- Prior experience leading a small technical team or serving as a technical lead for multi-functional infrastructure projects.
Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.
You will also be eligible for equity and benefits.
More jobs at Nvidia
Senior Program Manager, Logistics And Fulfillment Technology Enablement
Nvidia · Santa Clara, United States
USD 176,000-276,000 per year
Senior NPN Program Operations Analyst
Nvidia · Santa Clara, United States
USD 176,000-276,000 per year
Distinguished Engineer, Ai Safety And Security Engineering
Nvidia · Santa Clara, United States
USD 320,000-488,800 per year
Senior Manager, Ai Safety And Security Engineering
Nvidia · Santa Clara, United States
USD 272,000-431,200 per year
Senior Systems Software Engineer, Developer Productivity And Cloud Automation - GeForce NOW
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Similar jobs
Principal Engineer, AI Tooling and Workflows
Nvidia · Santa Clara, United States
USD 272,000-431,200 per year
Senior SWQA Test Development Engineer
Nvidia · Santa Clara, United States
USD 168,000-322,000 per year
Ncx Senior Engineer
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Security Engineer, AI Security
Reddit · United States
USD 190,800-267,100 per year
Senior Staff Software Engineer — AI Applications And Platform Foundations
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Software Engineer - Analytics Platform
Bloomberg · New York City, United States
USD 160,000-240,000 per year
Technical Support Engineer (4pm-12am PST)
Sentry · San Francisco, United States
USD 100,000-120,000 per year
Staff Software Engineer, EAA CX
Coinbase · United States
USD 218,000-256,500 per year