Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 3
Distributed Systems @ 3
GPU @ 3
HPC
Machine Learning
Networking
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
OpenAI’s Compute organization delivers the compute infrastructure behind its advanced AI models. The team works across software, hardware, facilities, operations, and engineering disciplines to make large-scale compute available, reliable, and efficient.
The role supports the design, construction, scaling, and operation of OpenAI’s compute infrastructure. Depending on background, work may involve large-scale distributed systems, ML infrastructure, hardware systems, manufacturing, supply chain, data center development, or the physical engineering systems required to bring compute capacity online.
Responsibilities
- Help build, scale, and operate OpenAI’s global compute infrastructure.
- Solve complex problems across software, hardware, manufacturing supply chain, and data center systems.
- Improve the reliability, performance, efficiency, and scalability of critical infrastructure.
- Partner with research, engineering, hardware, operations, and infrastructure teams to bring new compute capacity online quickly and reliably.
- Identify bottlenecks across technical, operational, and physical systems, and develop practical solutions.
- Build tools, processes, systems, or infrastructure that improve execution at scale.
- Contribute to the long-term architecture and operational maturity of OpenAI’s compute footprint.
Requirements
- Experience building, scaling, or operating complex technical systems.
- Ability to work on ambiguous, high-impact problems where the path forward is not always defined.
- Comfortable collaborating across software, hardware, operations, and physical infrastructure disciplines.
- Strong technical judgment and a bias toward execution.
- Commitment to reliability, speed, safety, and operational excellence.
- Interest in building infrastructure at unprecedented scale and supporting the development and deployment of frontier AI.
Preferred Skills
- Experience with AI infrastructure, high-performance computing, distributed systems, GPU clusters, or cloud-scale platforms.
- Experience with hardware systems, manufacturing, supply chain, data center development, or large capital infrastructure projects.
- Domain expertise in civil, controls, mechanical, hardware, electrical, thermal, power, networking, or facilities engineering.
- Experience bringing technical platforms, data centers, factories, or large-scale systems from concept to production.
- Experience working in fast-moving environments where technical depth and execution speed both matter.
Compensation And Benefits
- Base salary: $125,000–$400,000 per year.
- Equity, performance-related bonuses for eligible employees, and benefits.
- Medical, dental, and vision insurance, with employer contributions to Health Savings Accounts.
- Pre-tax FSA, dependent care, and commuter accounts.
- 401(k) retirement plan with employer match.
- Paid parental, medical, and caregiver leave.
- Paid time off, company holidays, and paid sick or safe time as required by applicable law.
- Mental health and wellness support.
- Employer-paid basic life and disability coverage.
- Annual learning and development stipend.
- Daily meals in offices and eligible meal delivery credits.
- Relocation support for eligible employees.
- Additional benefits may include charitable donation matching and wellness stipends.
More jobs at OpenAI
Strategic Delivery Lead, Intelligence Community
OpenAI · Washington, United States
USD 266,000-370,000 per year
Product Manager, Youth
OpenAI · San Francisco, United States
USD 293,000-385,000 per year
Software Engineer, API Safety
OpenAI · San Francisco, United States
USD 293,000-385,000 per year
Head of Marketplace
OpenAI · New York City, United States, San Francisco, United States
USD 400,000-445,000 per year
Research Engineer / Research Scientist, Health
OpenAI · San Francisco, United States
USD 295,000-555,000 per year
Similar jobs
Data Center Compute Infrastructure
OpenAI · United States, San Francisco, United States, Seattle, United States
USD 230,000-490,000 per year
Forward Deployed Engineer - Physical AI Cloud Platform
Nebius · United States, Austin, United States
USD 179,500-224,300 per year
Software Engineer, Workload Enablement
OpenAI · San Francisco, United States, Seattle, United States
USD 293,000-385,000 per year
Member of Technical Staff (AI Infrastructure Engineer)
Perplexity AI · Palo Alto, United States, San Francisco, United States
USD 220,000-405,000 per year
Senior Staff Platform Engineer
Nvidia · Santa Clara, United States
USD 200,000-322,000 per year
Senior Deep Learning Software Infrastructure Engineer
Nvidia · United States
USD 224,000-431,200 per year
Industrial Compute
OpenAI · United States
USD 150,000-300,000 per year
Senior Customer Success Engineer - DGX Cloud
Nvidia · Santa Clara, United States
USD 200,000-322,000 per year