Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
API
Azure @ 4
ChatGPT
Distributed Systems @ 4
Networking @ 4
Planning @ 7
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
The Applied team safely brings OpenAI’s technology to the world, powering products like ChatGPT and the APIs for GPT and more. The Applied Infrastructure Technical Program Management team partners across engineering to lead foundational programs that ensure OpenAI’s infrastructure can meet current and future demand.
This role drives critical infrastructure programs across the Applied organization, including compute capacity planning, process transformation, cost and quota attribution and optimization, and coordination across infrastructure and product stakeholders. The role also focuses on evolving infrastructure to support growth, scale, and new products.
Responsibilities
- Serve as the directly responsible individual for complex infrastructure programs spanning CPU planning, orchestration, networking, storage, and other resource management domains.
- Build and operationalize systems to capture demand signals, model future capacity needs, and align infrastructure planning across internal teams and external partners.
- Partner with Infrastructure, Product, and Finance teams to forecast infrastructure usage patterns and ensure supply and demand alignment.
- Lead cost attribution and quota enforcement programs to promote stability and equitable access to resources across teams.
- Drive simplification and standardization of infrastructure tooling and processes across Applied and Infrastructure organizations.
- Lead cross-functional programs to evolve infrastructure to support growth and scale.
- Work with external vendors and partners, including cloud providers, to manage delivery schedules, capacity projections, and cost growth.
- Communicate project risks, progress, and key decisions clearly to technical and executive stakeholders.
Requirements
- 7+ years of experience leading large-scale technical infrastructure programs, especially in compute resource planning, system scalability, or platform engineering.
- Understanding of how cloud infrastructure, distributed systems, and foundational infrastructure components—including compute, storage, and networking—work together to power complex products.
- Experience working with hyperscalers or cloud providers, such as Azure, to manage infrastructure delivery and planning cycles.
- Structured and detail-oriented approach, with the ability to untangle complex dependencies and drive clarity across ambiguous problem spaces.
- Ability to build durable processes, drive alignment across teams, and communicate clearly with technical and business audiences.
- Ability to navigate complex and ambiguous problems at the intersection of infrastructure and business.
Location
San Francisco, California. Hybrid schedule with three days per week in the office.
Benefits
- Base salary range of $257,000–$445,000 per year, plus equity.
- Medical, dental, and vision insurance, with employer contributions to Health Savings Accounts.
- Pre-tax accounts for Health FSA, Dependent Care FSA, and commuter expenses.
- 401(k) retirement plan with employer match.
- Paid parental, medical, and caregiver leave.
- Paid time off, paid company holidays, office closures, and paid sick or safe time.
- Mental health and wellness support.
- Employer-paid basic life and disability coverage.
- Annual learning and development stipend.
- Daily office meals and eligible meal delivery credits.
- Relocation support for eligible employees.
- Additional benefits may include charitable donation matching and wellness stipends.
More jobs at OpenAI
Agent Standards Specialist, Global Affairs
OpenAI · Washington, United States, New York City, United States
USD 171,000-280,000 per year
Technical Program Manager, Infrastructure Systems & Tooling
OpenAI · San Francisco, United States
USD 225,000-285,000 per year
Product Manager, Statsig
OpenAI · San Francisco, United States
USD 401,000-510,000 per year
Software Engineer, Host Assurance
OpenAI · United States, San Francisco, United States, Seattle, United States
USD 266,000-445,000 per year
Software Engineer, HSM Infrastructure Security, Consumer Devices
OpenAI · San Francisco, United States
USD 347,000-445,000 per year
Similar jobs
ChatGPT Performance Engineer
OpenAI · United States, New York City, United States, San Francisco, United States, Seattle, United States
USD 325,000-405,000 per year
Staff+ Software Engineer, Platform
Anthropic · New York City, United States, San Francisco, United States, Seattle, United States
USD 405,000-485,000 per year
TPM Manager, Infrastructure
Anthropic · New York City, United States, San Francisco, United States
USD 365,000-565,000 per year
Staff+ Software Engineer, Platform Distribution
Anthropic · New York City, United States, San Francisco, United States
USD 405,000-485,000 per year
Forward Deployed Engineer - Physical AI Cloud Platform
Nebius · United States, Austin, United States
USD 179,500-224,300 per year
Product Engineer, Ona
OpenAI · London, United Kingdom, San Francisco, United States
USD 255,000-445,000 per year
Network Operations Engineer, Network Activation
OpenAI · San Francisco, United States
USD 157,000-221,000 per year
Staff + Senior Software Engineer, Cloud Inference Launch Engineering
Anthropic · San Francisco, United States
USD 320,000-485,000 per year