Data Center Compute Infrastructure

at OpenAI
USD 230,000-490,000 per year
MIDDLE
✅ Remote ✅ Hybrid
✅ Relocation

Tech Stack

AI @ 3 Distributed Systems @ 3 GPU @ 3 HPC Machine Learning Networking

Details

OpenAI’s Compute organization delivers the infrastructure behind advanced AI models, working across software, hardware, facilities, operations, and engineering disciplines. The team focuses on distributed systems, ML infrastructure, GPU fleets, power, cooling, networking, manufacturing, supply chain, and data center delivery.

The role involves designing, building, scaling, and operating compute infrastructure. Depending on background, work may include large-scale distributed systems, ML infrastructure, hardware systems, manufacturing, supply chain, data center development, or the physical engineering systems required to bring large-scale compute capacity online.

Responsibilities

  • Help build, scale, and operate OpenAI’s global compute infrastructure.
  • Solve complex problems across software, hardware, manufacturing supply chain, and data center systems.
  • Improve the reliability, performance, efficiency, and scalability of critical infrastructure.
  • Partner with research, engineering, hardware, operations, and infrastructure teams to bring new compute capacity online quickly and reliably.
  • Identify bottlenecks across technical, operational, and physical systems and develop practical solutions.
  • Build tools, processes, systems, or infrastructure that improve execution at scale.
  • Contribute to the long-term architecture and operational maturity of OpenAI’s compute footprint.

Requirements

  • Experience building, scaling, or operating complex technical systems.
  • Ability to work on ambiguous, high-impact problems where the path forward is not always defined.
  • Comfort collaborating across software, hardware, operations, and physical infrastructure disciplines.
  • Strong technical judgment and a bias toward execution.
  • Commitment to reliability, speed, safety, and operational excellence.
  • Interest in building infrastructure at unprecedented scale and supporting the development and deployment of frontier AI.

Preferred Skills

  • Experience with AI infrastructure, high-performance computing, distributed systems, GPU clusters, or cloud-scale platforms.
  • Experience with hardware systems, manufacturing, supply chain, data center development, or large capital infrastructure projects.
  • Domain expertise in civil, controls, mechanical, hardware, electrical, thermal, power, networking, or facilities engineering.
  • Experience bringing technical platforms, data centers, factories, or large-scale systems from concept to production.
  • Experience working in fast-moving environments where technical depth and execution speed both matter.

Benefits

  • Base salary of $230,000–$490,000 per year, plus equity.
  • Medical, dental, and vision insurance, with employer contributions to Health Savings Accounts.
  • Pre-tax accounts for health, dependent care, and commuter expenses.
  • 401(k) retirement plan with employer match.
  • Paid parental, medical, and caregiver leave.
  • Paid time off, paid company holidays, office closures, and sick or safe time.
  • Mental health and wellness support.
  • Employer-paid basic life and disability coverage.
  • Annual learning and development stipend.
  • Daily office meals and eligible meal delivery credits.
  • Relocation support for eligible employees.
  • Additional benefits may include charitable donation matching and wellness stipends.

More jobs at OpenAI

Similar jobs