Technical Program Manager, Infrastructure Systems & Tooling

at OpenAI
USD 225,000-285,000 per year
MIDDLE
✅ Hybrid
✅ Relocation

Tech Stack

AI @ 3 API @ 5 BI @ 5 Prioritization @ 3 Reporting @ 3 SQL

Details

OpenAI's Industrial Compute organization is building and operating the infrastructure foundation for the next generation of AI. Infrastructure Operations works across facilities, hardware, network operations, incident management, data center engineering, delivery teams, and external partners to bring capacity online safely, understand its operational state, and improve it over time.

As OpenAI's data center portfolio grows across first-party and partner-delivered capacity, the organization needs clear goals, trusted data, repeatable processes, and systems that make ownership, risk, readiness, and performance visible. This role will help build the operating mechanisms that allow Infrastructure Operations to scale with rigor.

The Technical Program Manager will own the systems, data, reporting, governance, and program-management backbone for Infrastructure Operations. Reporting to the Delivery & Operations Lead, the role translates strategy into executable goals and operating cadences, turns operational needs into software and data solutions, and creates mechanisms that keep a rapidly evolving organization aligned and accountable.

The role also owns the current 1P+3P delivery-tracking layer within Operations, including milestones, delivery timelines, quantity forecasts, risks, decisions, and executive reporting. The position partners with 1P Delivery Program Management, Compute TPMs, Data Center Engineering, construction, commissioning, and operations leaders to ensure that delivery information becomes complete, usable input for readiness, handover, and ongoing operations.

Responsibilities

  • Establish and run Infrastructure Operations' goal-setting and operating cadence, including quarterly goals, weekly performance reviews, prioritization, action tracking, decision logs, escalation paths, and closure criteria.
  • Translate leadership priorities into clear programs with owners, milestones, dependencies, success metrics, and resourcing assumptions.
  • Maintain source-of-truth hygiene across goals, status, dates, risks, and decisions.
  • Build and operate dashboards, scorecards, and executive-ready reporting for operational health, capacity readiness, SLA and MTTR performance, delivery pipeline status, and material risks.
  • Define authoritative data models and reporting standards for milestones, timelines, delivery risk, capacity state, readiness, handover, exceptions, and operational performance.
  • Own the 1P+3P delivery-tracking program across sites and partners, integrating schedule and progress inputs, maintaining quantity and timeline forecasts, surfacing risks early, and driving cross-functional follow-through.
  • Define Operations' requirements for capacity acceptance and operational handover, including readiness evidence, risk and exception workflows, approvals, sign-offs, and post-handover action tracking.
  • Own the Information Governance Process and controlled-document lifecycle across relevant Operations workflows, including standards, procedures, work instructions, templates, ownership, approvals, versioning, exceptions, and obsolescence.
  • Lead operations software and tooling implementations from discovery through rollout, including workflow mapping, requirements, data and integration patterns, engineering and vendor partnerships, testing, and adoption.
  • Improve connections among reporting, ticketing, knowledge, delivery-tracking, sourcing, and operational systems so teams can work from consistent data instead of manual, fragmented updates.
  • Program-manage cross-functional initiatives across Infrastructure Operations and its interfaces with Data Center Engineering, Compute TPMs, 1P Delivery, construction, commissioning, sourcing, and external infrastructure partners.
  • Coordinate internal and external resources supporting systems, dashboards, process design, and document governance.
  • Capture lessons learned, identify recurring operational bottlenecks, and implement automation and process improvements that make the organization more predictable, scalable, and effective.

Requirements

  • 8+ years of experience in technical program management, operations program management, infrastructure delivery, operations transformation, or a comparable role in a complex technical environment.
  • Experience leading end-to-end software or systems implementations for an operations organization, including workflow discovery, requirements, data models, integrations, testing, rollout, adoption, and continuous improvement.
  • Fluency with dashboards, KPIs, operational data, and executive reporting, with the ability to turn incomplete inputs into clear definitions, trusted metrics, and decisions.
  • Experience building governance mechanisms such as goal-setting, operating reviews, intake and prioritization, risk and issue management, action tracking, decision logs, and escalation.
  • Experience with data centers, construction, commissioning, infrastructure operations, cloud, manufacturing, or another mission-critical physical-infrastructure environment.
  • Ability to influence senior leaders, technical DRIs, vendors, and partner organizations without relying on direct authority.
  • Ability to communicate clearly from working-team detail to executive summary.
  • Comfort navigating ambiguity, changing ownership boundaries, urgent timelines, and high operational stakes while maintaining rigor and momentum.
  • Bachelor's degree in Engineering, Computer Science, Information Systems, Construction Management, Operations Management, Business, or an equivalent combination of education and practical experience.

Preferred Skills

  • Experience with hyperscale or AI infrastructure, including 1P, 3P, colocation, or CSP delivery models.
  • Familiarity with project controls, scheduling, operational readiness, commissioning, capacity acceptance, SLA/MTTR reporting, incident or ticketing workflows, and handover governance.
  • Technical fluency with APIs, integration patterns, BI/reporting tools, workflow platforms, data quality controls, and automation.
  • SQL or equivalent analytical skills.
  • Experience managing vendors, consultants, or embedded support resources and converting ad hoc support into repeatable operating capability.

Benefits

  • Equity, performance-related bonus(es) for eligible employees, and benefits in addition to base salary.
  • Medical, dental, and vision insurance for employees and their families, with employer contributions to Health Savings Accounts.
  • Pre-tax accounts for Health FSA, Dependent Care FSA, and commuter expenses.
  • 401(k) retirement plan with employer match.
  • Paid parental leave, medical leave, and caregiver leave.
  • Paid time off, paid company holidays, office closures, and paid sick or safe time as required by applicable law.
  • Mental health and wellness support.
  • Employer-paid basic life and disability coverage.
  • Annual learning and development stipend.
  • Daily meals in offices and meal delivery credits as eligible.
  • Relocation support for eligible employees.

More jobs at OpenAI

Similar jobs