Full Stack Engineer, Fleet Scheduling

at OpenAI
USD 230,000-490,000 per year
MIDDLE
✅ Hybrid
✅ Relocation

Tech Stack

AI @ 3 API @ 6 Angular @ 3 Azure @ 6 Data Visualization @ 3 Distributed Systems @ 3 Docker @ 3 Go @ 3 GraphQL @ 6 Kubernetes @ 3 Machine Learning Node.js @ 3 Observability @ 3 Python @ 3 React @ 3 Security

Details

Full Stack engineers on the Fleet Scheduling team build intuitive and scalable interfaces that help researchers manage AI workloads across large-scale supercomputing systems. The team develops high-performance systems for real-time insights, resource tracking, resource allocation, and interaction with complex infrastructure.

The role involves designing, developing, and operating web-based systems for OpenAI’s supercomputing clusters. You will collaborate with researchers, product teams, and infrastructure teams to deliver scalable solutions for monitoring, job scheduling, and resource management. This role is based in San Francisco, California, with a hybrid work model requiring three days in the office per week.

Responsibilities

  • Design and develop full-stack web applications to track, monitor, and manage large-scale AI workloads in real time.
  • Translate complex operational needs into intuitive user interfaces and scalable backend systems in collaboration with researchers and infrastructure teams.
  • Build data visualization tools, including Gantt charts and dashboards, to provide insights into job scheduling and resource allocation.
  • Optimize backend services for massive data throughput, low-latency performance, and high availability.
  • Implement frontend components that interact with scheduling, storage, and compute systems.
  • Ensure system security, reliability, and scalability across globally distributed supercomputing infrastructure.

Requirements

  • Significant full-stack development experience.
  • Expertise with modern frontend frameworks such as React, Vue, or Angular.
  • Experience with backend technologies such as Python, Go, or Node.js.
  • Experience building scalable, high-performance web applications for complex distributed systems.
  • Strong understanding of RESTful and GraphQL APIs, distributed databases, and cloud infrastructure, especially Azure.
  • An execution-focused approach with attention to usability, performance, and scalability in enterprise-scale systems.
  • Ability to work effectively in fast-paced, highly collaborative environments with evolving priorities.

Preferred Qualifications

  • Experience with Kubernetes, Docker, and cloud-native application deployment.
  • Understanding of AI/ML workload scheduling and orchestration challenges.
  • Experience with real-time data processing, visualization libraries, and observability tooling.

About OpenAI

OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. The company is an equal opportunity employer and provides reasonable accommodations to applicants with disabilities. Background checks are administered in accordance with applicable law.

Benefits

  • Base salary range of $230,000–$490,000, plus equity.
  • Medical, dental, and vision insurance, with employer contributions to Health Savings Accounts.
  • Pre-tax accounts for health, dependent care, and commuter expenses.
  • 401(k) retirement plan with employer match.
  • Paid parental, medical, and caregiver leave.
  • Paid time off, company holidays, office closures, and applicable sick or safe time.
  • Mental health and wellness support.
  • Employer-paid basic life and disability coverage.
  • Annual learning and development stipend.
  • Daily meals in offices and eligible meal delivery credits.
  • Relocation support for eligible employees.
  • Additional benefits may include charitable donation matching and wellness stipends.

More jobs at OpenAI

Similar jobs