Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 6
API
ChatGPT @ 4
Codex
Communication @ 6
Data Science
Experimentation @ 4
LLM
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
The ChatGPT Model Flywheel team transforms model advancements into user experiences through reliable serving, rapid experimentation, safe deployment, and continuous improvement. The role focuses on AI infrastructure, model experimentation, deployment, measurement, and engineering team leadership.
Team Focus Areas
- Model experimentation: Enable rapid, safe model validation for ChatGPT and Codex through experiment automation and lifecycle management.
- Model deployment: Ensure safe, scalable deployment of model capabilities with robust rollout and operational tooling. Automate capacity management and incorporate platform-wide health monitors.
- Model measurement: Build comprehensive evaluation and measurement systems for model quality, including user signals and launch scorecards. Improve end-to-end feedback loops for continual model improvement.
Key Partnerships
Collaborate cross-functionally with Model Measurement Data Science, Research, Codex, Fleet, Inference, and API teams.
Responsibilities
- Elevate and consolidate ChatGPT’s harness, context management, and system prompt frameworks.
- Drive the expansion and improvement of multi-tier model experiences.
- Support and scale self-service experiment capabilities and automated guardrails.
- Lead model rollout automation, capacity management, and health monitoring.
- Shape end-to-end measurement systems, including evaluations, grader signals, and user feedback.
- Lead engineering teams in complex, cross-functional environments.
- Collaborate directly with engineering, research, and product stakeholders.
Requirements
- Proven experience leading engineering teams in complex, cross-functional environments.
- Demonstrated success shipping production systems at scale, ideally for AI or large backend services.
- Deep understanding of model-driven product development, deployment lifecycles, and measurement tooling.
- Excellent communication and collaboration skills.
- Experience with large language models, distributed infrastructure, or experimentation platforms is a plus.
Benefits
- Base salary range of $293,000–$385,000 per year, plus equity and performance-related bonuses for eligible employees.
- Medical, dental, and vision insurance, with employer contributions to Health Savings Accounts.
- Pre-tax Flexible Spending Accounts and commuter benefits.
- 401(k) retirement plan with employer match.
- Paid parental, medical, and caregiver leave.
- Paid time off, company holidays, office closures, and paid sick or safe time.
- Mental health and wellness support.
- Employer-paid basic life and disability coverage.
- Annual learning and development stipend.
- Daily office meals and eligible meal delivery credits.
- Relocation support for eligible employees.
- Additional benefits may include charitable donation matching and wellness stipends.
More jobs at OpenAI
GRC Program Manager, Assurance Engineering & Control Systems
OpenAI · San Francisco, United States
USD 216,000-252,000 per year
Android Systems Engineer, Consumer Devices
OpenAI · San Francisco, United States
USD 216,000-342,000 per year
Senior Staff Software Engineer, Identity
OpenAI · Mountain View, United States, San Francisco, United States
USD 345,000-405,000 per year
Analytics Engineer, GTM
OpenAI · San Francisco, United States, New York City, United States
USD 220,000-335,000 per year
Product Designer, Payments
OpenAI · San Francisco, United States
USD 245,000-310,000 per year
Similar jobs
Agent Post-Training, Personality
OpenAI · San Francisco, United States
USD 295,000-445,000 per year
Growth - Lifecycle Lead
OpenAI · San Francisco, United States, New York City, United States
USD 239,000-325,000 per year
Data Scientist, GTM
OpenAI · San Francisco, United States
USD 290,000-340,000 per year
AI Deployment Engineer, Plugins
OpenAI · San Francisco, United States
USD 197,000-280,000 per year
Agent Post-Training, Artifacts Research
OpenAI · San Francisco, United States
USD 295,000-445,000 per year
Research Engineer, Codex
OpenAI · San Francisco, United States
USD 295,000-445,000 per year
Data Scientist
Promptwatch · Amsterdam, Netherlands
EUR 60,000-85,000 per year
Lead - Advanced Analytics
Airbnb · Gurugram, India
INR 3,080,000-4,400,000 per year