Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 6
ChatGPT
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
Safety Systems manages the lifecycle of safety efforts for OpenAI’s frontier models, including system-level safeguards, model training, evaluation, and red-teaming. The Model Policy team designs policies that define safe and reliable model behavior in real-world environments.
Responsibilities
- Design and maintain model policies for audio, image, video, and omni-modal behavior.
- Translate theories of harm and threat models into behavioral safety policies, evaluation criteria, grading guidance, and safeguards.
- Identify and analyze safety regressions and failure patterns to identify gaps in existing policies and inform policy iteration.
- Develop policy artifacts supporting model training, evaluation, and deployment, including behavior instructions, human-data campaigns, golden sets, and evaluations.
- Partner with AI researchers, domain experts, and product teams to operationalize policy into measurable model behavior.
- Work with multimodal AI models and capabilities such as GPT-Live and ChatGPT Images.
Requirements
- Strong judgment about the real-world risks of advanced multimodal AI systems.
- Experience turning ambiguous safety questions into clear, data-driven policies, behavioral boundaries, and measurable evaluation criteria.
- Ability to treat policy as an end-to-end, measurable system by testing intended model behavior and diagnosing gaps across policy, data, graders, and safeguards.
- Strong technical judgment when designing policies around model behavior that can realistically be trained, measured, and supervised at scale.
- Strong technical fluency and experience using AI tools to accelerate policy development, evaluate model behavior, analyze failure patterns, and turn findings into actionable improvements.
- Hands-on experience with model data and evaluation results, including inspecting examples, analyzing failure patterns, assessing data quality, and distinguishing policy failures from grader, model, or system failures.
- Ability to work in fast-paced, collaborative research environments where priorities shift as models, evidence, and risks change.
- A pragmatic, evidence-driven approach to reducing risk while preserving beneficial uses of AI.
- Experience driving consensus and action in ambiguous spaces.
Benefits
- Base salary range of $266,000–$335,000 per year, plus equity.
- Medical, dental, and vision insurance, with employer contributions to Health Savings Accounts.
- Pre-tax accounts for health, dependent care, and commuter expenses.
- 401(k) retirement plan with employer match.
- Paid parental, medical, and caregiver leave.
- Paid time off, company holidays, office closures, and paid sick or safe time.
- Mental health and wellness support.
- Employer-paid basic life and disability coverage.
- Annual learning and development stipend.
- Daily office meals and eligible meal delivery credits.
- Relocation support for eligible employees.
- Additional benefits may include charitable donation matching and wellness stipends.
This role is based in San Francisco, California, and follows a hybrid work model requiring three days in the office per week. OpenAI is an equal opportunity employer and provides reasonable accommodations to applicants with disabilities.
More jobs at OpenAI
Model Policy Manager, National Security
OpenAI · San Francisco, United States
USD 266,000-335,000 per year
Engineering Manager, Library
OpenAI · Seattle, United States
USD 401,000-445,000 per year
Dedicated Support Engineering Lead
OpenAI · San Francisco, United States
USD 347,000-385,000 per year
Forward Deployed Engineer (FDE), Financial Services – New York City
OpenAI · New York City, United States
USD 185,000-300,000 per year
Infrastructure Operations & Sustainability Lead
OpenAI · United States
USD 223,000-385,000 per year
Similar jobs
Growth - Lifecycle Lead
OpenAI · New York City, United States, San Francisco, United States
USD 239,000-325,000 per year
Technical Program Manager, Multimodal
OpenAI · San Francisco, United States
USD 207,000-445,000 per year
Staff Advanced Analytics, Community Blueprint & Quality
Airbnb · United States
USD 180,000-221,000 per year
Software Engineer, Healthcare
OpenAI · San Francisco, United States
USD 347,000-385,000 per year
Software Engineer, Applied Emerging Talent (2027)
OpenAI · San Francisco, United States
USD 180,000 per year
Engineering Manager, ChatGPT Search Infrastructure
OpenAI · San Francisco, United States
USD 401,000-445,000 per year
Full Stack Software Engineer, Product Explorations
OpenAI · San Francisco, United States
USD 347,000-385,000 per year
Software Engineer, Native Learning Experiences
OpenAI · San Francisco, United States
USD 266,000-385,000 per year