Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 6
Deep Learning @ 7
Machine Learning
Security
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
The Safety Systems team works to ensure that OpenAI's models can be safely deployed in the real world. The Model Safety Research team advances methods for implementing robust, safe behavior in AI models and addresses safety challenges related to nuanced safety policies, adversarial robustness, privacy, security, and trustworthy AI in safety-critical domains.
Responsibilities
- Conduct state-of-the-art research on AI safety topics, including RLHF, adversarial training, and robustness.
- Implement new methods in OpenAI's core model training and launch safety improvements in OpenAI's products.
- Set research directions and strategies to make AI systems safer, more aligned, and more robust.
- Coordinate and collaborate with cross-functional teams, including trust and safety, legal, policy, and other research teams, to ensure products meet high safety standards.
- Evaluate and understand the safety of models and systems, identify areas of risk, and propose mitigation strategies.
Requirements
- Passion for AI safety and making cutting-edge AI models safer for real-world use.
- 4+ years of experience in AI safety, especially in areas such as RLHF, adversarial training, robustness, fairness, and bias.
- Ph.D. or another degree in computer science, machine learning, or a related field.
- Experience conducting safety work for AI model deployment.
- In-depth understanding of deep learning research and/or strong engineering skills.
- Ability to work collaboratively as part of a team.
- Alignment with OpenAI's mission of building safe, universally beneficial AGI and its charter.
About OpenAI
OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. The company is an equal opportunity employer and provides reasonable accommodations to applicants with disabilities. Background checks are administered in accordance with applicable law.
Benefits
- Base salary plus equity and performance-related bonuses for eligible employees.
- Medical, dental, and vision insurance, with employer contributions to Health Savings Accounts.
- Pre-tax accounts for health, dependent care, and commuter expenses.
- 401(k) retirement plan with employer match.
- Paid parental, medical, and caregiver leave.
- Paid time off, company holidays, office closures, and paid sick or safe time as required by law.
- Mental health and wellness support.
- Employer-paid basic life and disability coverage.
- Annual learning and development stipend.
- Daily office meals and eligible meal delivery credits.
- Relocation support for eligible employees.
- Additional benefits may include charitable donation matching and wellness stipends.
More jobs at OpenAI
Strategic Delivery Lead, Intelligence Community
OpenAI · Washington, United States
USD 266,000-370,000 per year
Product Manager, Youth
OpenAI · San Francisco, United States
USD 293,000-385,000 per year
Software Engineer, API Safety
OpenAI · San Francisco, United States
USD 293,000-385,000 per year
Head of Marketplace
OpenAI · New York City, United States, San Francisco, United States
USD 400,000-445,000 per year
Research Engineer / Research Scientist, Health
OpenAI · San Francisco, United States
USD 295,000-555,000 per year
Similar jobs
Anthropic Fellows Program
Anthropic · Canada, United States, London, United Kingdom, Berkeley, United States, San Francisco, United States
USD 200,200 per year
Senior MLOps Engineer - DSX Enablement
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Machine Learning Engineer, GenAI Security
Reddit · United States
USD 216,700-303,400 per year
Senior Full-Stack Lead Engineer
Nvidia · Santa Clara, United States
USD 224,000-356,500 per year
Senior Systems Performance Engineer
Nvidia · Santa Clara, United States
USD 136,000-258,800 per year
Developer Technology Engineer, Public Sector - New College Grad 2026
Nvidia · Santa Clara, United States
USD 124,000-241,500 per year
Senior Deep Learning Frameworks Sustaining Engineer
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Machine Learning Engineer
Coinbase · India
INR 4,408,400 per year