Researcher, Robustness & Safety Training

at OpenAI
USD 295,000-445,000 per year
SENIOR
✅ On-site
✅ Relocation

Tech Stack

AI @ 6 Deep Learning @ 7 Machine Learning Security

Details

The Safety Systems team works to ensure that OpenAI's models can be safely deployed in the real world. The Model Safety Research team advances methods for implementing robust, safe behavior in AI models and addresses safety challenges related to nuanced safety policies, adversarial robustness, privacy, security, and trustworthy AI in safety-critical domains.

Responsibilities

  • Conduct state-of-the-art research on AI safety topics, including RLHF, adversarial training, and robustness.
  • Implement new methods in OpenAI's core model training and launch safety improvements in OpenAI's products.
  • Set research directions and strategies to make AI systems safer, more aligned, and more robust.
  • Coordinate and collaborate with cross-functional teams, including trust and safety, legal, policy, and other research teams, to ensure products meet high safety standards.
  • Evaluate and understand the safety of models and systems, identify areas of risk, and propose mitigation strategies.

Requirements

  • Passion for AI safety and making cutting-edge AI models safer for real-world use.
  • 4+ years of experience in AI safety, especially in areas such as RLHF, adversarial training, robustness, fairness, and bias.
  • Ph.D. or another degree in computer science, machine learning, or a related field.
  • Experience conducting safety work for AI model deployment.
  • In-depth understanding of deep learning research and/or strong engineering skills.
  • Ability to work collaboratively as part of a team.
  • Alignment with OpenAI's mission of building safe, universally beneficial AGI and its charter.

About OpenAI

OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. The company is an equal opportunity employer and provides reasonable accommodations to applicants with disabilities. Background checks are administered in accordance with applicable law.

Benefits

  • Base salary plus equity and performance-related bonuses for eligible employees.
  • Medical, dental, and vision insurance, with employer contributions to Health Savings Accounts.
  • Pre-tax accounts for health, dependent care, and commuter expenses.
  • 401(k) retirement plan with employer match.
  • Paid parental, medical, and caregiver leave.
  • Paid time off, company holidays, office closures, and paid sick or safe time as required by law.
  • Mental health and wellness support.
  • Employer-paid basic life and disability coverage.
  • Annual learning and development stipend.
  • Daily office meals and eligible meal delivery credits.
  • Relocation support for eligible employees.
  • Additional benefits may include charitable donation matching and wellness stipends.

More jobs at OpenAI

Similar jobs