Researcher, Trustworthy Ai

at OpenAI
USD 380,000 per year
MIDDLE
✅ Hybrid
✅ Relocation

Tech Stack

AI @ 3 LLM @ 5 Python @ 5

Details

About the team

The Safety Systems team is responsible for various safety work to ensure our best models can be safely deployed to the real world to benefit society, and is at the forefront of OpenAI's mission to build and deploy safe AGI, driving our commitment to AI safety and fostering a culture of trust and transparency.

The Trustworthy AI team works on action relevant or decision relevant research to ensure we shape A(G)I keeping societal impacts in mind. This includes work on full stack policy problems such as building methods for public inputs into model values and understanding impacts of anthropomorphism of AI. We aim to translate nebulous policy problems to be technically tractable and measurable. We then use this work to inform and build interventions that increase societal readiness for increasingly intelligent systems.

Our team also works on external assurances for AI with an aim for increasing independent checks and forming additional layers of validation.

About the role

We are looking to hire exceptional research scientists/engineers that can push the rigor of work needed to increase societal readiness for AGI. Specifically, we are looking for those that will enable us to translate nebulous policy problems to be technically tractable and measurable.

This role is based in our San Francisco HQ. We offer relocation assistance to new employees.

In this role, you will enable:

Responsibilities

  • Set research and strategies to study societal impacts of our models in an action-relevant manner and figure out how to tie this back into model design
  • Build creative methods and run experiments that enable public input into model values
  • Increase rigor of external assurances by turning external findings into robust evaluations
  • Facilitate and grow our ability to effectively de-risk flagship model deployments in a timely manner

Requirements

You might thrive in this role if you:

  • Are excited about OpenAI’s mission of building safe, universally beneficial AGI and are aligned with OpenAI’s charter
  • Demonstrate a passion for AI safety and making cutting-edge AI models safer for real-world use
  • Possess 3+ years of research experience (industry or similar academic experience) and proficiency in Python or similar languages
  • Thrive in environments involving large-scale AI systems and multimodal datasets
  • Enjoy working on large-scale, difficult, and nebulous problems in a well-resourced environment
  • Exhibit proficiency in the field of AI safety, focusing on topics like RLHF, adversarial training, robustness, LLM evaluations
  • Have past experience in interdisciplinary research
  • Show enthusiasm for socio-technical topics

Benefits

  • Medical, dental, and vision insurance for you and your family, with employer contributions to Health Savings Accounts
  • Pre-tax accounts for Health FSA, Dependent Care FSA, and commuter expenses (parking and transit)
  • 401(k) retirement plan with employer match
  • Paid parental leave (up to 24 weeks for birth parents and 20 weeks for non-birthing parents), plus paid medical and caregiver leave (up to 8 weeks)
  • Paid time off: flexible PTO for exempt employees and up to 15 days annually for non-exempt employees
  • 13+ paid company holidays, and multiple paid coordinated company office closures throughout the year for focus and recharge, plus paid sick or safe time (1 hour per 30 hours worked, or more, as required by applicable state or local law)
  • Mental health and wellness support
  • Employer-paid basic life and disability coverage
  • Annual learning and development stipend
  • Daily meals in our offices, and meal delivery credits as eligible
  • Relocation support for eligible employees
  • Additional taxable fringe benefits, such as charitable donation matching and wellness stipends, may also be provided

More details about our benefits are available to candidates during the hiring process.

Notes on compensation

The base pay offered may vary depending on multiple individualized factors, including market location, job-related knowledge, skills, and experience. If the role is non-exempt, overtime pay will be provided consistent with applicable laws. In addition to the salary range listed above, total compensation also includes generous equity, performance-related bonus(es) for eligible employees, and the benefits listed above.

This role is at-will and OpenAI reserves the right to modify base pay and other compensation components at any time based on individual performance, team or company results, or market conditions.

More jobs at OpenAI

Similar jobs

Safeguards Enforcement Analyst, Integrity & Authenticity
Anthropic · Washington, United States, United States, New York City, United States, San Francisco, United States
USD 285,000-330,000 per year
Safeguards Enforcement Analyst, Cyber Harm
Anthropic · Washington, United States, United States, New York City, United States, San Francisco, United States
USD 285,000-330,000 per year
Member of Technical Staff (Software Engineer, Connector Platform)
Perplexity AI · Palo Alto, United States, San Francisco, United States, New York City, United States, Seattle, United States
USD 220,000-405,000 per year
Member Of Technical Staff (Forward Deployed Engineer, Applied Ai)
Perplexity AI · London, United Kingdom, New York City, United States, San Francisco, United States, Seattle, United States, Palo Alto, United States
USD 205,000-335,000 per year
Applied AI Architect, Enterprise Tech
Anthropic · San Francisco, United States, New York City, United States, Seattle, United States, Boston, United States
USD 240,000-315,000 per year
Security Engineer, Detection & Response
Anthropic · Washington, United States, New York City, United States, San Francisco, United States, Seattle, United States
USD 300,000-405,000 per year
Safeguards Enforcement Analyst, Violence & Extremism
Anthropic · Washington, United States, New York City, United States, San Francisco, United States, World
USD 285,000-330,000 per year
Staff + Senior Software Engineer, Inference Deployment
Anthropic · San Francisco, United States, New York City, United States, Seattle, United States
USD 320,000-485,000 per year