Product Manager, Safeguards (Verticals)

USD 305,000-385,000 per year
MIDDLE
✅ Hybrid
✅ Visa Sponsorship

Tech Stack

AI @ 3 API Communication @ 3 Machine Learning Product Management @ 5

Details

Anthropic's Safeguards team builds protections for new AI features and products, including detections, evaluations, interventions, and tools to measure and mitigate deployment and user risks. The Product Manager will own the ideation, design, development, and deployment of Safeguards systems and related product user experiences across Claude.ai, the first-party API, and external cloud providers. This role works closely with research, product, policy, enforcement, engineering, and other cross-functional teams to advance frontier AI models safely.

The ideal candidate is deeply committed to making AI safe and beneficial, has technical expertise in the development, deployment, and measurement of Safeguards systems, and thrives in rapidly changing and ambiguous environments.

Responsibilities

  • Determine how to build safety by design upstream and leverage downstream defenses for Anthropic's frontier models, AI products, and customers across Claude.ai, the first-party API, and external cloud providers.
  • Write safety evaluations and communicate externally about safety.
  • Drive impact through prioritization by clearly defining problems, solution options, business and technical tradeoffs, and requirements for minimum viable products versus ideal states.
  • Align and collaborate with policy, enforcement, research, engineering, and cross-functional stakeholders.
  • Understand the AI landscape and ecosystem to plan mitigations for deployment risks involving increasingly powerful models and determined adversaries.
  • Lead the development of metrics to understand the area, system performance, and blind spots, helping inform future project planning.

Requirements

  • Ability to make technical tradeoff decisions, ideally with experience working across policy experts, AI/ML research engineers, and software engineering teams to design and build state-of-the-art safety systems.
  • Strong understanding of how products are used, their Safeguards concerns, and how to provide effective solutions.
  • Demonstrated ability to build product and engineering strategy across multiple cross-functional teams in a rapidly changing environment.
  • Demonstrated experience designing and building metrics to evaluate risks, system performance, and user impact, and making clear tradeoffs.
  • Very strong ability to navigate and prioritize rapidly changing product specifications and flex into different domains to provide clarity and execute.
  • Evidence of sound judgment and decision-making in ambiguous situations.
  • Experience planning, building, launching, and measuring new products or systems in a zero-to-one environment.
  • Ability to clearly articulate complex technical concepts to non-technical audiences in written and verbal communication.
  • Ability to think creatively about the risks and benefits of new technologies and beyond existing checklists and playbooks.
  • 5+ years of product management experience, with a focus on quickly understanding problems, building roadmaps with tractable progress, and working in detail on data, detection and interventions, infrastructure and tools, and/or evaluations.
  • Bachelor's degree or an equivalent combination of education, training, and experience. The field of study must be relevant to the role as demonstrated through coursework, training, or professional experience.

Benefits

Anthropic offers competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and an office space for collaboration. Anthropic currently expects staff to be in one of its offices at least 25% of the time, although some roles may require more time in the office.

More jobs at Anthropic

Similar jobs