Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 3
API
Machine Learning
Product Management @ 5
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
Anthropic is seeking a Product Manager for the Safeguards team to own the ideation, design, development, and deployment of safeguards systems and related product user experiences. The role focuses on advancing frontier models safely across Claude.ai, the first-party API, and external cloud providers. You will work with research and product teams to develop detections, evaluations, interventions, and tools that measure and mitigate deployment and user risks.
The ideal candidate is deeply committed to making AI safe and beneficial, has technical expertise in the development, deployment, and measurement of safeguards systems, and thrives in rapidly changing and ambiguous environments.
Responsibilities
- Determine how to build safety by design upstream and leverage downstream defenses for frontier models, AI products, and customers across Claude.ai, the first-party API, and external cloud providers.
- Write safety evaluations and communicate externally about safety.
- Drive impact through prioritization by defining problems, evaluating solution options, clarifying business and technical tradeoffs, and establishing requirements for minimum viable products and ideal states.
- Align and collaborate with policy, enforcement, research, engineering, and cross-functional stakeholders.
- Understand the AI landscape and ecosystem to plan mitigations for deployment risks associated with increasingly powerful models and determined adversaries.
- Lead the development of metrics to understand the area, system performance, and blind spots, informing future project planning.
Requirements
- Ability to make technical tradeoff decisions, ideally with experience working across policy experts, AI/ML research engineers, and software engineering teams to design and build state-of-the-art safety systems.
- Strong understanding of how products are used, their safeguards concerns, and how to provide effective solutions.
- Demonstrated ability to build product and engineering strategy across multiple cross-functional teams in a rapidly changing environment.
- Experience designing and building metrics to evaluate risks, system performance, and user impact, and to make clear tradeoffs.
- Ability to navigate and prioritize rapidly changing product specifications and flex into different domains to bring clarity and execute.
- Evidence of sound judgment and decision-making in ambiguous situations.
- Experience planning, building, launching, and measuring new products or systems in a zero-to-one environment.
- Ability to clearly communicate complex technical concepts to non-technical audiences in writing and verbally.
- Ability to think creatively about the risks and benefits of new technologies and beyond existing checklists and playbooks.
- 5+ years of product management experience, with a focus on understanding problems quickly, building roadmaps with tractable progress, and working with data, detections and interventions, infrastructure and tools, and/or evaluations.
- Bachelor's degree or an equivalent combination of education, training, and experience. The field of study must be relevant to the role through coursework, training, or professional experience.
Benefits
Anthropic offers competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and an office space for collaboration. Anthropic currently expects staff to work from one of its offices at least 25% of the time, though some roles may require more in-office time.