Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 6
Agentic Systems @ 4
Communication @ 6
Compliance @ 4
Data Science
Engineering Management
Fraud @ 4
Hiring @ 4
LLM
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
Anthropic’s Safeguards team ensures that its models and products are developed and deployed safely. The Review Tooling team builds systems that human reviewers—and increasingly Claude—use to investigate potential harms and take enforcement actions across Anthropic’s first-party products and third-party cloud platforms.
The role owns the tools safety investigators use to understand activity on Anthropic’s platforms and take action, as well as the underlying platform. This includes analytics capabilities, privacy-preserving primitives that keep review workflows compatible with data-retention commitments, and a sandbox for rapidly developing and iterating on review interfaces and workflows. The role will also drive the use of automation to scale review, enabling Claude to extend human reviewers’ capabilities while keeping people involved where their judgment is most important.
Responsibilities
- Lead, grow, and develop a team of engineers building investigation, review, and enforcement tooling for first-party and third-party platform surfaces.
- Define the vision and roadmap for the review tooling platform, including analytics, privacy-compatible data-access primitives, and a sandbox for rapidly developing new review interfaces.
- Drive the strategy for scaling review through automation, including enabling reviewers to use Claude effectively and building toward Claude-assisted and Claude-driven review workflows.
- Partner with policy, operations, legal, privacy, and data science stakeholders to translate enforcement and investigation needs into reliable, well-designed systems.
- Ensure review tooling evolves alongside new privacy primitives and data-retention commitments.
- Create clarity for the team and stakeholders in an ambiguous and evolving environment.
- Take an inclusive and equitable approach to hiring and coaching technical talent while maintaining a high-performing team.
- Contribute to engineering-wide initiatives as a member of Anthropic’s engineering management community.
Requirements
- Experience managing software engineering teams, including hiring, coaching, and developing engineers.
- A technical background in full-stack or platform engineering, with the ability to engage deeply in architecture and design discussions.
- Experience shipping internal tools or platforms for demanding operational users, with a track record of measurably improving their workflows.
- Experience working cross-functionally with non-engineering partners such as operations, policy, or legal teams.
- Excellent communication skills, including the ability to explain technical tradeoffs to non-technical stakeholders.
- Interest in the societal impacts of AI and making powerful systems safer.
- Bachelor’s degree or an equivalent combination of education, training, and experience. The field of study must be relevant to the role as demonstrated through coursework, training, or professional experience.
Preferred Qualifications
- 4+ years of management experience and 10+ years of industry software engineering experience.
- Experience building trust and safety, integrity, fraud, abuse-prevention, or other human-review tooling at scale.
- Experience designing systems under strict privacy, compliance, or data-governance constraints, such as zero-data-retention environments.
- Experience integrating LLMs or agentic systems into operational workflows, or building human-in-the-loop automation.
- Experience building developer platforms or extensible tooling frameworks used by other teams.
- Experience supporting enforcement or moderation systems across multiple product surfaces, including enterprise or cloud-platform contexts.
Compensation
The annual salary range is £325,000–£390,000 GBP.
Work Policy and Sponsorship
Anthropic currently expects staff to work from one of its offices at least 25% of the time, although some roles may require more office time. Anthropic sponsors visas for eligible roles and candidates and makes reasonable efforts to obtain visas, with assistance from an immigration lawyer.
Anthropic offers competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and office space for collaboration.