Safeguards Enforcement Analyst, Integrity & Authenticity
📍 Washington, United States
📍 New York City, United States
📍 San Francisco, United States
Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 5
Data Analysis @ 5
Data Science
GDPR @ 3
GenAI
Generative AI @ 3
LLM
Profiling @ 5
Python @ 5
SQL @ 5
Security @ 2
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
Anthropic is seeking a Safeguards Enforcement Analyst focused on Integrity & Authenticity to build and execute enforcement workflows for its products and services. The role focuses on detecting and mitigating misuse of AI systems for coordinated inauthentic behavior, election manipulation, influence operations, disinformation, and the targeting, tracking, surveillance, or profiling of individuals and groups.
The position may involve exposure to explicit political, violent, or psychologically disturbing content and may require responding to escalations during weekends and holidays, particularly around major electoral events.
Responsibilities
- Design and architect automated enforcement systems and scalable review workflows while maintaining high accuracy.
- Partner with Engineering and Data Science teams to optimize detection models and automated enforcement systems.
- Review flagged content to drive enforcement and policy improvements.
- Enforce usage policies related to AI-enabled influence operations, coordinated inauthentic behavior, election interference, and targeting, tracking, or surveillance.
- Support Safeguards policy design by providing detailed feedback on policy gaps based on real enforcement scenarios.
- Monitor emerging AI policy enforcement practices, threat actor tactics, and regulatory developments related to elections, privacy, and surveillance.
Requirements
- Experience in trust and safety, policy enforcement, threat intelligence, or a closely related field, particularly involving influence operations, disinformation, coordinated inauthentic behavior, election integrity, or privacy and surveillance harms.
- Experience establishing and scaling policy enforcement or content review workflows.
- Proficiency in SQL and/or other data analysis tools for extracting insights from large datasets.
- Experience identifying emerging risks and threat actors and communicating findings to Product, Policy, Engineering, Legal, and other stakeholders.
- Experience working with generative AI products, including writing effective prompts for content review and enforcement.
- Understanding of implementing product policies at scale, including content moderation.
Preferred Qualifications
- Experience conducting cross-platform investigations into influence operations, coordinated inauthentic behavior, or disinformation campaigns.
- Familiarity with open-source intelligence (OSINT) techniques and tools for threat actor tracking and network analysis.
- Working knowledge of privacy law, surveillance technology, or data broker ecosystems.
- Experience with large language models and AI misuse scenarios involving synthetic personas, fabricated quotes, or automated persuasion.
- Familiarity with election security frameworks, campaign finance law, or electoral integrity standards.
- Experience with relevant regulatory frameworks, including the DSA, EU AI Act, FEC regulations, and GDPR.
- Experience working with election bodies, civil society organizations, or government agencies on integrity or disinformation issues.
- Proficiency in Python for data analysis and automation.
- Experience with dark web monitoring or tracking threat actors across surface, deep, and dark web environments.
Education And Experience
- Bachelor's degree or equivalent combination of education, training, and experience.
- Required field of study: a field relevant to the role, as demonstrated through coursework, training, or professional experience.
- Required years of experience correlate with the internal job level requirements.
Benefits
Anthropic offers competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and office space for collaboration.