Safeguards Enforcement Analyst, Ban Evasion & Recidivism
at Anthropic
📍 United States
📍 Washington, United States
📍 New York City, United States
📍 San Francisco, United States
📍 Washington, United States
📍 New York City, United States
📍 San Francisco, United States
USD 245,000-285,000 per year
Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 3
Communication @ 6
Data Science @ 3
Fraud @ 3
GenAI
Generative AI @ 3
SQL @ 5
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
Anthropic is seeking a Safeguards Enforcement Analyst to join the account abuse team and build and execute enforcement workflows that keep its products safe. The initial focus will be recidivism: detecting when banned actors return, linking accounts across identities, and closing important re-registration paths, including evasion of child-safety enforcement bans. The role may expand into broader areas of policy enforcement over time.
Responsibilities
- Investigate evasion clusters end to end, from individual appeals or signal anomalies to full linked actor networks.
- Convert individual findings into durable systemic controls and detection proposals.
- Operationalize re-registration controls for high-severity ban populations.
- Partner with Engineering and Data Science teams on account-linking signals to connect returning actors across identities.
- Build a recidivism measurement framework covering return frequency, detection speed, and the effectiveness of controls.
- Author playbooks for contractor-supported evasion review and perform quality assurance against a gold standard.
- Keep up to date with emerging AI policy enforcement best practices and apply them to decision-making and workflows.
Requirements
- Experience investigating ban evasion, multi-accounting, or repeat fraud actors on a platform with adversarial users.
- Fluency in SQL and comfort building analyses across large account and event datasets.
- Experience with fraud or identity-linking signals and an understanding of their precision and recall tradeoffs.
- Rigor regarding evidence standards and the asymmetric cost of false positives in severe-harm enforcement.
- A track record of turning one-off investigations into repeatable detection logic and policy.
- Strong written communication skills and experience producing clear briefs and recommendations for technical and non-technical stakeholders.
- Excellent judgment and the ability to collaborate while navigating rapidly evolving priorities and workstreams.
Preferred Qualifications
- Experience using payment or network risk signals in an enforcement context.
- Experience with child-safety or other high-severity integrity enforcement.
- Experience collaborating directly with detection engineering or data science teams on rule deployment.
- Deep interest in AI safety and responsible technology development.
- Experience writing effective prompts for generative AI systems in a content review or enforcement context.
Education And Experience
- Bachelor's degree or an equivalent combination of education, training, and experience.
- Relevant field of study demonstrated through coursework, training, or professional experience.
- Required years of experience correlate with the internal job level requirements.
Compensation And Logistics
- Annual salary: $245,000–$285,000 USD.
- Remote-friendly role in the United States.
- Staff are currently expected to work from an Anthropic office at least 25% of the time; some roles may require more office time.
- Anthropic sponsors visas, although sponsorship is not guaranteed for every role or candidate.
- Anthropic offers competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and office space for collaboration.
More jobs at Anthropic
Technical Program Manager, Silicon
Anthropic · San Francisco, United States, New York City, United States
USD 365,000-435,000 per year
Researcher, Cybersecurity Products
Anthropic · San Francisco, United States
USD 320,000-405,000 per year
Technical Program Manager, RL Research
Anthropic · San Francisco, United States, New York City, United States
USD 365,000-435,000 per year
Product Design Manager
Anthropic · San Francisco, United States, New York City, United States, Seattle, United States
USD 385,000-460,000 per year
Insider Risk Investigator
Anthropic · Washington, United States, Boston, United States, New York City, United States
USD 245,000-305,000 per year
Similar jobs
Safeguards Enforcement Analyst, Account Takeover & Credential Abuse
Anthropic · Washington, United States, New York City, United States, San Francisco, United States
USD 245,000-285,000 per year
Safeguards Enforcement Analyst, Fraud & Scams
Anthropic · Washington, United States, New York City, United States, San Francisco, United States, United States
USD 245,000-285,000 per year
Safeguards Enforcement Analyst, Age-Appropriate Design
Anthropic · Washington, United States, New York City, United States, San Francisco, United States, United States
USD 245,000-285,000 per year
Safeguards Enforcement Analyst, Bio Harms
Anthropic · Washington, United States, United States, New York City, United States, San Francisco, United States
USD 245,000-285,000 per year
Safeguards Enforcement Analyst, Radiological & Nuclear Harms
Anthropic · Washington, United States, New York City, United States, San Francisco, United States
USD 245,000-285,000 per year
Safeguards Enforcement Analyst, Chem & Explosives Harms
Anthropic · Washington, United States, New York City, United States, San Francisco, United States
USD 245,000-285,000 per year
Senior DFT Methodology - Data Analytics and Applied AI Engineer
Nvidia · Santa Clara, United States
USD 152,000-264,500 per year
Analytics Engineer, Safety Systems
OpenAI · San Francisco, United States
USD 210,000-260,000 per year