Safeguards Enforcement Analyst, Ban Evasion & Recidivism

USD 245,000-285,000 per year
MIDDLE
✅ Remote ✅ Hybrid
✅ Visa Sponsorship

Tech Stack

AI @ 3 Communication @ 6 Data Science @ 3 Fraud @ 3 GenAI Generative AI @ 3 SQL @ 5

Details

Anthropic is seeking a Safeguards Enforcement Analyst to join the account abuse team and build and execute enforcement workflows that keep its products safe. The initial focus will be recidivism: detecting when banned actors return, linking accounts across identities, and closing important re-registration paths, including evasion of child-safety enforcement bans. The role may expand into broader areas of policy enforcement over time.

Responsibilities

  • Investigate evasion clusters end to end, from individual appeals or signal anomalies to full linked actor networks.
  • Convert individual findings into durable systemic controls and detection proposals.
  • Operationalize re-registration controls for high-severity ban populations.
  • Partner with Engineering and Data Science teams on account-linking signals to connect returning actors across identities.
  • Build a recidivism measurement framework covering return frequency, detection speed, and the effectiveness of controls.
  • Author playbooks for contractor-supported evasion review and perform quality assurance against a gold standard.
  • Keep up to date with emerging AI policy enforcement best practices and apply them to decision-making and workflows.

Requirements

  • Experience investigating ban evasion, multi-accounting, or repeat fraud actors on a platform with adversarial users.
  • Fluency in SQL and comfort building analyses across large account and event datasets.
  • Experience with fraud or identity-linking signals and an understanding of their precision and recall tradeoffs.
  • Rigor regarding evidence standards and the asymmetric cost of false positives in severe-harm enforcement.
  • A track record of turning one-off investigations into repeatable detection logic and policy.
  • Strong written communication skills and experience producing clear briefs and recommendations for technical and non-technical stakeholders.
  • Excellent judgment and the ability to collaborate while navigating rapidly evolving priorities and workstreams.

Preferred Qualifications

  • Experience using payment or network risk signals in an enforcement context.
  • Experience with child-safety or other high-severity integrity enforcement.
  • Experience collaborating directly with detection engineering or data science teams on rule deployment.
  • Deep interest in AI safety and responsible technology development.
  • Experience writing effective prompts for generative AI systems in a content review or enforcement context.

Education And Experience

  • Bachelor's degree or an equivalent combination of education, training, and experience.
  • Relevant field of study demonstrated through coursework, training, or professional experience.
  • Required years of experience correlate with the internal job level requirements.

Compensation And Logistics

  • Annual salary: $245,000–$285,000 USD.
  • Remote-friendly role in the United States.
  • Staff are currently expected to work from an Anthropic office at least 25% of the time; some roles may require more office time.
  • Anthropic sponsors visas, although sponsorship is not guaranteed for every role or candidate.
  • Anthropic offers competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and office space for collaboration.

More jobs at Anthropic

Similar jobs