Safeguards Policy Analyst, Cyber Harms

USD 190,000-285,000 per year
MIDDLE
✅ Hybrid
✅ Visa Sponsorship

Tech Stack

AI @ 2 Security @ 3

Details

Anthropic is seeking an analyst to support cyber product policy work, including usage policy language, help-center and enforcement guidance, launch policy notes, controlled-access frameworks, and policy commitments related to cyber-relevant use of AI systems. The role involves engaging deeply with technical material such as probes, classifiers, threat assessments, and trusted-access programs, and translating technical evaluation and safeguard work into clear policy.

Responsibilities

  • Contribute to cyber product policy artifacts, including usage policy language, help-center and enforcement guidance, and launch policy notes.
  • Ensure policy artifacts remain consistent with Anthropic's constitution and access tiers.
  • Analyze the constitution and cyber-related usage policies to verify that enforcement decisions and safeguards align with policy.
  • Maintain a gap log and propose text fixes.
  • Work with threat intelligence and enforcement teams to ensure cyber safety standards are met.
  • Regularly review samples of enforcement decisions against policy text and report policy drift.
  • Coordinate cyber policy inputs for model releases and regulatory requirements alongside the Senior Cyber Policy Lead.
  • Prepare the cyber policy section of launch and model-card reviews and regulator pre-briefs.
  • Support Anthropic's access-requirements policy and help consolidate the controlled-access framework across existing and emerging access programs.
  • Keep cyber-related policy commitments current as industry and regulatory standards evolve, ensuring they map cleanly to technical safeguards.
  • Draft inputs for reporting to partners and regulators.
  • Translate technical evaluation and safeguard work into policy positions and internal guidance.
  • Engage with technical teams to understand probes and classifiers relevant to access policy.

Requirements

  • Demonstrated ability to write clear policy analysis in a professional or academic setting.
  • Familiarity with how AI developers or platforms enforce usage policies, or with cybersecurity policy, standards, or regulatory frameworks such as usage policies/AUPs, trust and safety enforcement, NIST CSF, or coordinated vulnerability disclosure.
  • Ability to read and interpret technical security material, including vulnerability reports and threat assessments.
  • Ability to work across threat intelligence, enforcement, engineering, and policy teams and communicate clearly in writing.
  • A bachelor's degree or equivalent combination of education, training, and experience.
  • Relevant field of study demonstrated through coursework, training, or professional experience.

Preferred Qualifications

  • Experience in cybersecurity policy, including familiarity with coordinated vulnerability disclosure.
  • Exposure to government information-sharing and incident-notification frameworks.
  • Experience supporting engagement with government agencies, regulators, or standards bodies on cybersecurity or AI matters.
  • Experience with legal, technical, or policy aspects of vulnerabilities and disclosure.
  • Experience with model-release or product-launch review processes.
  • Familiarity with trust-and-safety or product-policy work at a platform, including usage policies, enforcement appeals, and policy communications.

Benefits

Anthropic offers competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and office space for collaboration.

Staff are currently expected to work from one of Anthropic's offices at least 25% of the time, although some roles may require more office time. Anthropic sponsors visas where possible and retains an immigration lawyer to assist with the process.

More jobs at Anthropic

Similar jobs