Lead, Frontier Red Team (Cyber)

USD 485,000-755,000 per year
SENIOR
✅ Hybrid
✅ Visa Sponsorship

Tech Stack

AI Security @ 6

Details

Anthropic is forming a team focused on the impact of frontier AI models on cybersecurity. The team will research advanced models' offensive and defensive capabilities, publish research, and build and deploy defenses to help secure the world.

The Lead will oversee Anthropic's mission-driven research on defending the world in an era of advanced cybersecurity capabilities. This includes leading research into Claude's offensive and defensive capabilities, prototyping defenses, informing how Claude is trained and safeguarded to favor defenders, and engaging with the public and government. The role will work closely with the Product, Training, Security, and Safeguards teams.

Responsibilities

  • Define Anthropic's overall defensive cybersecurity vision.
  • Hire and lead a world-class team of cybersecurity researchers and program managers.
  • Design and execute an AGI cybersecurity research program at the scale, speed, and quality of a frontier lab.
  • Work with Training, Safeguards, Policy, and other Anthropic teams on cybersecurity opportunities and challenges.
  • Identify field partners, including maintainers, researchers, companies, and governments, for joint projects.
  • Develop strategies to publish and share research for maximum impact.
  • Translate technical findings into demonstrations and artifacts for policymakers and the public.
  • Identify critical strategic decisions, marshal resources, and support decision-making at the company and AI-lab ecosystem level.
  • Build and deploy defenses into the world and communicate a vision for a more secure future to the company, industry, researchers, and government.

Sample Projects

  • Developing a strategy for how a frontier AI lab can use frontier models and resources to defend the world.
  • Building frameworks and tools that enable AI models to autonomously find and patch vulnerabilities using tens of trillions of tokens.
  • Running purple-team simulations in which AI defenders compete against AI attackers in network environments.
  • Using autonomous AI systems to address real-world security challenges, including bug bounties and CTFs, to characterize risks and defensive potential and compare results with human experts.
  • Building demonstrations of frontier AI cyber capabilities for policy stakeholders.

Requirements

  • Bachelor's degree or an equivalent combination of education, training, and experience.
  • A field of study relevant to the role, demonstrated through coursework, training, or professional experience.
  • Experience appropriate to the internal job level requirements.
  • Ability to lead frontier cybersecurity research, build and manage a research organization, communicate technical findings, and collaborate across research, product, training, security, safeguards, policy, industry, and government stakeholders.

Work Arrangement

The role is based in San Francisco, California, with expected travel at least to Washington, DC. Anthropic currently expects staff to work from one of its offices at least 25% of the time, although some roles may require more office time.

Compensation and Benefits

  • Annual salary: $485,000–$755,000 USD.
  • Competitive compensation and benefits.
  • Optional equity donation matching.
  • Generous vacation and parental leave.
  • Flexible working hours.
  • Office space for collaboration.

Anthropic sponsors visas for this role and states that it will make every reasonable effort to obtain a visa for successful candidates, with support from an immigration lawyer.

More jobs at Anthropic

Similar jobs