Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 4
Security @ 4
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
Anthropic is seeking a Capabilities Researcher to join the team building Claude Security. The role focuses on identifying security capabilities in frontier models that are ready to build on, measuring their performance, and making them useful to customers who are not security experts.
The researcher will investigate which capabilities are reliable enough to depend on, how they perform in realistic conditions, how they behave when adversaries are involved, and where their limits are. The role combines rapid prototyping, rigorous evaluation, and product development in collaboration with engineers and researchers.
Responsibilities
- Prototype rapidly to define the AI frontier for cybersecurity work.
- Design evaluations that measure model performance on work security teams actually do.
- Build the datasets, harnesses, and scoring systems required for evaluations.
- Engage with the cybersecurity community to help define where AI can make the greatest impact.
- Work with engineers and researchers to operationalize promising capabilities into products customers can rely on.
- Track how model capabilities for security are changing and determine what that means for future product development.
- Share findings that inform product direction and partner with product leadership on priorities.
- Design scaffolding, tooling, and defaults that help non-experts use strong model capabilities effectively.
Requirements
- Deep expertise in one or more security domains, such as vulnerability research, exploit development, reverse engineering, malware analysis, incident response, or offensive security.
- Experience building AI-powered tools or capabilities for security work.
- Ability to move quickly from an idea to a working prototype and abandon approaches that do not hold up.
- Comfort designing rigorous evaluations and interpreting results honestly.
- Ability to write and communicate clearly about technical findings.
- At least 7 years of experience in security research, security engineering, or a closely related field.
- Bachelor's degree or an equivalent combination of education, training, and experience.
- Education or professional experience in a field relevant to the role.
Preferred Qualifications
- Published research, CTF results, CVEs, or open-source security tooling.
- Experience with model evaluation, benchmarking, or red teaming.
- Experience building agentic applications.
- Familiarity with the safety considerations of AI in security contexts.
Compensation
- Annual salary: $405,000–$485,000 USD.
Benefits
- Hybrid policy requiring staff to work from an Anthropic office at least 25% of the time; some roles may require more office time.
- Competitive compensation and benefits.
- Optional equity donation matching.
- Generous vacation and parental leave.
- Flexible working hours.
- Office space for collaboration.
- Visa sponsorship is available, although eligibility depends on the role and candidate.
More jobs at Anthropic
Salesforce Developer
Anthropic · New York City, United States, San Francisco, United States, Seattle, United States
USD 270,000-345,000 per year
Business Systems Analyst, New Product Introduction
Anthropic · New York City, United States, San Francisco, United States, Seattle, United States
USD 270,000-315,000 per year
Research Engineer, Takeoff Intel
Anthropic · San Francisco, United States
USD 350,000-850,000 per year
Lead, Security Controls Assurance - SOX
Anthropic · Washington, United States, New York City, United States, San Francisco, United States, Seattle, United States
USD 410,000-510,000 per year
Lead Technical Instructor
Anthropic · New York City, United States
USD 270,000-310,000 per year
Similar jobs
Member of Technical Staff (Software Engineer, Infrastructure)
Perplexity AI · United States, Toronto, Canada, Berlin, Germany, London, United Kingdom, Austin, United States, New York City, United States, San Francisco, United States, Seattle, United States
USD 220,000-405,000 per year
Software Engineer, Host Assurance
OpenAI · United States, San Francisco, United States, Seattle, United States
USD 266,000-445,000 per year
Safeguards Enforcement Analyst, Conventional Weapons
Anthropic · Washington, United States, New York City, United States, San Francisco, United States
USD 245,000-330,000 per year
Agent Standards Specialist, Global Affairs
OpenAI · Washington, United States, New York City, United States
USD 171,000-280,000 per year
Application Security Consultant
SentinelOne · United States
USD 132,000-160,000 per year
Senior System Software Engineer - Platform Software
Nvidia · Santa Clara, United States
USD 224,000-356,500 per year
Manager, Software Engineering Productivity and Release Engineering
Nvidia · Santa Clara, United States
USD 224,000-431,200 per year
Principal Security Researcher
GitLab · Canada, United Kingdom, Israel, United States
USD 203,200-275,000 per year