Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 4
Security @ 4
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
Anthropic is seeking a Capabilities Researcher to join the team building Claude Security. The role focuses on identifying security capabilities in frontier models that are ready to build on, measuring their performance, and making them useful to customers who are not security experts.
The researcher will investigate which capabilities are reliable enough to depend on, how they perform in realistic conditions, how they behave when adversaries are involved, and where their limits are. The role combines rapid prototyping, rigorous evaluation, and product development in collaboration with engineers and researchers.
Responsibilities
- Prototype rapidly to define the AI frontier for cybersecurity work.
- Design evaluations that measure model performance on work security teams actually do.
- Build the datasets, harnesses, and scoring systems required for evaluations.
- Engage with the cybersecurity community to help define where AI can make the greatest impact.
- Work with engineers and researchers to operationalize promising capabilities into products customers can rely on.
- Track how model capabilities for security are changing and determine what that means for future product development.
- Share findings that inform product direction and partner with product leadership on priorities.
- Design scaffolding, tooling, and defaults that help non-experts use strong model capabilities effectively.
Requirements
- Deep expertise in one or more security domains, such as vulnerability research, exploit development, reverse engineering, malware analysis, incident response, or offensive security.
- Experience building AI-powered tools or capabilities for security work.
- Ability to move quickly from an idea to a working prototype and abandon approaches that do not hold up.
- Comfort designing rigorous evaluations and interpreting results honestly.
- Ability to write and communicate clearly about technical findings.
- At least 7 years of experience in security research, security engineering, or a closely related field.
- Bachelor's degree or an equivalent combination of education, training, and experience.
- Education or professional experience in a field relevant to the role.
Preferred Qualifications
- Published research, CTF results, CVEs, or open-source security tooling.
- Experience with model evaluation, benchmarking, or red teaming.
- Experience building agentic applications.
- Familiarity with the safety considerations of AI in security contexts.
Compensation
- Annual salary: $405,000–$485,000 USD.
Benefits
- Hybrid policy requiring staff to work from an Anthropic office at least 25% of the time; some roles may require more office time.
- Competitive compensation and benefits.
- Optional equity donation matching.
- Generous vacation and parental leave.
- Flexible working hours.
- Office space for collaboration.
- Visa sponsorship is available, although eligibility depends on the role and candidate.
More jobs at Anthropic
TPM Manager, Infrastructure
Anthropic · New York City, United States, San Francisco, United States
USD 365,000-565,000 per year
Manager, Technical Deployment (Financial Services)
Anthropic · New York City, United States
USD 240,000-450,000 per year
Research Manager, Biological Safety
Anthropic · San Francisco, United States
USD 405,000-485,000 per year
Senior Manager, IT SOX
Anthropic · San Francisco, United States
USD 230,000-300,000 per year
Technical Program Manager, Enterprise Readiness
Anthropic · New York City, United States, San Francisco, United States
USD 365,000-435,000 per year
Similar jobs
Safeguards Enforcement Analyst, Conventional Weapons
Anthropic · Washington, United States, New York City, United States, San Francisco, United States
USD 245,000-330,000 per year
Product Manager, Multi-Cloud Trust & Safety
Anthropic · New York City, United States, San Francisco, United States, Seattle, United States
USD 305,000-385,000 per year
Product Manager, Claude Science
Anthropic · New York City, United States, San Francisco, United States, Seattle, United States
USD 305,000-385,000 per year
Applied AI Architect, Cyber
Anthropic · New York City, United States, San Francisco, United States
USD 240,000-315,000 per year
Cyber Evaluations Engineer
Anthropic · Washington, United States, San Francisco, United States
USD 300,000-405,000 per year
Lead Product Designer, Enterprise
Perplexity AI · New York City, United States, San Francisco, United States
USD 180,000-300,000 per year
Senior Security Assurance Engineer
GitLab · United States
USD 139,200-196,000 per year
Senior Staff Forward Deployed Engineer, AI
SentinelOne · United States
USD 184,000-253,000 per year