Research Engineer, Cybersecurity RL (Reinforcement Learning)

USD 300,000-405,000 per year
MIDDLE
✅ Hybrid
✅ Visa Sponsorship

Tech Stack

AI @ 3 LLM Machine Learning @ 3 Reinforcement Learning @ 2 Security @ 3

Details

Anthropic is hiring a Research Engineer for the Cybersecurity Reinforcement Learning team within Horizons. The role focuses on safely advancing model capabilities in secure coding, vulnerability remediation, and other areas of defensive cybersecurity.

The position combines research and engineering. Responsibilities include developing novel approaches, implementing them in code, designing and implementing reinforcement learning environments, conducting experiments and evaluations, delivering work into production training runs, and collaborating with researchers, engineers, and cybersecurity specialists within and outside Anthropic.

Responsibilities

  • Develop research approaches for defensive cybersecurity and safe AI systems.
  • Design and implement reinforcement learning environments.
  • Conduct experiments and evaluations.
  • Deliver research and engineering work into production training runs.
  • Collaborate with researchers, engineers, and cybersecurity specialists.
  • Advance model capabilities in secure coding and vulnerability remediation.

Requirements

  • Experience in cybersecurity research.
  • Experience with machine learning.
  • Strong software engineering skills.
  • Ability to balance research exploration with engineering implementation.
  • Passion for AI's potential and commitment to developing safe and beneficial systems.
  • Bachelor's degree or an equivalent combination of education, training, and/or experience.
  • A field of study relevant to the role, as demonstrated through coursework, training, or professional experience.

Preferred Qualifications

  • Professional experience in security engineering, fuzzing, detection and response, or other applied defensive work.
  • Experience participating in or building Capture the Flag competitions and cyber ranges.
  • Academic research experience in cybersecurity.
  • Familiarity with reinforcement learning techniques and environments.
  • Familiarity with large language model training methodologies.

Benefits

Anthropic offers competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and office space for collaboration. Staff are expected to work from one of Anthropic's offices at least 25% of the time, although some roles may require more office time.

Anthropic sponsors visas, though sponsorship may not be possible for every role or candidate. The company retains an immigration lawyer and makes every reasonable effort to assist with visas when an offer is made.

More jobs at Anthropic

Similar jobs