Research Engineer, Cybersecurity RL (Reinforcement Learning)
Anthropic- Compensation
- $300k–$405k Published range · Top quartile for Engineering (582 listings)
- Location
- Hybrid - San Francisco, CA or New York City, NY, at least 25% in office Remote eligibility
- Employment
- Full-time Mid-level
About the job
Anthropic is a public benefit corporation focused on creating reliable, interpretable, and steerable AI systems. The Horizons team leads reinforcement learning (RL) research and development, contributing to every Claude release. The Cybersecurity RL team within Horizons is hiring a Research Engineer to advance model capabilities in secure coding, vulnerability remediation, and defensive cybersecurity.
This role blends research and engineering, involving designing and implementing RL environments, conducting experiments and evaluations, delivering work into production training runs, and collaborating with researchers, engineers, and cybersecurity specialists.
Ideal candidates have experience in cybersecurity research, machine learning, and strong software engineering skills. Strong candidates may also have professional experience in security engineering, fuzzing, detection and response, CTF competitions, cyber ranges, academic research in cybersecurity, familiarity with RL techniques and environments, and LLM training methodologies.
Annual salary: $300,000—$405,000 USD. Anthropic offers competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space.
Location: San Francisco, CA or New York City, NY. Hybrid policy: all staff expected in office at least 25% of the time.
Visa sponsorship is available. Apply via the provided link.
Skills & tags
Compare the essentials before you leave: pay, remote scope, employment type, source, and the employer apply destination.