Skip to main content

Lead, Frontier Red Team (Cyber)

Anthropic
Hybrid - San Francisco, CA (at least 25% in office)Updated 1d ago
Compensation
$485k–$755k
Published range · Top quartile for Security (62 listings)
Location
Hybrid - San Francisco, CA (at least 25% in office)
Remote eligibility
Employment
Full-time
Lead / Manager
Role family
Security
AI / ML
Role skills
Apply on job-boards.greenhouse.io
Job actionsApply now
Job actionsApply now

About the job

About Anthropic

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole.

About the Team

As part of Anthropic’s work on the frontier of cybersecurity, we are forming a new team focused on steering the world through the next generation of cybersecurity. This team will research the impacts of advanced models on security and publish research and build defenses.

About the Role

The Lead will oversee all of Anthropic’s mission-driven research on defending the world in an era of advanced cybersecurity capabilities. They will lead research to uncover offensive and defensive capabilities of Claude, prototype defenses, inform how Claude is trained, and engage with public and government. They will hire and lead a world-class research team, design an AGI cybersecurity research program, and work with other teams.

What You'll Do

  • Be responsible for Anthropic’s overall defensive cybersecurity vision
  • Hire and lead a world-class research team of AGI-pilled cybersecurity researchers and program managers
  • Design an AGI cybersecurity research program that aims to secure the world
  • Lead the team to execute on this research at the scale, speed, and quality of a frontier lab
  • Work with other Anthropic teams, such as Training, Safeguards, and Policy, to address their biggest opportunities and challenges in Anthropic’s cybersecurity mission
  • Identify field partners – maintainers, researchers, companies, governments – to do joint projects with
  • Develop a strategy to publish and share research with the world for maximum impact; shut down and avoid low-impact research or research communication
  • Translate technical findings into compelling demonstrations and artifacts that inform policymakers and the public
  • Proactively identify critical strategic decisions for the company and the AI lab ecosystem, raise them, marshall resources to address them, and help the decisions be made as smoothly as possible

Sample Projects

  • Writing the strategy for how a frontier AI lab can use frontier models and resources to most defend the world
  • Building frameworks and tools that enable AI models to autonomously find and patch vulnerabilities using tens of trillions of tokens
  • Running purple-team simulations where AI defenders compete against AI attackers in network environments
  • Pointing autonomous AI systems at real-world security challenges (bug bounties, CTFs etc.) to characterize risks, defensive potential, and compare to human experts
  • Building demonstrations of frontier AI cyber capabilities for policy stakeholders

Compensation

Annual Salary: $485,000—$755,000 USD

Logistics

Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience. Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience. Years of experience required will correlate with the internal job level requirements for the position. Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.

How we're different

We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time.

Skills & tags

What you can verify before applying

Compare the essentials before you leave: pay, remote scope, employment type, source, and the employer apply destination.