Safeguards Enforcement Analyst, Ban Evasion & Recidivism
Anthropic- Compensation
- $245k–$285k Published range
- Location
- Hybrid - San Francisco, CA | New York City, NY | Washington, DC Remote eligibility
- Employment
- Full-time Mid-level
About the job
About the role
As a Safeguards Enforcement Analyst on the account abuse team, you'll build and execute enforcement workflows that keep our products safe, with a focus on detecting and mitigating potential harm. Your initial focus will be recidivism: a ban that an actor can evade in five minutes isn't enforcement — it's friction. You'll own detecting when banned actors return, linking accounts across identities, and closing the re-registration paths that matter most. The mandate includes our highest-stakes populations, including preventing evasion of child-safety enforcement bans, where the cost of a missed return is unacceptable. This position may expand into broader areas of enforcement over time.
Key responsibilities
- Investigate evasion clusters end to end — from a single appeal or signal anomaly to the full linked actor network
- Convert individual findings into durable systemic controls and detection proposals
- Operationalize re-registration controls for high-severity ban populations
- Partner with Engineering and Data Science teams on account-linking signals to connect returning actors across identities
- Build the recidivism measurement framework: how often banned actors return, how fast we catch them, and which controls reduce return rates
- Author playbooks for contractor-supported evasion review with QA against your own gold standard
- Keep up to date with emerging AI policy enforcement best practices, and use these to inform our decision-making and workflows
Minimum qualifications
- Experience investigating ban evasion, multi-accounting, or repeat fraud actors at a platform with adversarial users
- Fluency in SQL and comfort building your own analyses across large account and event datasets
- Experience working with fraud or identity-linking signals and a working understanding of their precision/recall tradeoffs
- Rigor about evidence standards — comfort with the asymmetric cost of false positives in severe-harm enforcement
- A track record of turning one-off investigations into repeatable detection logic and policy
- Strong written communication skills, with experience producing clear briefs and recommendations for technical and non-technical stakeholders
- Excellent judgment and the ability to collaborate with team members while navigating rapidly evolving priorities and workstreams
Preferred qualifications
- Experience using payment or network risk signals in an enforcement context
- Experience with child-safety or other high-severity integrity enforcement
- Experience collaborating directly with detection engineering or data science teams on rule deployment
- A deep interest in AI safety and responsible technology development
- Experience writing effective prompts for generative AI systems in a content review or enforcement context
Logistics
- Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience
- Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience
- Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position
- Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices.
- Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.
How we're different
We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills.
Come work with us!
Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues.
Skills & tags
Compare the essentials before you leave: pay, remote scope, employment type, source, and the employer apply destination.