Skip to main content

Safeguards Enforcement Analyst, Age-Appropriate Design

Anthropic
Hybrid - US (SF, NYC, DC), 25% in-officeUpdated 30d ago
Compensation
$245k–$285k
Published range
Location
Hybrid - US (SF, NYC, DC), 25% in-office
Remote eligibility
Employment
Full-time
Mid-level
Role family
Security
AI / ML
Apply on job-boards.greenhouse.io
Job actionsApply now
Job actionsApply now

About the job

About Anthropic

Anthropic is a public benefit corporation headquartered in San Francisco, focused on creating reliable, interpretable, and steerable AI systems. The company works as a single cohesive team on large-scale research efforts, valuing impact and collaboration.

About the Role

As a Safeguards Enforcement Analyst on the user well-being team, you will build and execute enforcement workflows to keep products safe, with a focus on detecting and mitigating potential harm. Your initial focus will be on age-appropriate design, ensuring consumer products reach the right audiences through detection signals, verification paths, and appeals workflows. You will also support third-party developers using the API and may expand into broader user well-being enforcement over time. This role includes exposure to explicit content and on-call responsibilities.

Key Responsibilities

  • Design and architect automated enforcement systems and review workflows that scale effectively while maintaining high accuracy.
  • Partner with Engineering and Data Science teams to optimize detection models for policy violations and automated enforcement.
  • Review flagged content to drive enforcement and policy improvements.
  • Enforce usage policies with a focus on detecting and mitigating potential harmful use of AI systems.
  • Work with Legal, Public Policy, and Privacy stakeholders to keep age assurance proportionate, privacy-preserving, and responsive to regulation.
  • Support the Safeguards policy design team by providing feedback on policy gaps based on real enforcement scenarios.
  • Keep up to date with emerging AI policy enforcement best practices.
  • Own Anthropic's layered age assurance approach (self-declaration, behavioral signals, verification, ban appeals) for first-party consumer products.
  • Handle adjacent user well-being enforcement where age is a key factor (e.g., sexual content, illicit substances).

Minimum Qualifications

  • Experience in trust and safety, online child safety, age assurance, privacy, product policy, or a related field.
  • Subject matter expertise in age assurance/verification, age-appropriate design, child online safety, privacy-preserving verification, or content classification for young people.
  • Experience driving cross-functional initiatives with Product, Engineering, Legal, and Policy partners, especially where safety, privacy, and usability tradeoffs are involved.
  • Experience navigating evolving regulatory landscapes (UK AADC, COPPA, DSA) and enforcement best practices for age assurance, CSAM/CSEM, NCII, and digital well-being.
  • A thoughtful perspective on privacy/safety tradeoffs in age verification.
  • Strong written communication skills, with experience producing clear briefs for technical and non-technical stakeholders.
  • Comfort using data (SQL or similar tools) to measure and inform decisions.
  • Excellent judgment and collaboration skills in rapidly evolving environments.

Preferred Qualifications

  • Experience building or operating age-gating flows, age estimation signals, underage account detection, or appeals workflows.
  • Experience advising or partnering with third-party platforms on deploying safely to younger users.
  • Experience working with or evaluating third-party age verification providers.
  • A deep interest in AI safety and responsible technology development.
  • Experience writing effective prompts for generative AI systems in a content review or enforcement context.

Compensation & Benefits

Annual salary: $245,000—$285,000 USD. Anthropic offers competitive compensation and benefits, including optional equity donation matching, generous vacation and parental leave, flexible working hours, and a collaborative office space.

Skills & tags

What you can verify before applying

Compare the essentials before you leave: pay, remote scope, employment type, source, and the employer apply destination.