Safeguards Enforcement Analyst, Age-Appropriate Design
Anthropic- Compensation
- $245k–$285k Published range
- Location
- Hybrid - US (SF, NYC, DC), 25% in-office Remote eligibility
- Employment
- Full-time Mid-level
About the job
About Anthropic
Anthropic is a public benefit corporation headquartered in San Francisco, focused on creating reliable, interpretable, and steerable AI systems. The company works as a single cohesive team on large-scale research efforts, valuing impact and collaboration.
About the Role
As a Safeguards Enforcement Analyst on the user well-being team, you will build and execute enforcement workflows to keep products safe, with a focus on detecting and mitigating potential harm. Your initial focus will be on age-appropriate design, ensuring consumer products reach the right audiences through detection signals, verification paths, and appeals workflows. You will also support third-party developers using the API and may expand into broader user well-being enforcement over time. This role includes exposure to explicit content and on-call responsibilities.
Key Responsibilities
- Design and architect automated enforcement systems and review workflows that scale effectively while maintaining high accuracy.
- Partner with Engineering and Data Science teams to optimize detection models for policy violations and automated enforcement.
- Review flagged content to drive enforcement and policy improvements.
- Enforce usage policies with a focus on detecting and mitigating potential harmful use of AI systems.
- Work with Legal, Public Policy, and Privacy stakeholders to keep age assurance proportionate, privacy-preserving, and responsive to regulation.
- Support the Safeguards policy design team by providing feedback on policy gaps based on real enforcement scenarios.
- Keep up to date with emerging AI policy enforcement best practices.
- Own Anthropic's layered age assurance approach (self-declaration, behavioral signals, verification, ban appeals) for first-party consumer products.
- Handle adjacent user well-being enforcement where age is a key factor (e.g., sexual content, illicit substances).
Minimum Qualifications
- Experience in trust and safety, online child safety, age assurance, privacy, product policy, or a related field.
- Subject matter expertise in age assurance/verification, age-appropriate design, child online safety, privacy-preserving verification, or content classification for young people.
- Experience driving cross-functional initiatives with Product, Engineering, Legal, and Policy partners, especially where safety, privacy, and usability tradeoffs are involved.
- Experience navigating evolving regulatory landscapes (UK AADC, COPPA, DSA) and enforcement best practices for age assurance, CSAM/CSEM, NCII, and digital well-being.
- A thoughtful perspective on privacy/safety tradeoffs in age verification.
- Strong written communication skills, with experience producing clear briefs for technical and non-technical stakeholders.
- Comfort using data (SQL or similar tools) to measure and inform decisions.
- Excellent judgment and collaboration skills in rapidly evolving environments.
Preferred Qualifications
- Experience building or operating age-gating flows, age estimation signals, underage account detection, or appeals workflows.
- Experience advising or partnering with third-party platforms on deploying safely to younger users.
- Experience working with or evaluating third-party age verification providers.
- A deep interest in AI safety and responsible technology development.
- Experience writing effective prompts for generative AI systems in a content review or enforcement context.
Compensation & Benefits
Annual salary: $245,000—$285,000 USD. Anthropic offers competitive compensation and benefits, including optional equity donation matching, generous vacation and parental leave, flexible working hours, and a collaborative office space.
Skills & tags
Compare the essentials before you leave: pay, remote scope, employment type, source, and the employer apply destination.