Staff / Principal
Anthropic- Compensation
- $320k–$485k Published range · Top quartile for Engineering (813 listings)
- Location
- Hybrid - New York City, NY or San Francisco, CA, 25% in office Remote eligibility
- Employment
- Full-time Staff / Principal
About the job
About the role
The Safeguards team is responsible for ensuring Anthropic's models and products are developed and deployed safely. The Review Tooling team builds the systems that humans — and increasingly Claude — use to investigate potential harms and take enforcement actions across Anthropic's first-party products and third-party cloud platforms. You'll own the tools safety investigators rely on, including the platform underneath those tools: data analysis capabilities, privacy-preserving primitives, and a sandbox environment for building new review interfaces and workflows.
Key responsibilities
- Build investigation, review, and enforcement tooling for first-party and third-party platform surfaces, including case queues, investigation views, decision and audit logging, and account-actioning workflows.
- Develop the platform layer of reusable APIs, data storage, and backend services for quickly and safely standing up new review workflows.
- Stand up and run deployments across multiple clouds, including inside cloud-provider partner environments where data must stay in place, keeping deployments consistent through shared pipelines, smoke tests, observability, and alerting.
- Scale review through automation, including enabling reviewers to use Claude effectively and building toward Claude-assisted and Claude-driven review workflows.
- Partner with policy, operations, legal, privacy, and data science stakeholders to translate enforcement and investigation needs into reliable systems that reduce handling time and decision error.
- Instrument tools to surface metrics on queue health, reviewer throughput, and decision quality.
Minimum qualifications
- Technical background in full-stack or platform engineering, with ability to engage deeply in architecture and design discussions.
- Experience shipping internal tools or platforms with demanding operational users, and a track record of improving their workflows measurably.
- Experience working cross-functionally with non-engineering partners such as operations, policy, or legal teams.
- Excellent communication skills, including explaining technical tradeoffs to non-technical stakeholders.
- Care about the societal impacts of AI and want your work to make powerful systems safer.
Preferred qualifications
- 8+ years of industry software engineering experience.
- Experience building data labeling, trust and safety, integrity, fraud, or abuse-prevention tooling, or other systems supporting human review at scale.
- Experience with sensitive or regulated data, including access control, auditability, retention, and data residency.
- Experience across multiple cloud providers or building provider-agnostic infrastructure.
- A product-minded approach to internal users.
Compensation & benefits
Annual salary: $320,000—$485,000 USD. Anthropic offers competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, and flexible working hours.
Work location & visa
Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.
Application instructions
Apply via the Greenhouse job posting. Anthropic recruiters only contact from @anthropic.com addresses; be cautious of other domains.
Skills & tags
Compare the essentials before you leave: pay, remote scope, employment type, source, and the employer apply destination.