Primary record

Safeguards Enforcement Analyst, Ban Evasion & Recidivism

Anthropic Indexed employerRemote-Friendly, United States; San Francisco, CA | New York City, NY | Washington, DC · San Francisco, California, United States
Source-hosted applyChecked 5h agoContract
Apply at Anthropic

Anthropic receives this application through Greenhouse. Babu Careers does not claim delivery.

Workplace

hybrid

Employment

Contract

Published

Aug 11, 2026

Closes

No date supplied

The role

About Anthropic

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.

About the role

As a Safeguards Enforcement Analyst on the account abuse team, you'll build and execute enforcement workflows that keep our products safe, with a focus on detecting and mitigating potential harm. Your initial focus will be recidivism: a ban that an actor can evade in five minutes isn't enforcement — it's friction. You'll own detecting when banned actors return, linking accounts across identities, and closing the re-registration paths that matter most. The mandate includes our highest-stakes populations, including preventing evasion of child-safety enforcement bans, where the cost of a missed return is unacceptable.

This position may expand into broader areas of enforcement over time. Safety is core to our mission, and you'll help shape policy enforcement so that our users can safely interact with and build on top of our products in a harmless, helpful, and honest way.

Key responsibilities

• Investigate evasion clusters end to end — from a single appeal or signal anomaly to the full linked actor network

• Convert individual findings into durable systemic controls and detection proposals

• Operationalize re-registration controls for high-severity ban populations

• Partner with Engineering and Data Science teams on account-linking signals to connect returning actors across identities

• Build the recidivism measurement framework: how often banned actors return, how fast we catch them, and which controls reduce return rates

• Author playbooks for contract

Requirements

Department: Safeguards (Trust & Safety)