Safeguards Enforcement Lead, User Well-Being
Posted
$285,000 USD
Tech Stack
Responsibilities
- Manage a team of individual contributors across multiple policy areas under the User Well-Being banner
- Serve as the primary point of contact for review partners conducting content review, including onboarding, training, and quality assurance
- Design and improve enforcement workflows to scale effectively as volume grows while maintaining accuracy
- Partner with Engineering and Data Science teams to optimize detection models and automated enforcement systems
- Coordinate reporting obligations to relevant external bodies (e.g., NCMEC) in accordance with applicable law and Anthropic policy
Benefits
- Equity
- Gym Membership
- Learning Budget
- Parental Leave
Culture
Cross-Functional TeamsInclusive HiringFlexible Hours
Requirements
Required: Bachelor’s degree or an equivalent combination of education, training, and/or experience
Regions: Us
Get jobs like this in your inbox
Weekly AWS, Git, Python hiring trends and salary data — free.
Join 8 engineers getting weekly insights
Get job market intelligence in your inbox
Free weekly insights on tech hiring trends, salaries, and in-demand stacks.
Already a subscriber? Sign in
About Anthropic
Industry: saas
Size: medium
Anthropic is an AI safety and research company that builds reliable, beneficial, and steerable AI systems like Claude.
View company profile →Compensation
Base salary: $285,000 USD
Similar Jobs
Safeguards Enforcement Lead, Cyber Harms
Anthropic · San Francisco, CA | New York City, NY | Washington, DC
$285k
Data Scientist, Safeguards
Anthropic · New York City, NY; San Francisco, CA; Seattle, WA
$285k
Technical Cyber Threat Investigator
Anthropic · Remote-Friendly (Travel-Required) | San Francisco, CA | Washington, DC
Remote
$230k
Security Controls Assurance Lead
Anthropic · San Francisco, CA | New York City, NY | Washington, DC
$270k
Data Scientist, Policy
Anthropic · New York City, NY; San Francisco, CA | New York City, NY; Washington, DC
$285k