AI Safety Policy Evaluator, Violence & Threats | Seattle Onsite
Posted
Tech Stack
Responsibilities
- Evaluate user requests and AI model responses involving violence, weapons, threats, and dark fiction
- Distinguish fictional, educational, historical, and defensive violence from real-world intent to harm
- Select the most defensible classification when a case is genuinely ambiguous and write concise rationales
- Write and refine adversarial or borderline prompts that probe where a model draws the line
- Identify policy gaps, contradictions, and emerging edge cases and raise them with policy teams
Benefits
- 401k
- Gym Membership
- Health Insurance
- Learning Budget
Culture
Collaborative SpaceFast-PacedHigh Growth
Requirements
Regions: Us
Get jobs like this in your inbox
Weekly AWS, Express, Git hiring trends and salary data — free.
Join 8 engineers getting weekly insights
Get job market intelligence in your inbox
Free weekly insights on tech hiring trends, salaries, and in-demand stacks.
Already a subscriber? Sign in
About handshake
Industry: saas
Size: large
Handshake AI partners with leading AI research labs to make models safer and more robust through red teaming operations.
View company profile →