AI Safety Policy Evaluator, Violence & Threats | Seattle Onsite

Posted

handshakeSeattle, WAfull-time

Tech Stack

Responsibilities

  • Evaluate user requests and AI model responses involving violence, weapons, threats, and dark fiction
  • Distinguish fictional, educational, historical, and defensive violence from real-world intent to harm
  • Select the most defensible classification when a case is genuinely ambiguous and write concise rationales
  • Write and refine adversarial or borderline prompts that probe where a model draws the line
  • Identify policy gaps, contradictions, and emerging edge cases and raise them with policy teams

Benefits

  • 401k
  • Gym Membership
  • Health Insurance
  • Learning Budget

Culture

Collaborative SpaceFast-PacedHigh Growth

Requirements

Regions: Us

Get jobs like this in your inbox

Weekly AWS, Express, Git hiring trends and salary data — free.

Join 8 engineers getting weekly insights

Get job market intelligence in your inbox

Free weekly insights on tech hiring trends, salaries, and in-demand stacks.

Already a subscriber? Sign in