AI Red Teamer, LLM Generalist

Posted

handshakeSeattle, WAfull-time

Tech Stack

Responsibilities

  • Craft creative prompts and multi-turn scenarios to stress-test AI guardrails across diverse risk categories.
  • Discover ways around safety filters, restrictions, and defenses using jailbreak, evasion, and prompt injection techniques.
  • Evaluate and score model responses against structured harm taxonomies and severity rubrics.
  • Document experiments clearly, including what you tried, why you tried it, and what it revealed.
  • Collaborate with engineers, data scientists, and researchers to share findings and strengthen defenses.

Benefits

  • Equity
  • Learning Budget

Culture

Customer-ObsessedCross-Functional TeamsFast-Paced

Requirements

Regions: Us

Get jobs like this in your inbox

Weekly Go, Next.js, Python hiring trends and salary data — free.

Join 8 engineers getting weekly insights

Get market intelligence in your inbox

Free weekly insights on tech hiring trends, salaries, and in-demand stacks.

Already a subscriber? Sign in

About handshake
Industry: saas
Size: large

Handshake powers 25 million job seekers, 1 million+ employers, and 1,600 educational institutions, and has built a fast-growing AI data business supporting frontier AI labs.

View company profile →