Research Lead, Training Insights

Posted

AnthropicRemote-Friendly (Travel Required) | San Francisco, CA; San Francisco, CA | New York City, NYRemotefull-timestaff

$850,000 USD

Tech Stack

Responsibilities

  • Develop the strategy and lead execution on measuring and characterizing model capabilities across training and deployment.
  • Drive original research into new evaluation methodologies and lead a small team of researchers and research engineers.
  • Research and build new long-horizon evaluations, develop novel approaches to measuring emerging capabilities, and deepen understanding of capability development.
  • Work across Reinforcement Learning, Pretraining, Inference, Product, Alignment, and Safeguards teams to identify critical gaps in evaluation coverage.
  • Shape the evaluation narrative for model releases and contribute to how Anthropic communicates about its models internally and externally.

Benefits

  • Equity
  • Learning Budget
  • Parental Leave

Culture

Cross-Functional TeamsMission-DrivenImpact-OrientedCollaborative SpaceWork-Life BalanceInclusive HiringWritten-First Culture

Requirements

Required: Bachelor’s degree or an equivalent combination of education, training, and/or experience
Regions: Us

Get jobs like this in your inbox

Weekly AWS, Git, Rust hiring trends and salary data — free.

Join 8 engineers getting weekly insights

Get market intelligence in your inbox

Free weekly insights on tech hiring trends, salaries, and in-demand stacks.

Already a subscriber? Sign in