Agent Post-Training, Computer Use Research

Posted

OpenAISan Franciscofull-time

Tech Stack

Responsibilities

  • Design and run experiments to improve agentic model behavior for complex computer use, including desktop and browser.
  • Own end-to-end improvements to the post-training stack, including RL, data pipelines, graders, reward signals, evals, diagnostics, and model-behavior analysis.
  • Build evals and environments that expose new model failures and translate them into training data, product fixes, or new research directions.
  • Partner with product teams to understand user needs and translate product signals into model improvements.
  • Improve the machinery for large-scale training and launch, focusing on experiment velocity, reliability, observability, reproducibility, cost, latency, and production readiness.

Culture

Mission-DrivenCustomer-ObsessedCross-Functional TeamsInclusive Hiring

Requirements

Regions: Us

Get jobs like this in your inbox

Weekly AWS, Go, Next.js hiring trends and salary data — free.

Join 8 engineers getting weekly insights

Get job market intelligence in your inbox

Free weekly insights on tech hiring trends, salaries, and in-demand stacks.

Already a subscriber? Sign in