Agent Post-Training, Context Research

Posted

OpenAISan Franciscofull-time

Tech Stack

Responsibilities

  • Design and run experiments that improve scaling of compute on context.
  • Own end-to-end improvements to the post-training stack, including RL, data pipelines, graders, reward signals, evals, diagnostics, and model-behavior analysis.
  • Build evals and environments that expose the next set of model failures, then turn those failures into training data, product fixes, or new research directions.
  • Partner with Codex and ChatGPT product teams to understand what users need and translate product signal into model improvements.
  • Work on early-training and alignment interventions, including data mixtures, objectives, synthetic data, and eval loops that shape downstream agent behavior.

Culture

Mission-DrivenCustomer-ObsessedImpact-OrientedAutonomous Teams

Requirements

Regions: Us

Get jobs like this in your inbox

Weekly AWS, Go, Next.js hiring trends and salary data — free.

Join 8 engineers getting weekly insights

Get job market intelligence in your inbox

Free weekly insights on tech hiring trends, salaries, and in-demand stacks.

Already a subscriber? Sign in