Tech Stack
Responsibilities
- Design and run experiments to improve agentic model behavior for complex software and plugins.
- Own end-to-end improvements to the post-training stack, including RL, data pipelines, graders, reward signals, evals, and diagnostics.
- Build evals and environments to expose model failures, then turn those failures into training data, product fixes, or new research directions.
- Partner with product teams to translate user needs into model improvements.
- Improve the machinery for large-scale training and launch, focusing on experiment velocity, reliability, and cost.
Culture
Cross-Functional TeamsCustomer-ObsessedMission-Driven
Requirements
Regions: Us
Get jobs like this in your inbox
Weekly AWS, Git, Go hiring trends and salary data — free.
Join 8 engineers getting weekly insights
Get job market intelligence in your inbox
Free weekly insights on tech hiring trends, salaries, and in-demand stacks.
Already a subscriber? Sign in
About OpenAI
Industry: artificial intelligence
Size: large
OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity.
View company profile →Similar Jobs
Agent Post-Training, Computer Use Research
OpenAI · San Francisco
Agent Post-Training, Artifacts Research
OpenAI · San Francisco
Agent Post-Training, Context Research
OpenAI · San Francisco
Agent Post-Training, API & Power Users
OpenAI · San Francisco
Agent Post-Training, Frontier Evals and Environments Research
OpenAI · San Francisco