Tech Stack
Responsibilities
- Design and run experiments to improve agentic model behavior for complex computer use, including desktop and browser.
- Own end-to-end improvements to the post-training stack, including RL, data pipelines, graders, reward signals, evals, diagnostics, and model-behavior analysis.
- Build evals and environments that expose new model failures and translate them into training data, product fixes, or new research directions.
- Partner with product teams to understand user needs and translate product signals into model improvements.
- Improve the machinery for large-scale training and launch, focusing on experiment velocity, reliability, observability, reproducibility, cost, latency, and production readiness.
Culture
Mission-DrivenCustomer-ObsessedCross-Functional TeamsInclusive Hiring
Requirements
Regions: Us
Get jobs like this in your inbox
Weekly AWS, Go, Next.js hiring trends and salary data — free.
Join 8 engineers getting weekly insights
Get job market intelligence in your inbox
Free weekly insights on tech hiring trends, salaries, and in-demand stacks.
Already a subscriber? Sign in
About OpenAI
Industry: artificial intelligence
Size: large
OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity.
View company profile →Similar Jobs
Agent Post-Training, Artifacts Research
OpenAI · San Francisco
Agent Post-Training, API & Power Users
OpenAI · San Francisco
Agent Post-Training, Connectors Research
OpenAI · San Francisco
Agent Post-Training, Context Research
OpenAI · San Francisco
Agent Post-Training, Frontier Evals and Environments Research
OpenAI · San Francisco