Tech Stack
Solid badges = required, outlined = preferred
Responsibilities
- Build statistical and machine learning models to profile and improve GPU utilization, latency, throughput, and overall fleet efficiency.
- Develop forecasting models for inference demand across products, regions, and model families.
- Analyze production workloads to identify latency bottlenecks and capacity constraints.
- Partner with Capacity Systems Engineering to inform infrastructure planning and long-term GPU investment strategies.
- Design experiments and simulations to evaluate scheduling policies, serving strategies, and infrastructure tradeoffs.
Soft Skills
StatisticsCausal AnalysisStatistical InferenceForecastingOperations ResearchCapacity PlanningTime-Series ForecastingQueueing TheoryReinforcement LearningCost OptimizationExecutive Communication
Culture
Mission-DrivenInclusive HiringCross-Functional Teams
Requirements
Required: MS or PhD in Statistics, Computer Science, Operations Research, Applied Mathematics, Economics, or related quantitative discipline
Regions: Us
Get jobs like this in your inbox
Weekly Python, SQL, Distributed Systems hiring trends and salary data — free.
Join 8 engineers getting weekly insights
Get market intelligence in your inbox
Free weekly insights on tech hiring trends, salaries, and in-demand stacks.
Already a subscriber? Sign in
About OpenAI
Industry: saas
Size: large
OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity.
View company profile →