Tech Stack
Responsibilities
- Execute the roadmap for the inference platform including request handling, rate limits, usage controls, and reliability
- Coordinate onboarding, launch readiness, and rollout for new models and capacity with model providers and internal teams
- Drive execution metrics around latency, throughput, uptime, and cost-efficiency
- Run the operating model for model-release and optimization programs across performance, infrastructure, and product teams
- Partner with GPU capacity and compute teams to reconcile decisions against cost, capacity, and vendor constraints
Culture
Cross-Functional TeamsAgile/ScrumFast-Paced
Requirements
Regions: Us
Get jobs like this in your inbox
Weekly TypeScript hiring trends and salary data — free.
Join 8 engineers getting weekly insights
Get job market intelligence in your inbox
Free weekly insights on tech hiring trends, salaries, and in-demand stacks.
Already a subscriber? Sign in
About perplexity
Industry: saas
Size: small
Perplexity is an AI-powered conversational search engine and answer engine that delivers fast, accurate answers to questions.
View company profile →Similar Jobs
Member of Technical Staff (Answer Quality & Evals)
perplexity · San Francisco
Product Technical Program Management - Weights & Biases
CoreWeave · San Francisco, CA / Remote - US
Remote
$188k
Member of Technical Staff (Model Behavior)
perplexity · San Francisco
Member of Technical Staff - Technical PM
Basis AI · New York Office
Member of Technical Staff (Data Scientist, Evals)
perplexity · San Francisco