Tech Stack
Responsibilities
- Define and drive the end-to-end infrastructure architecture for Deepgram's AI/ML workloads across production inference and research training.
- Design multi-cloud and hybrid infrastructure strategies that balance performance, reliability, cost, and vendor flexibility.
- Architect compute orchestration systems that efficiently schedule and manage GPU and CPU workloads across heterogeneous infrastructure.
- Design storage architectures that handle massive datasets for speech and audio ML, from high-throughput training to low-latency model serving.
- Lead capacity planning across all infrastructure dimensions, modeling growth and ensuring Deepgram can scale ahead of demand.
Culture
AI-First MindsetFast-PacedExperimentationAdaptabilityContinuous LearningWork-Life BalanceFlexible ScheduleLearning/Education StipendInternal MobilityMentorship ProgramTech TalksConference BudgetErg/Affinity GroupsInclusive Hiring
Requirements
Regions: Us
Get jobs like this in your inbox
Weekly AWS, Git, Kubernetes hiring trends and salary data — free.
Join 8 engineers getting weekly insights
Get job market intelligence in your inbox
Free weekly insights on tech hiring trends, salaries, and in-demand stacks.
Already a subscriber? Sign in
About Deepgram
Industry: saas
Size: medium
Deepgram is a leading voice AI platform providing real-time APIs for speech-to-text, text-to-speech, and production-grade voice agents.
View company profile →Similar Jobs
Solutions Architect - HPC/AI/ML
CoreWeave · Livingston, NJ / New York, NY / Sunnyvale, CA / Bellevue, WA
$165k
Machine Learning Infrastructure Tech Lead
reducto · San Francisco Office
HPC/ GPU Cluster Architect
sfcompute · San Francisco, CA
Senior Software Architect
sevenai · Boston, MA
Senior Solutions Architect - Weights & Biases
CoreWeave · Livingston, NJ / New York, NY / San Francisco, CA / Sunnyvale, CA / Bellevue, WA
$180k