Tech Stack
Distributed SystemsObservabilityCapacity PlanningMonitoringIncident ManagementSystem ResilienceDeployment SafetyOperational QualityRunbooksAI
Solid badges = required, outlined = preferred
Responsibilities
- Architect and own reliability and infrastructure strategy across multiple teams and services.
- Own the observability, capacity planning, and monitoring strategy for complex distributed systems.
- Define and evolve SLI/SLO frameworks, error budgets, and production readiness standards org-wide.
- Lead incident management practices, escalation design, and drive systemic improvements from post-mortems.
- Champion toil reduction, on-call sustainability, and long-term system resilience.
Soft Skills
Post-MortemsToil ReductionOn-Call SustainabilityDesign ReviewsTechnical Decision-MakingCross-Functional CollaborationWritten CommunicationSRELegal Tech
Benefits
- 401k
- Commuter Benefits
- Dental
- Disability Insurance
- HSA/FSA
- Health Insurance
- Life Insurance
- Parental Leave
- Unlimited PTO
- Vision
Culture
Cross-Functional TeamsTeam LeadershipMission-DrivenFast-PacedInclusive Hiring
Requirements
Regions: Us
Get jobs like this in your inbox
Weekly Distributed Systems, Observability, Capacity Planning hiring trends and salary data — free.
Join 8 engineers getting weekly insights
Get market intelligence in your inbox
Free weekly insights on tech hiring trends, salaries, and in-demand stacks.
Already a subscriber? Sign in
About Legora
Industry: saas
Size: medium
Legora is an AI-native legal tech company that automates and accelerates complex document analysis and workflows for leading global legal teams.
View company profile →Similar Jobs