Staff Site Reliability Engineer

earninSet a reliability strategy with AI at the center. Define SLIs, SLOs, and error budgets across critical services. Use AI to surface trends, predict capacity risks, and auto-generate reliability scorecards so teams act on data.Hybridfull_timestaff

$252,000 – $308,000 USD

Tech Stack

Solid badges = required, outlined = preferred

Responsibilities

  • Set a reliability strategy with AI at the center, defining SLIs, SLOs, and error budgets across critical services.
  • Redesign the incident lifecycle around AI-assisted speed, leading high-severity incident response and building AI-driven alert correlation and triage.
  • Improve on-call fundamentally better through automation by building AI agents that draft runbook responses, pull relevant context, and recommend remediation steps.
  • Partner with product engineering teams to embed AI-assisted investigation, alerting, and production readiness into their workflows.
  • Architect for resilience at scale, guiding service designs for graceful degradation, failure isolation, and capacity planning across EarnIn's AWS footprint.

Soft Skills

Slos/SlisError BudgetsIncident CommandBlameless PostmortemsFintechSOC 2PciFinops

Benefits

  • Equity

Culture

Blameless PostmortemsMentorship ProgramWork-Life Balance

Requirements

Regions: Us

Get jobs like this in your inbox

Weekly Python, Go, Datadog hiring trends and salary data — free.

Join 8 engineers getting weekly insights

Get market intelligence in your inbox

Free weekly insights on tech hiring trends, salaries, and in-demand stacks.

Already a subscriber? Sign in