Lead Site Reliability Engineer

Added
12 hours ago
Type
Full time
Salary
Salary not provided

Related skills

datadog terraform bash aws prometheus

πŸ“‹ Description

  • Lead SRE team within core Platform/Production ownership
  • Stay hands-on with incident response and on-call duties
  • Operate and improve Kubernetes clusters and cloud infra
  • Improve observability with dashboards, alerts, logs, traces
  • Automate repetitive toil and refine tooling
  • Contribute to deployment strategies and runbooks

🎯 Requirements

  • 7+ years in SRE/DevOps/operations-heavy roles
  • Experience hiring and coaching engineers
  • Deep experience with production systems and on-call
  • Comfort debugging live systems under pressure
  • Strong cloud infra experience (AWS preferred)
  • Hands-on with Kubernetes and containerized workloads

🎁 Benefits

  • Annual learning and development budget ($1,000)
  • Health/wellness allowance ($150/month)
  • Home office budget ($500)
  • Parental leave (26 weeks primary, 18 weeks secondary)
  • Fertility support up to $10,000
  • Work from anywhere for 4 weeks/year
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’