Site Reliability Engineering Lead

Added
5 hours ago
Type
Full time
Salary
Salary not provided

Related skills

linux networking distributed systems storage incident management

πŸ“‹ Description

  • Build and develop the SRE team from initial formation through full 24x7x365 production operations
  • Mentor engineers and team leads, develop successors, and build an organization capable of operating
  • Work with leadership to forecast staffing requirements as the platform grows from initial
  • Define the operating model for an SRE organization, including staffing and coverage model
  • Establish clear operational interfaces with Datacenter Operations, engineering teams, vendors, and
  • Establish and continuously improve production readiness standards, runbooks, operational

🎯 Requirements

  • Significant experience leading or building an SRE, Production Engineering, Infrastructure
  • Experience taking a new or rapidly evolving platform through production readiness, launch
  • Experience building and operating sustainable 24x7x365 production support or on-call organizations.
  • Strong understanding of modern Site Reliability Engineering principles, including SLOs, incident
  • Strong incident leadership experience, including managing high-severity, multi-team production
  • A strong systems engineering background with a working understanding of Linux, networking, storage

🎁 Benefits

  • Competitive salary
  • Flexible working
  • Medical coverage
  • Dental coverage
  • Vision coverage
  • Flexible Spending Accounts (FSAs)
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’