Senior Engineering Manager, Site Reliability

Added
18 days ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

datadog aws grafana prometheus kubernetes

๐Ÿ“‹ Description

  • Lead a team focused on incident management, observability, operational readiness, and reliability engineering
  • Translate reliability strategy into plans, ownership, and measurable outcomes
  • Partner with engineering leaders to establish reliability expectations
  • Balance immediate operations with durable improvements that reduce risk
  • Develop engineers and leaders who can independently own complex reliability initiatives
  • Drive cross-functional collaboration to protect customers and the business

๐ŸŽฏ Requirements

  • 5+ years reliability engineering management
  • 7+ years in software engineering, site reliability, or infrastructure
  • Direct experience leading an SRE or production engineering function
  • Strong depth in distributed systems and cloud
  • Experience leading high-severity incident response at scale
  • Proven ability to translate strategy into focused plans with ownership

๐ŸŽ Benefits

  • Competitive compensation and equity
  • 401(k) with company match
  • Comprehensive health coverage (medical, dental, vision)
  • Paid time off, holidays, and parental leave
  • Employee assistance program and wellness resources
  • Remote-friendly with regular on-site collaboration in the US
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest โ€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs โ†’