Site Reliability Engineering Lead

Added
3 days ago
Type
Full time
Salary
Salary not provided

Related skills

azure docker ansible terraform grafana

๐Ÿ“‹ Description

  • Build the SRE team from scratch: define roles, hire, onboard and mentor.
  • Define the SRE direction and operational roadmap aligned with product priorities.
  • Lead on-call rotations, incident management, and blameless post-mortems.
  • Define and track SLOs, SLIs, and error budgets with engineering and product.
  • Hands-on contributor during ramp-up, operating critical Azure infrastructure.

๐ŸŽฏ Requirements

  • 3+ years in an SRE role
  • 1+ year leading or coordinating a technical team
  • Proven IT operations or infrastructure engineering experience
  • Hands-on with Azure cloud services
  • Experience with monitoring/observability tools (Prometheus, Grafana, Azure Monitor, ELK)
  • IaC tools such as Terraform or Ansible
  • Docker and Kubernetes (AKS)
  • Define and implement SLOs, SLIs and error budgets in production

๐ŸŽ Benefits

  • Collaborative, innovative culture
  • Opportunity to build a team from scratch
  • Ownership mindset and drive outcomes
  • One Team: focusing on diversity and teamwork
  • Fast decision and delivery culture
  • Resilience: steady under pressure
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest โ€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs โ†’