Staff Site Reliability Engineer

Added
10 days ago
Type
Full time
Salary
Upgrade to Premium to se...

πŸ“‹ Description

  • Define, measure, and continuously improve the reliability, availability, and performance of CDI
  • Establish and drive best practices for incident management, root cause analysis, postmortems, and
  • Build and evolve comprehensive monitoring, logging, tracing, and alerting capabilities to enable
  • Identify operational inefficiencies and develop automation, self-service capabilities, and
  • Design and optimize cloud-native infrastructure and services to support growing business demands
  • Drive Infrastructure as Code (IaC), platform standardization, and deployment automation to improve

🎯 Requirements

  • 12+ years of experience in Site Reliability Engineering (SRE)
  • Deep expertise in Kubernetes and GCP
  • Strong Infrastructure as Code (IaC) experience, preferably with Terraform
  • Solid foundations in Linux system administration and networking
  • Strong working knowledge of distributed systems concepts
  • Advanced proficiency in Go, Python, Java, or Shell

🎁 Benefits

  • 401(k) plan
  • Employer matched retirement savings
  • Flexible time off
  • Medical, dental, vision, and life insurance
  • Health saving accounts
  • Professional development budget
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’