Senior Site Reliability Engineer

Added
10 days ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

terraform linux aws grafana prometheus
JobCopilot logo
Meet JobCopilot: Your Personal Al Job Hunter
Automatically Apply to Your Dream Jobs While You Sleep
Try it now β†’

πŸ“‹ Description

  • 100% hands-on across infra and software for a multi-cloud, multi-region platform
  • Support and inform evolution of major systems to scale from 25B to 50B daily requests
  • Collaborate with cross-team groups to ensure reliable delivery to clients
  • Build, maintain, and support core applications
  • Improve tooling and automation to minimize manual work and incidents
  • Monitor capacity, performance, and troubleshoot issues

🎯 Requirements

  • 4+ years experience as SRE or Software Engineer with AWS/GCP
  • Experience building tooling/automation and diagnosing performance issues
  • Experience architecting large-scale observability platforms (Prometheus, Thanos, Grafana, Loki, Tempo)
  • On-call experience, improving monitoring/alerting and creating runbooks
  • Infrastructure as code experience; Terraform is a major plus
  • Kubernetes experience with EKS/GKE

🎁 Benefits

  • Medical, dental, and vision coverage
  • Competitive compensation and bonus potential
  • Flexible/remote-friendly work arrangements
  • Inclusive, diverse culture and equal opportunity employer
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’