Senior Site Reliability Engineer

Added
1 day ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

terraform aws grafana prometheus kubernetes

πŸ“‹ Description

  • Lead discovery and design of reliability projects.
  • Shape platform architecture and tooling roadmaps.
  • Define SLOs, SLIs, error budgets, alerting, observability.
  • Operate and scale Kubernetes and container infra.
  • Build cloud infra with AWS; emphasize reliability and scalability.
  • Develop infrastructure as code with Terraform and CI/CD workflows.

🎯 Requirements

  • Solid SRE/DevOps/platform engineering experience.
  • Production Kubernetes with Docker and container ecosystem.
  • Production cloud infra with AWS or similar provider.
  • Deep Terraform and IaC experience.
  • Experience with SLOs/SLIs, alerting, incident mgmt.
  • Observability with OpenTelemetry, Grafana, Prometheus, or similar.

🎁 Benefits

  • 100% remote work, work from anywhere.
  • Async-first environment with flexible hours.
  • Flexible PTO and 16 weeks parental leave.
  • Budget for co-working, learning, wellness; gym memberships.
  • Mental health support services.
  • Stock options; home office budget and IT equipment.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’