Senior Site Reliability Engineer

Added
1 hour ago
Type
Full time
Salary
Salary not provided

Related skills

postgresql prometheus kubernetes elasticsearch google cloud platform

πŸ“‹ Description

  • Design, build, and operate GCP infra, Kubernetes, networking, CI/CD, and observability.
  • Diagnose and troubleshoot complex distributed systems at high request volume.
  • Ensure observability and analyze behavior of our stack.
  • Modernize edge, caching, and gateway layers with Fastly and improved observability.
  • Raise reliability with dashboards, alerting, paging standards, and on-call readiness.
  • Make deployments boring: golden paths, readiness checks, safe rollouts, and automation.

🎯 Requirements

  • Based in the United States with overlap with European hours.
  • 5+ years of SRE/on-call experience.
  • Kubernetes for container orchestration in cloud environments.
  • CI/CD pipeline experience.
  • Observability stack expertise (Prometheus, etc.).
  • Comfortable with incidents and high-pressure scenarios.

🎁 Benefits

  • Highly skilled, inspiring, and supportive team.
  • Real infrastructure scale and meaningful, hands-on work.
  • Flexible, trust-based environment with growth opportunities.
  • Global, diverse group of colleagues and customers.
  • Comprehensive health plans and perks.
  • Healthy work-life balance for individuals and families.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’