Staff Site Reliability Engineer

Added
2 hours ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

terraform aws grafana kubernetes ci/cd

๐Ÿ“‹ Description

  • Owning the architecture and evolution of our Kubernetes platform for the Dublin/EMEA team โ€” cluster
  • Leading the technical design and rollout of self-service platform capabilities, golden paths, and
  • Driving the platform's readiness to host AI-driven internal workflows reliably at scale
  • Taking ownership of the hardest, highest-ambiguity reliability problems: complex production
  • Participating in a follow-the-sun on-call rotation and driving root-cause resolution and postmortems
  • Setting technical standards for SLIs/SLOs, error budgets, and on-call practice

๐ŸŽฏ Requirements

  • 8+ years of experience in Site Reliability Engineering, Platform Engineering, or Infrastructure
  • Deep, hands-on expertise in Kubernetes at production scale โ€” architecture, multi-tenancy
  • Experience running large-scale internal infrastructure platforms in a public cloud, preferably AWS
  • Strong expertise in cloud-native architectures, infrastructure-as-code (Terraform), and CI/CD
  • Experience building or operating infrastructure that hosts AI/ML or agentic workflows โ€” model
  • Deep experience with observability platforms and monitoring tools (Grafana, Splunk, APM, or

๐ŸŽ Benefits

  • Equity (where applicable) and bonus
  • Comprehensive healthcare coverage
  • Financial benefits including paid time off and parental leave
  • Immersive, in-person onboarding experience
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest โ€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs โ†’