Senior Site Reliability Engineer — Observability

Added
14 days ago
Location
Type
Contract
Salary
Upgrade to Premium to se...

Related skills

kubernetes gcp observability platform engineering google cloud platform
JobCopilot logo
Meet JobCopilot: Your Personal Al Job Hunter
Automatically Apply to Your Dream Jobs While You Sleep
Try it now →

📋 Description

  • Design, build, and improve OpenTelemetry pipelines for metrics, traces, and logs.
  • Strengthen observability across services running on GCP and GKE.
  • Integrate telemetry with tools such as Google Cloud Monitoring, Cloud Trace, Managed Service for
  • Define and implement SLIs and SLOs based on customer and service reliability goals.
  • Build actionable burn-rate alerts that identify reliability risks without generating unnecessary
  • Manage observability infrastructure through Terraform, Helm, and Kubernetes.

🎯 Requirements

  • 5+ years of professional experience in Site Reliability Engineering, Platform Engineering
  • Proven, hands-on experience building an OpenTelemetry pipeline in a production GCP environment.
  • Strong experience with Google Kubernetes Engine (GKE).
  • Experience with Google Cloud Monitoring, Cloud Trace, and Managed Service for Prometheus.
  • Strong knowledge of Grafana, Prometheus, distributed tracing, metrics, logging, and telemetry
  • Practical experience defining SLIs and SLOs and implementing multi-window, multi-burn-rate alerting.

🎁 Benefits

  • Remote work within Europe.
  • Strong potential to convert to a permanent role.
  • Fast-moving interview process, no take-home assessment.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest — finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs →