Senior Site Reliability Engineer, Observability

Added
17 days ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

azure terraform powershell aws windows
JobCopilot logo
Meet JobCopilot: Your Personal Al Job Hunter
Automatically Apply to Your Dream Jobs While You Sleep
Try it now →

📋 Description

  • Design and implement NR monitoring, alerts, and dashboards across Azure/AWS.
  • Define SLOs/SLIs and error budgets; coach teams on reliability.
  • Lead alert noise reduction; tune thresholds; actionable alerts only.
  • Terraform IaC for observability resources; enforce governance.
  • Support incident management: MTTR/MTTD tracking and postmortems.
  • Collaborate with teams to raise observability maturity and practices.

🎯 Requirements

  • 7+ years in SRE/DevOps with an observability focus.
  • Expert in New Relic APM/Infra/Logs/Synthetics; NRQL.
  • Strong logging, metrics (RED/USE), tracing; dashboards and alerts.
  • Terraform IaC for cloud monitoring; governance patterns.
  • PowerShell scripting; Windows-heavy environments; Azure DevOps; Octopus Deploy.
  • Incident management platforms (Incident.IO, PagerDuty, OpsGenie).

🎁 Benefits

  • Competitive benefits covering physical and mental healthcare.
  • Generous vacation, parental leave, and family planning benefits.
  • Wellness reimbursement and weekly onsite/virtual programs.
  • R&R days to rest and recharge.
  • Catered lunches and fully stocked kitchens.
  • Employee giving match and mobile phone stipend.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest — finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs →