Senior Site Reliability Engineer, Observability

Added
16 days ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

azure terraform powershell aws windows
JobCopilot logo
Meet JobCopilot: Your Personal Al Job Hunter
Automatically Apply to Your Dream Jobs While You Sleep
Try it now →

📋 Description

  • Design and implement monitoring, alerting, and dashboards in New Relic.
  • Define and implement SLOs/SLIs and error budgets; coach teams.
  • Lead alert noise reduction; tune thresholds; actionable alerts.
  • Optimize observability costs; manage log ingestion and NR config.
  • Improve observability maturity: structured logging, metrics, tracing.
  • Develop Terraform IaC for monitoring resources and governance.

🎯 Requirements

  • 7+ years in SRE/DevOps with observability focus.
  • Hands-on delivery while coaching teams; Agile experience.
  • New Relic (APM/Infra/Logs) and NRQL proficiency.
  • Structured logging, RED/USE metrics, tracing, dashboards.
  • SLOs/SLIs and error budgets for reliability.
  • Incident management tools: Incident.IO, PagerDuty, OpsGenie.
  • Terraform IaC for monitoring infra; governance standards.
  • PowerShell, Azure DevOps; Windows-centric environments; Octopus Deploy.

🎁 Benefits

  • R&R days to rest and recharge.
  • Wellness reimbursement and weekly onsite/virtual programs.
  • Generous vacation policy; flexible time off when needed.
  • Parental leave and family planning benefits.
  • Catered lunches and premium snacks; team events.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest — finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs →