Site Reliability Engineer

Added
4 days ago
Type
Full time
Salary
Salary not provided

Related skills

devops grafana python splunk new relic

πŸ“‹ Description

  • Own the day-to-day operation of a monitoring platform, including dashboards, monitors, log
  • Proactively sift through logs, error tracking, traces, and metrics to find failures, regressions
  • Tune alert thresholds, monitor logic, and notification routing to reduce noise and false positives
  • Instrument new and existing services with meaningful metrics, structured logs, and distributed
  • Write post-incident reviews, track remediation items to completion, and feed lessons learned back
  • Define, measure, and report on service level indicators for key customer-facing services.

🎯 Requirements

  • 3+ years of experience in site reliability engineering, DevOps, platform engineering, or a
  • Hands-on experience administering and building in platform such as New Relic, Grafana, Splunk, or
  • Strong troubleshooting and root-cause analysis skills, with a demonstrated ability to work through
  • Working proficiency in at least one scripting or programming language such as Python
  • Experience operating services in a major cloud provider (AWS, GCP, or Azure), with a solid grasp of
  • Familiarity with infrastructure as code (Terraform, CloudFormation, or Pulumi) and CI/CD tooling

🎁 Benefits

  • Health and dental insurance
  • Meal and food allowance
  • Childcare assistance
  • Extended paternity leave
  • Partnership with gyms and health and wellness professionals via Wellhub (Gympass) TotalPass
  • Profit Sharing and Results Participation (PLR)
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’