Site Reliability Engineer

Added
28 days ago
Type
Full time
Salary
Salary not provided

Related skills

redshift java linux aws snowflake

πŸ“‹ Description

  • Establish the SRE function: define SLIs, SLOs, and error budgets.
  • Build observability: instrument services for availability and latency.
  • Drive data-driven reliability: measure change impact with metrics.
  • Own production reliability end-to-end: design to steady-state operations.
  • Eliminate manual tasks via automation and infrastructure as code.
  • Plan for scale: capacity planning and performance analysis.

🎯 Requirements

  • Bachelor's or Master's in CS, related field, or equivalent.
  • 6+ years software/systems engineering with β‰₯4 years in SRE.
  • 4+ years designing, analyzing, and troubleshooting distributed systems.
  • Unix/Linux internals and networking basics (TCP/IP, DNS, load bal).
  • Experience establishing SRE practices: SLIs/SLOs, monitors, automation.
  • Chaos engineering and incident management experience.

🎁 Benefits

  • High-performance culture focused on accountability and collaboration.
  • Discretionary Time Off with no maximum limits.
  • Industry-leading health, vision, and dental benefits.
  • Competitive compensation package.
  • 16 weeks of fully paid parental leave.
  • Flexible, hybrid work arrangement with wellness programs.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’