Staff Site Reliability Engineer - Release Engineering

Added
1 hour ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

prometheus kubernetes go observability service mesh

📋 Description

  • Define and scale reliability practices across product engineering.
  • Architect SLOs and error budgets for Plaid teams.
  • Drive progressive delivery and automated safety gates.
  • Turn complex production needs into self-service tooling.
  • Shape deployment systems for fast, safe AI-enabled velocity.
  • Lead responses to incidents and drive post-mortem improvements.

🎯 Requirements

  • 8+ years in backend, SRE, or platform engineering.
  • Proven reliability program design with cross-team adoption.
  • Experience with canary rollouts, metric-gated analysis, automated rollback.
  • Go or similar systems language proficiency.
  • Ability to drive organizational change without formal authority.
  • Experience with Kubernetes, service mesh, Prometheus, or ArgoCD is a strong asset.

🎁 Benefits

  • Equity or commissions dependent on position.
  • Medical, dental, vision, and 401(k) benefits.
  • Accommodations for candidates with disabilities in recruiting.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest — finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs →