Software Dev Senior Engineer

Added
10 days ago
Type
Full time
Salary
Salary not provided

Related skills

terraform grafana prometheus python kubernetes

๐Ÿ“‹ Description

  • Lead on call response and triage โ€“ help design and lead the 24x7 response team for triage and rally
  • Define, publish, and continuously refine Service Level Indicators (SLIs), Service Level Objectives
  • Own the error management practices, including monitoring and alerting, and surfacing notable errors
  • Drive toil reduction initiatives by identifying and automating repetitive, manual operational work
  • Design and execute chaos engineering programs to proactively uncover reliability weaknesses in our
  • Lead blameless postmortem culture: facilitate incident retrospectives, extract systemic learnings

๐ŸŽฏ Requirements

  • 8+ years of experience in scalable, distributed systems architecture.
  • 3+ years of hands-on Site Reliability Engineering experience, including ownership of SLOs and error
  • 4+ years of experience with Cloud Platforms, including GCP.
  • 4+ years of experience in infrastructure as code (Terraform, AWS CDK).
  • 5+ years of experience in scripting using Python, Shell, or a similar language.
  • 4+ years of experience with orchestration technologies, including Kubernetes.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest โ€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs โ†’