Manager, Site Reliability Engineer

Added
5 minutes ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

datadog azure ansible terraform aws
JobCopilot logo
Meet JobCopilot: Your Personal Al Job Hunter
Automatically Apply to Your Dream Jobs While You Sleep
Try it now โ†’

๐Ÿ“‹ Description

  • Manage Forge's Site Reliability Engineering team responsible for keeping Forge systems highly
  • Drive strong incident management practices in partnership with engineering teams, including
  • Build, improve, and manage observability infrastructure in partnership with Platform Engineering
  • Improve monitoring coverage and alert quality to reduce noise, shorten time to detect, and support
  • Champion reliability best practices across engineering, including service ownership, operational
  • Contribute to technical design, architecture, automation, infrastructure, and overall team delivery.

๐ŸŽฏ Requirements

  • 5+ years of experience leading a Site Reliability Engineering, DevOps, Cloud Operations, or similar
  • 10+ years of total software engineering, infrastructure, platform, cloud, or production operations
  • Bachelor's degree in Computer Science, Engineering, or a closely related field, or equivalent
  • Experience building, operating, and maintaining large-scale cloud infrastructure and distributed
  • Hands-on experience with observability, monitoring, alerting, incident response, troubleshooting
  • Experience with CI/CD, infrastructure automation, cloud platforms, and operational tooling.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest โ€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs โ†’