Site Reliability Engineer

Added
15 days ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

datadog azure terraform powershell python
JobCopilot logo
Meet JobCopilot: Your Personal Al Job Hunter
Automatically Apply to Your Dream Jobs While You Sleep
Try it now β†’

πŸ“‹ Description

  • Own the availability and performance of production SaaS applications running on Azure (AKS, App
  • Lead troubleshooting and resolution of cloud infrastructure and application issues, including AKS
  • Participate in an on-call rotation (including weekends) and drive incident response from detection
  • Drive improvements to disaster recovery, failover, and incident management processes across
  • Build and maintain automation scripts and monitoring tools to reduce manual toil and streamline
  • Author post-incident reviews (RCAs), identify root causes, and drive preventive action items to

🎯 Requirements

  • 5+ years of relevant experience in Site Reliability Engineering, DevOps, or Cloud Administration
  • Hands-on experience administering Azure environments, including AKS (Kubernetes), core Azure
  • Solid understanding of monitoring, logging, and alerting practices (e.g., Datadog, Azure Monitor
  • Familiarity with networking fundamentals: firewalls, load balancers, VPNs, DNS, and routing.
  • Experience with automation and scripting (PowerShell, Python, or similar).
  • Practical understanding of backup, redundancy, and disaster recovery strategies in cloud

🎁 Benefits

  • Competitive salaries
  • Meaningful bonus program
  • Excellent benefits including healthcare insurance
  • Pension/retirement matching
  • Comprehensive life insurance
  • Employee assistance program
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’