Site Reliability Engineer III

Added
1 hour ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

terraform postgresql mysql kubernetes gcp

πŸ“‹ Description

  • Consolidate our Terraform, which has grown into inconsistent patterns across our infrastructure and
  • Normalize environments, improve build and deploy automation in GitHub Actions, and add drift
  • Apply overdue patches and upgrades across our Cloud SQL databases and application runtimes.
  • Right-size compute and database workloads for growth, including connection pooling and scaling
  • Evaluate our Kubernetes architecture as we grow, including whether and when to move to a
  • Improve monitoring and observability in Datadog and Cloud Monitoring so we catch issues before they

🎯 Requirements

  • Bachelor's degree at a minimum.
  • 5+ years of experience in SRE, DevOps, or infrastructure engineering, with real ownership of
  • Deep hands-on Terraform experience, including structuring modules and managing state across
  • Strong working knowledge of GCP, including GKE, Cloud SQL (MySQL and PostgreSQL), IAM, networking
  • Production Kubernetes experience, including autoscaling, resource management, and judgment about
  • Hands-on experience building monitoring, alerting, and dashboards with tools like Datadog or Cloud

🎁 Benefits

  • Experience as an early or first SRE hire.
  • Experience refactoring or consolidating a large, organically grown Terraform codebase.
  • Experience improving observability from a less mature baseline.
  • Experience in a HIPAA-regulated or other compliance-driven environment.
  • CI/CD experience with GitHub Actions.
  • Experience running Django applications or Airflow in production on Kubernetes.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’