Site Reliability Engineer (US - Central/Eastern Timezone)

Added
20 days ago
Type
Full time
Salary
Salary not provided

Related skills

gitops terraform github actions linux aws

๐Ÿ“‹ Description

  • Turn fast-growing stateful systems into a predictable automated platform.
  • Reduce ops stress with safe automation for traffic-heavy workloads.
  • Build tooling and patterns to scale with less human effort.
  • Tackle scale: petabytes of data, thousands of cores, multi-region AWS.
  • Operate EKS with Karpenter, Cilium, and ArgoCD GitOps deployments.
  • Manage multi-AWS accounts: provisioning, networking, IAM.

๐ŸŽฏ Requirements

  • Deep hands-on Kubernetes production experience (EKS preferred).
  • Strong AWS production infra experience across multiple accounts (IAM, networking).
  • Terraform/Terragrunt at scale: modules and state management.
  • Linux systems knowledge: disks, memory, networking, failure modes.
  • Experience supporting stateful systems (databases, queues, storage).
  • Ability to debug and reason about performance and reliability in production.
  • Comfortable owning systems end-to-end, including on-call.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest โ€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs โ†’