Staff Site Reliability Engineer

Added
14 days ago
Type
Full time
Salary
Salary not provided

Related skills

azure terraform aws kubernetes ci/cd

πŸ“‹ Description

  • Fully remote opportunity for an SRE leader to shape reliability across AI-driven environments.
  • Act as senior technical authority, designing resilient infrastructure and establishing SRE
  • Span cloud infrastructure, Kubernetes, observability, CI/CD, data platforms, and ML systems.
  • Collaborate across platform, product, data, and ML engineering to improve availability
  • Productionize AI workloads and standardize customer environments for enterprise deployments.
  • Influence architecture and technology direction across engineering while emphasizing ownership and

🎯 Requirements

  • Extensive hands-on SRE/Production Engineering experience in large, distributed environments.
  • Proven ability to scale SRE practices in high-growth settings.
  • Deep AWS or Azure cloud experience and cloud-native architectures.
  • Strong Kubernetes production experience (migration, scaling, security hardening).
  • Advanced Infrastructure-as-Code with Terraform or similar.
  • End-to-end CI/CD design and optimization expertise.

🎁 Benefits

  • Full-time, permanent employment.
  • Fully remote within European time zones.
  • Ownership over reliability practices and architecture.
  • Exposure to cloud infrastructure, Kubernetes, data platforms, and ML operations.
  • Work with experienced engineers, product teams, and AI/ML specialists.
  • Opportunity to shape reliability standards as the organization scales.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’