Staff Site Reliability Engineer

Added
12 days ago
Type
Full time
Salary
Salary not provided

Related skills

azure production terraform aws kubernetes
JobCopilot logo
Meet JobCopilot: Your Personal Al Job Hunter
Automatically Apply to Your Dream Jobs While You Sleep
Try it now β†’

πŸ“‹ Description

  • Fully remote SRE leadership shaping reliability across AI-driven production environments.
  • Architect, deploy, operate scalable AWS-based infrastructure with Kubernetes, observability, and
  • Lead reliability practices, CI/CD pipelines, and incident response across multiple streams.
  • Collaborate with product, data, and ML teams to incorporate telemetry and operational data.
  • Productize ML workloads with MLOps, monitoring, and lifecycle management.
  • Ensure security, compliance, and enterprise-grade reliability throughout the stack.

🎯 Requirements

  • Extensive hands-on SRE/Production engineering experience in large-scale environments.
  • Deep AWS or Azure expertise with modern cloud-native architectures.
  • Strong Kubernetes production experience (migration, scaling, security hardening).
  • Advanced Infrastructure-as-Code experience (Terraform or equivalent).
  • End-to-end CI/CD design and optimization skills.
  • Observability expertise across distributed systems and data platforms.

🎁 Benefits

  • Full-time, permanent employment.
  • Fully remote position within European time zones.
  • Ownership and influence over reliability practices and architecture.
  • Work across cloud, Kubernetes, distributed systems, data platforms, and ML ops.
  • Exposure to enterprise-scale deployments and AI/ML specialists.
  • Collaborative environment with experienced engineers and architects.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’