Site Reliability Engineer

Added
28 days ago
Type
Full time
Salary
Salary not provided

Related skills

helm grafana prometheus kubernetes opentelemetry

šŸ“‹ Description

  • Design, implement, and manage on‑premise Kubernetes infrastructure.
  • Build cloud‑native platforms on‑premises for scalable services.
  • Establish observability with Grafana, Prometheus, and distributed tracing.
  • Architect secure, multi‑tenant clusters with policy‑as‑code and zero‑trust networking.
  • Develop and maintain MLOps platforms to deploy and monitor ML models.
  • Collaborate with Security teams on supply chain security and runtime protection.

šŸŽÆ Requirements

  • Scripting: Python, Go, Rust or Bash/Shell for automation.
  • Experience with GitOps workflows and CI/CD automation.
  • Kubernetes production experience, custom controllers/operators, service mesh (Istio/Linkerd).
  • Cloud‑Native Tech: Helm, ArgoCD, Flux; container security tools like Falco.
  • Observability: Grafana, Prometheus, Loki, Tempo, OpenTelemetry; dashboards and SLIs/SLAs.
  • Networking: strong networking concepts and security.

šŸŽ Benefits

  • Focus on outcomes, not time‑tracking.
  • Competitive compensation and VSOP options.
  • Relocation support.
  • Social and education allowances.
  • Regular company events and all‑hands across Europe.
  • Infraduction onboarding to learn our stack and tools from day one.

🚚 Relocation support

Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest — finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs →