Senior Site Reliability Engineer

Added
18 days ago
Location
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

datadog kubernetes machine learning distributed systems autoscaling
JobCopilot logo
Meet JobCopilot: Your Personal Al Job Hunter
Automatically Apply to Your Dream Jobs While You Sleep
Try it now →

📋 Description

  • Own ML inference infra powering GPU AI requests.
  • Deploy and run ML workloads; ensure reliability and scalability.
  • Design cloud-agnostic infra for current and future AI workloads.
  • Infra lifecycle: architecture, deployment, monitoring, incidents; Datadog.
  • Build load balancing, autoscaling, queuing, and orchestration.
  • Collaborate with ML, Product, Mobile, and Web teams to speed deployment.

🎯 Requirements

  • Experience designing and operating large-scale distributed systems.
  • Hands-on with load balancing, autoscaling, queues and traffic management.
  • Worked on low-latency real-time backends; optimize perf/throughput.
  • Built resilient, redundant infra that handles failures gracefully.
  • Designed platforms that deploy containerised workloads at high scale.
  • Fluent in English; French not required.

🎁 Benefits

  • Work flexibly from core European countries.
  • In-person onboarding in Paris; quarterly team meetups.
  • 30 days annual leave plus local public holidays.
  • Equity package with stock options.
  • €1,000 home office grant or €400/month coworking stipend.
  • €1,000 annual learning budget and private health insurance.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest — finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs →