AI Infrastructure & Platform Operations Engineer (remote in the EU)

Added
less than a minute ago
Type
Full time
Salary
Salary not provided

Related skills

linux grafana prometheus kubernetes opentelemetry
JobCopilot logo
Meet JobCopilot: Your Personal Al Job Hunter
Automatically Apply to Your Dream Jobs While You Sleep
Try it now โ†’

๐Ÿ“‹ Description

  • Monitor, operate, and support production AI infrastructure platforms.
  • Investigate and resolve infrastructure, networking, hardware, and platform incidents.
  • Support NVIDIA GPU infrastructure and related platform services.
  • Monitor and troubleshoot Kubernetes-based environments.
  • Collaborate with engineering teams, vendors, datacenters, and service delivery teams to resolve technical issues.
  • Participate in incident response and root cause analysis.

๐ŸŽฏ Requirements

  • 3+ years in infrastructure, platform, network ops, SRE, cloud, or datacenter roles.
  • Strong Linux administration and troubleshooting skills.
  • Good understanding of networking concepts and diagnosing infra issues.
  • Working knowledge of Kubernetes in production environments.
  • Experience supporting production infrastructure and services.
  • Ability to work in a shift-based operational environment.

๐ŸŽ Benefits

  • Work with advanced AI infrastructure in production today.
  • Exposure to NVIDIA GPU tech, Kubernetes, and high-performance networking.
  • Define how next-generation AI infrastructure is operated and supported.
  • Be part of a team shaping AI-powered operations through k0rdent AI.
  • Join a growing organization investing in AI infrastructure and platform services.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest โ€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs โ†’