Senior Technical Operations & Deployment Engineer (GPU Cloud Infrastructure)

Added
1 day ago
Type
Full time
Salary
Salary not provided

Related skills

terraform linux kubernetes driver cuda

๐Ÿ“‹ Description

  • Deploy and operate GPU cloud infra in datacenters
  • Coordinate hardware, networking, storage, and platform software
  • Validate BIOS, firmware, NICs, DPUs, GPUs, NVMe, and PCIe topology
  • Troubleshoot across hardware and software layers
  • Establish deployment standards and runbooks
  • Collaborate with engineering, network, and service teams

๐ŸŽฏ Requirements

  • Datacenter infra experience in GPU/AI cloud envs
  • Bare-metal bring-up and hardware validation
  • GPU servers, drivers, NVLink/NVSwitch familiarity
  • Linux troubleshooting and kernel/driver experience
  • Networking: VLANs, BGP, EVPN, OVS/OVN
  • Observability tools: Prometheus, Grafana, Zabbix

๐ŸŽ Benefits

  • Competitive compensation
  • Full-time or contract depending on arrangement
  • Europe-based remote work environment
  • Hands-on GPU platform exposure and career growth
  • International, diverse team
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest โ€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs โ†’