Customer Success Engineer (CSE), GPU Cluster

Added
30 minutes ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

bash grafana prometheus python ethernet

πŸ“‹ Description

  • Named technical owner for a strategic customer across compute, networking, storage, and facilities.
  • Drive engagement with regular cadences: status reports, steering meetings, QBRs, and EBRs.
  • Translate customer feedback into input for Engineering, Product, and Infrastructure roadmaps.
  • Lead issue lifecycle management, escalation, and RCA across all infrastructure domains.
  • Own end-to-end RMA coordination and hardware lifecycle for large GPU deployments.
  • Maintain expertise across GPU compute, fabric, and large-scale storage; advise on config and incident resolution.

🎯 Requirements

  • 5+ years in a customer-facing technical role, with 2+ years in dedicated technical account management or solutions architecture for large-scale AI or HPC infrastructure
  • Deep expertise in GPU infrastructure β€” GPU health diagnostics, RMA workflows, and hardware acceptance testing
  • Hands-on experience with large-scale Ethernet and InfiniBand fabric architecture
  • Working knowledge of enterprise storage systems, including high-density NVMe, parallel file systems, and metadata infrastructure
  • Experience with DC operations, facilities coordination, and hosting provider SLA management
  • Proficiency in Python, Bash, or infrastructure automation tools preferred

🎁 Benefits

  • Startup equity
  • Health insurance
  • Other benefits
  • Remote work flexibility
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Customer Support Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Customer Support Jobs

See more Customer Support jobs β†’