Intelligent Infrastructure Engineer

Added
1 day ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

python kubernetes go gpu ray

πŸ“‹ Description

  • Design and operate infra platforms for large-scale AI training and inference.
  • Manage GPU clusters, distributed training, and scheduling for ML workloads.
  • Improve reliability, performance, and scalability of AI infrastructure.
  • Develop automation tools using Python and Go or C++.
  • Configure Kubernetes, Slurm, Ray for ML workloads.
  • Implement strong software practices: testing, CI/CD, code reviews.

🎯 Requirements

  • Bachelor's or Master's in CS, Engineering, or related field.
  • 6+ years in infrastructure, platform engineering, or HPC.
  • Hands-on with GPU clusters or large-scale ML infra.
  • Python plus Go or C++ systems programming.
  • Kubernetes, Slurm, Ray, or similar schedulers.
  • Linux internals, networking, and high-performance storage.

🎁 Benefits

  • Competitive salary: $100k–$150k per year
  • Fully remote within the United States
  • Full-time direct employment
  • Work on advanced AI infrastructure and ML systems
  • Career growth opportunities
  • Exposure to cloud, AI, and distributed computing
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’