Site Reliability Engineer - Datacenter

Added
5 days ago
Type
Full time
Salary
Salary not provided

Related skills

data analysis bash python monitoring c/c++

πŸ“‹ Description

  • Role focuses on hardware, firmware, vendor relations, and failure analysis in the datacenter.
  • Handle RMA processes and forward-looking hardware evaluation.
  • Collaborate with Datacenter Operations Technicians to troubleshoot in real time.
  • Develop monitoring tools and scripts to detect hardware anomalies.
  • Participate in on-call rotations for hardware-related incidents in Memphis area.
  • Wear multiple hats in a flat org; contribute directly to mission with strong communication.

🎯 Requirements

  • Bachelor's in Systems/EE/CS or related field, or equivalent experience.
  • 2+ years in hardware reliability engineering in HPC/datacenter.
  • Firmware analysis, hardware specs review, release validation.
  • Experience with RMA processes and vendor negotiations.
  • Diagnose complex hardware failures with data-driven approach.
  • Familiarity with datacenter hardware (servers, GPUs, networking).

🎁 Benefits

  • Equal opportunity employer.
  • Recruitment Privacy Notice link provided in posting.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’