Sr. Staff Site Reliability Engineer

Added
24 minutes ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

ansible terraform aws python gcp

πŸ“‹ Description

  • Develop production-grade automation in Ansible, Python and/or Go for bare-metal provisioning
  • Design and implement fault-tolerant provisioning systems with self-healing and auto-remediation
  • Develop telemetry pipelines for hardware and system monitoring β€” integrating metrics, logs, and
  • Collaborating with cross-functional teams to deliver integrated solutions and contribute to
  • Partner directly with SWE teams to define and enforce production-readiness criteria β€” including

🎯 Requirements

  • Foundational understanding of AI/ML technologies and experience leveraging, securing, or
  • 7+ years of Production Engineering, Platform Engineering, or Infrastructure Engineering experience
  • Deep Linux/Unix systems mastery β€” OS internals, kernel networking, boot pipelines and low-level
  • Hands-on operational experience across GCP and/or AWS alongside on-prem datacenter operations, with
  • Expertise building durable, idempotent, and repeatable automation using Python and/or Go

🎁 Benefits

  • Various health plans
  • Time off plans for vacation and sick time
  • Parental leave options
  • Retirement options
  • Education reimbursement
  • In-office perks, and more!

🚚 Relocation support

πŸ›ƒ Visa sponsorship

Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’