Sr. Production Engineer

Added
3 days ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

azure aws prometheus python kubernetes

πŸ“‹ Description

  • Design and implement highly available infrastructure across AWS, Azure, GCP
  • Drive automation to eliminate toil and build self-healing systems
  • Observe and improve observability, define SLIs/SLOs
  • Lead incident response and post-incident analyses
  • Partner with teams to improve operability reviews

🎯 Requirements

  • Foundational understanding of AI/ML and AI-driven solutions
  • 4+ years in reliability and scalable production services
  • Strong programming in Python/Go/C/C++
  • Networking, Linux/FreeBSD, distributed systems
  • Experience with 24/7 on-call and incident management
  • ITIL-based operability and problem management

🎁 Benefits

  • Hybrid work model (3 days/wk in San Jose or Bellevue)
  • Salary range: $103,600 β€” $148,000 USD
  • Comprehensive health, time off, parental leave
  • Retirement options and education reimbursement
  • In-office perks and resources
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’