Sr. Staff Production Engineer

Added
6 days ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

azure aws grafana prometheus python
JobCopilot logo
Meet JobCopilot: Your Personal Al Job Hunter
Automatically Apply to Your Dream Jobs While You Sleep
Try it now β†’

πŸ“‹ Description

  • Hybrid role: 3 days/wk in San Jose, CA or Bellevue, WA
  • Own and drive reliability for a global cloud infra platform
  • Lead automation-first mindset and self-healing systems
  • Improve observability, SLOs, MTTR/MTTM, and incident response
  • Collaborate with SRE, cloud infra, and partner teams
  • Contribute to scalable multi-cloud infrastructure across environments

🎯 Requirements

  • 8+ years in reliability/scalability for large-scale services
  • Strong programming: Python/Go or C/C++
  • Experience with AWS, Azure, GCP; IaC (Terraform/Ansible)
  • Expertise in Prometheus, Grafana, OpenTelemetry
  • Knowledge of incident management and on-call practices
  • Networking basics: BGP, GRE/IPSec, L7 proxies

🎁 Benefits

  • Comprehensive health plans
  • Time off and parental leave options
  • Retirement options and education reimbursement
  • In-office perks and hybrid work support
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’