Senior Site Reliability Engineer - SDN

Added
25 days ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

terraform python kubernetes go sdn

๐Ÿ“‹ Description

  • Operate and scale Lambda's multi-tenant cloud networking and SDN infra
  • Operate and improve Kubernetes-based control plane and SmartNIC dataplane
  • Develop tooling to reduce toil and improve reliability
  • Collaborate with software, platform, and networking teams to improve reliability
  • Deploy and maintain network monitoring, observability, and management tools
  • Hybrid work: presence in SF/SJ/Bellevue 4 days/week; WFH Tue

๐ŸŽฏ Requirements

  • 5+ years in Site Reliability Engineering, Production Engineering, or a similar role
  • Experience operating large-scale distributed systems in production
  • Kubernetes lifecycle management, upgrades, troubleshooting
  • On-call rotations and incident response experience
  • Strong troubleshooting across Linux, Kubernetes, distributed systems, and networking
  • Experience with observability platforms, monitoring, alerting, and metrics

๐ŸŽ Benefits

  • Generous cash and equity compensation
  • Health, dental, and vision coverage
  • Wellness and commuter stipends
  • 401k plan with 2% company match
  • Flexible paid time off
  • Equal opportunity employer
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest โ€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs โ†’