Site Reliability Engineer in Network Infrastructure

Added
21 days ago
Type
Full time
Salary
Salary not provided

Related skills

linux networking python kubernetes go

📋 Description

  • Improve reliability, performance, and operating maturity of large-scale network systems.
  • Define and manage SLI/SLOs, availability targets, and error budgets.
  • Drive reliability improvements across network infrastructure, including site readiness and
  • Own incident response, lead investigations, and implement long-term solutions.
  • Build observability systems with metrics, logs, traces, and alerts.
  • Design safer infrastructure change processes with automation and CI/CD.

🎯 Requirements

  • Strong experience in site reliability engineering, infrastructure operations, and network systems.
  • Production Linux environments and structured troubleshooting.
  • Networking fundamentals across control/data planes, latency, and failure domains.
  • Experience operating high-availability systems and improving reliability over time.
  • Automation and infrastructure software development; Go preferred, Python welcomed.
  • Experience with Infrastructure as Code, CI/CD, and container platforms.

🎁 Benefits

  • Competitive compensation package.
  • Career growth and continuous learning opportunities.
  • Flexible working environment with ownership and autonomy.
  • Opportunity to work on impactful cloud and AI infrastructure projects.
  • Collaborative culture with international engineering teams.
  • Exposure to cutting-edge networking and cloud technologies.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest — finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs →