Systems Reliability Engineer

Added
13 minutes ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

java linux grafana prometheus python
JobCopilot logo
Meet JobCopilot: Your Personal Al Job Hunter
Automatically Apply to Your Dream Jobs While You Sleep
Try it now β†’

πŸ“‹ Description

  • Design, build, and maintain reliable distributed systems for higher availability.
  • Develop automation tools using Python, Go, or Java.
  • Operate and troubleshoot Linux systems at scale (networking, perf).
  • Manage Kubernetes and containerized workloads in production.
  • Build and improve CI/CD pipelines for apps and infrastructure.
  • Implement observability with Prometheus, Grafana, OpenTelemetry, ELK/EFK.

🎯 Requirements

  • Bachelor's degree in Computer Science, Engineering, or related technical discipline.
  • 5+ years in SRE, DevOps, or production engineering with large-scale distributed systems.
  • Strong programming in Python, Go, or Java.
  • Linux system management with networking, troubleshooting, and performance tuning.
  • Production experience with Kubernetes and container-based architectures.
  • Observability/monitoring with Prometheus, Grafana, OpenTelemetry, ELK/EFK.

🎁 Benefits

  • Competitive annual salary of $100,000–$150,000.
  • Fully remote work environment within the United States.
  • Full-time employment with career growth potential.
  • Work on large-scale distributed systems and advanced tech projects.
  • Collaborative environment focused on engineering excellence.
  • Exposure to modern cloud, automation, observability, and reliability practices.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to DevOps Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related DevOps Jobs

See more DevOps jobs β†’