Added
12 minutes ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

datadog sre aws kubernetes observability
JobCopilot logo
Meet JobCopilot: Your Personal Al Job Hunter
Automatically Apply to Your Dream Jobs While You Sleep
Try it now β†’

πŸ“‹ Description

  • Lead observability, reliability, and operational excellence across cloud-native environments.
  • Define SRE standards, governance, and monitoring strategies for distributed systems.
  • Partner with product and engineering to improve system reliability and performance.
  • Design and implement monitoring, alerting, dashboards, and reporting solutions.
  • Collaborate across distributed teams in a multilingual and global context.
  • Drive technical initiatives, prioritization, and timely delivery within budgets.

🎯 Requirements

  • Significant experience in Site Reliability Engineering (SRE) in large-scale production.
  • Strong expertise in observability platforms, SLIs/SLOs, and alerting governance.
  • Hands-on experience with Dynatrace, OpenTelemetry, and distributed tracing.
  • Proficient with Dynatrace, Datadog, New Relic, AppDynamics, and related APM/tools.
  • Experience with cloud and container tech (AWS, Kubernetes), and monitoring tooling (RUM, APM).
  • Strong collaboration, stakeholder management, and communication in English; French/English

🎁 Benefits

  • Competitive salary: CAD 100,000 - 150,000 CAD annually.
  • Comprehensive benefits after three months: health insurance, disability coverage, EAP, and mental
  • Personal tech reimbursement, DPSP, flexible vacation, and retirement matching.
  • Distributed team with remote-friendly setup and a borderless global framework.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Engineering Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Engineering Jobs

See more Engineering jobs β†’