Senior Site Reliability Engineer

Compunnel Software Group — Canada · Posted ~22 hours ago

Senior Hybrid

Skills

Kubernetes Linux Command line Production systems operations Debugging Troubleshooting Public cloud Azure AWS Grafana Prometheus Loki Tempo Python Java CI/CD Helm Terraform

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A senior Site Reliability Engineer role responsible for operating production Kubernetes environments and troubleshooting complex systems from the application layer through underlying infrastructure. Strong Linux fundamentals and public cloud experience are essential, while observability, scripting, CI/CD, and Infrastructure as Code experience are valuable additions.

Highlights

Senior-level reliability engineering role with hands-on production Kubernetes, complex systems troubleshooting, and public cloud exposure. The role also offers opportunities to work with modern observability, automation, and cloud-native tooling.

Description

Job Title: Site Reliability Engineer Experience Level: Level 3 (senior): 5-7 years Location: Montreal (Day 1 onboarding onsite/in office presence 3x/week) Required Skills: Hands on experience operating Kubernetes in production, ideally with Service Mesh. Strong Linux and command line fundamentals. Confident debugging & troubleshooting complex systems, from the application layer through to lower-level infrastructure. Experience working with a public cloud provider, preferable Azure or AWS. Working knowledge of Grafana, Prometheus, Loki and Tempo is a plus. Scripting or coding in Python or Java is a strong plus. CI/CD, infrastructure as code such as Helm or Terraform is a plus A financial services background is not required.