Site Reliability Engineer

Sohosquaresolutions — Canada · Posted ~2 hours ago

Mid Onsite

Skills

Kubernetes Linux Service Mesh Debugging Troubleshooting Cloud Provider (Azure/AWS) Azure AWS Grafana Prometheus Python Java Helm Terraform

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

We are seeking a Site Reliability Engineer for an in-person role in Montreal. The ideal candidate has hands-on experience operating Kubernetes in production, ideally with Service Mesh. Strong Linux and command line fundamentals are essential. You should be confident in debugging and troubleshooting complex systems, from the application layer through to lower-level infrastructure. Experience working with a public cloud provider, preferably Azure or AWS, is required. Knowledge of monitoring tools and scripting skills are a plus.

Highlights

Local opportunity in Montreal, hands-on role with production Kubernetes, complex systems debugging

Description

ONLY Locals to Montreal In Person Interview Must Required Skills: Hands on experience operating Kubernetes in production, ideally with Service Mesh.Strong Linux and command line fundamentals.Confident debugging & troubleshooting complex systems, from the application layer through to lower-level infrastructure.Experience working with a public cloud provider, preferable Azure or AWS.Working knowledge of Grafana, Prometheus, Loki and Tempo is a plus.Scripting or coding in Python or Java is a strong plus.CI/CD, infrastructure as code such as Helm or Terraform is a plusA financial services background is not required.