Senior Site Reliability Engineer

Systemsofttech β€” Canada Β· Posted ~7 hours ago

Senior Full-time Remote

Skills

Kubernetes Cloud-native platforms CI/CD SRE Infrastructure as Code Observability Incident response CNCF Backstage Prometheus Grafana

πŸ”“ Log in to save this job, tailor your resume & track your apply process β€” 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A senior SRE position responsible for building and operating cloud-native developer platforms, improving reliability, and automating infrastructure workflows using modern technologies.

Highlights

Fully remote opportunity focused on cloud-native platforms, automation, reliability engineering, and modern AI-enabled operations.

Description

πŸš€ Senior Platform Engineer (SRE) - AI Solutions πŸ“ Montreal, Quebec, Canada (Remote) | Quebec, Canada or Slovakia Are you passionate about cloud-native platforms, Kubernetes, AI, and building the foundation that enables development teams to move fast and operate with confidence? Join us, a global leader in managed cloud and enterprise application services, as we transform engineering teams through AI-powered development and operations. What You'll Do βœ… Build and operate a cloud-native developer platform leveraging Kubernetes and the CNCF ecosystem βœ… Design self-service golden-path blueprints for provisioning, CI/CD, deployment, and operations βœ… Own and enhance the internal developer portal (Backstage) βœ… Drive observability with Prometheus and Grafana βœ… Lead SRE initiatives including SLOs, incident response, reliability, performance, and automation βœ… Use AI-powered and agentic tools to produce production-quality Infrastructure as Code and advance AI-driven operations What We're Looking For βœ” Deep expertise with Kubernetes and cloud-native technologies βœ” Hands-on experience with AWS, Azure, or GCP βœ” Strong SRE background including reliability engineering and incident response βœ” Experience with Terraform, Helm, GitOps, and developer enablement βœ” Backstage, Prometheus, and Grafana experience βœ” Passion for AI-assisted development and operations workflows Nice to Have ⭐ Multi-cloud expertise (AWS, Azure, GCP) ⭐ Open-source or CNCF contributions ⭐ AI-assisted operations, self-healing platforms, policy-as-code, or supply chain security