GCP Infrastructure and Cloud Site Reliability Engineer

Jmgroupinc — Canada · Posted ~3 hours ago

Senior

Skills

Google Cloud Platform Google Kubernetes Engine Kubernetes Terraform Infrastructure as Code GitHub Actions Jenkins Argo CD Site reliability engineering Observability Monitoring Incident management Linux Python Shell scripting Cloud networking Identity and access management CI/CD GitOps Cloud security Automation ArgoCD GKE Shell IAM

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A technology organization is seeking a cloud infrastructure and site reliability engineer to build and manage enterprise-scale platforms. You will implement Kubernetes-based infrastructure, automate provisioning with infrastructure-as-code tools, maintain CI/CD and GitOps workflows, and establish monitoring and reliability standards. Responsibilities also include incident response, cloud security, networking, disaster recovery, performance optimization, and cost management.

Highlights

Build and operate enterprise-scale cloud platforms with modern infrastructure automation and Kubernetes technologies. The role offers broad ownership across reliability, security, observability, deployment automation, disaster recovery, and cloud cost optimization while working with multidisciplinary engineering teams.

Description

Build and manage enterprise-scale cloud platforms on Google Cloud Platform (GCP). Design and implement GKE/Kubernetes-based container platforms. Develop Infrastructure as Code (IaC) solutions using Terraform. Build and maintain CI/CD pipelines and GitOps deployment frameworks. Define and manage SLI/SLOs, monitoring, observability, and reliability standards. Automate platform provisioning, operations, and incident response. Implement cloud security, IAM, networking, governance, and compliance controls. Drive high availability, disaster recovery, performance optimization, and cost management. Partner with development, security, and infrastructure teams to enable self-service platform capabilities. Required Skills Strong hands-on experience with GCP Services, GKE, Compute Engine, Networking, IAM. Expertise in Kubernetes, Terraform, GitHub Actions/Jenkins, ArgoCD. Experience with SRE practices, observability, incident management, and automation. Proficiency in Linux, Python, Shell Scripting. Experience building cloud landing zones, developer platforms, and reusable infrastructure services. Knowledge of FinOps, cloud security, and enterprise operations. Preferred GCP Professional Cloud Architect/DevOps Engineer certification. Banking/Financial Services domain experience.