Cloud Site Reliability Engineer Specialist

Scotiabank — Colombia · Posted ~22 hours ago

Senior Full-time

Skills

Google Cloud Platform Site reliability engineering Reliability engineering Cloud infrastructure Software engineering Systems engineering Distributed systems Monitoring

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Join a high-performing technology organization as a Cloud Site Reliability Engineer focused on the availability, performance, and resilience of critical applications. You will combine software and systems engineering to build and operate large-scale, fault-tolerant cloud infrastructure.

Highlights

Work on critical cloud infrastructure and large-scale distributed systems, with a strong focus on reliability, resilience, availability, performance, and proactive monitoring.

Description

Requisition ID: 271054 Thanks for your interest in ScotiaTech, Scotiabank's new and innovative Technology hub in Bogota. Join a purpose driven winning team that promotes creativity and innovation in a fast-paced environment, where we’re always committed to results, in an inclusive, diverse, and high-performing culture. Purpose The System Reliability Engineering team focuses on System Reliability, Reliability Engineering, and Resilience across Global Corporate Functions, with a specific emphasis on Google Cloud Platform engineering and onboarding to Cloud infrastructure. The Cloud Site Reliability Engineer Specialist is responsible for the availability, performance, and reliability of critical corporate function applications hosted on Google Cloud Platform. This role combines software and systems engineering to build and run large-scale, distributed, fault-tolerant systems. The primary objective is to enhance system resilience and reduce operational work through proactive monitoring, automation, and continuous improvement. The position's mandate is to drive the adoption of SRE principles and practices, collaborating with development teams to engineer scalable and reliable solutions from inception through to production. Accountabilities Design, build, and maintain scalable and resilient infrastructure on Google Cloud Platform using Terraform. Develop and manage automation solutions with Python to eliminate operational toil and improve system efficiency. Deploy, manage, and scale containerized applications using Kubernetes, ensuring optimal performance and availability. Implement and administer secure cloud networking architectures, including virtual private clouds, firewall rules, and load balancing. Establish and enforce robust security protocols and secret management practices within the cloud environment. Drive the adoption of GitOps methodologies using tools like Argo CD or Flux for declarative infrastructure and application management. Define, track, and report on Service Level Indicators and Service Level Objectives to maintain reliability targets. Construct and optimize CI/CD pipelines to enable rapid, reliable, and automated software delivery. Lead incident response efforts for Google Cloud Platform, facilitate blameless post-mortems, and implement corrective actions to prevent future occurrences. Collaborate with application development teams to integrate reliability best practices into the software development lifecycle. Education / Experience / Other Information Completion of a post-secondary degree in Computer Science, Engineering, or a related technical field. Experience working within the financial services or a similarly regulated industry. Expert-level knowledge of Google Cloud Platform services, cloud networking, and security principles. Familiarity with regulatory and compliance standards applicable to the banking sector. Extensive experience with Infrastructure as Code, specifically using Terraform. Advanced proficiency in scripting and automation using Python. Demonstrated expertise in container orchestration with Kubernetes. Strong understanding and practical application of GitOps principles and tools such as Argo CD or Flux. Proven experience designing, building, and maintaining CI/CD pipelines for automated deployments. In-depth knowledge of SRE principles, including Service Level Objectives, error budgets, and blameless post-mortems. Working Conditions Work in a standard office-based environment; non-standard hours are a common occurrence Location(s): Colombia : Bogota : Bogota ScotiaTech is a business unit within ScotiaGBS, a Scotiabank Group company located in Bogota, Colombia. The ScotiaTech hub was created to support different technology systems and processes of the Bank. We offer an inclusive, positive work environment, and competitive benefits. At ScotiaTech, we value the unique skills and experiences each individual brings and are committed to creating and maintaining an inclusive and accessible environment for everyone. Candidates must apply directly online to be considered for this role. We thank all applicants for their interest in a career at ScotiaTech; however, only those candidates who are selected for an interview will be contacted. Note: All postings in me@Scotiabank will remain live for a minimum of 5 days.