Site Reliability Engineer

Infotek Consulting Services Inc — Canada · Posted ~4 hours ago

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Description

Site Reliability Specialist Position Overview We are seeking a Senior Cloud Platform Engineer / Site Reliability Specialist to design, build, automate, secure, and support enterprise Azure and Databricks platforms. This is a hands-on engineering role focused on cloud infrastructure, DevOps, automation, security, reliability, and troubleshooting. Key Responsibilities Design, build, deploy, and support enterprise Azure infrastructure across Windows and Linux environments.Build and maintain Azure environments across development, testing, production, and disaster recovery.Deploy, configure, secure, and support Azure Databricks platforms.Manage Azure VMs, App Services, Storage, Networking, Monitoring, Backup, and DR solutions.Design highly available, scalable, secure, and resilient cloud architectures.Troubleshoot complex infrastructure, networking, security, and performance issues.Develop Infrastructure-as-Code using Terraform, Ansible, Bicep, or similar tools.Build and maintain CI/CD pipelines and automated deployment processes.Develop PowerShell and Python automation to improve efficiency and reliability.Support cloud migration, modernization, security, governance, and operational initiatives.Maintain technical documentation, runbooks, and infrastructure code using Git.Required Qualifications 7+ years of experience in cloud, infrastructure, platform, or SRE engineering.Strong hands-on Azure experience, including building enterprise environments from the ground up.Strong Azure Databricks experience.Experience with Infrastructure-as-Code, automation, and CI/CD.Experience supporting both Windows and Linux environments.Strong troubleshooting and production support skills.Ability to design and implement secure, highly available, and resilient cloud platforms.Nice to Have Azure certifications.Experience with Terraform, Ansible, Bicep, PowerShell, or Python.Cloud migration/modernization experience.Disaster recovery and cloud governance experience.Contract Details 12-month contract with potential extension and conversionHybrid – multiple days onsite per weekMonday–Friday, core business hoursOccasional overtime may be requiredNo travel required