Summary
A technology services organization is seeking a Platform and SRE Engineer to design and operate scalable containerized applications. You will manage Kubernetes platforms, build CI/CD pipelines, improve observability, support authentication and authorization, troubleshoot production issues, and automate operational tasks.
Highlights
Own and improve scalable cloud-native platforms with strong emphasis on reliability, observability, automation, and production performance. The role provides broad exposure to Kubernetes, security, CI/CD, infrastructure automation, and system design.
Description
Job Title: Platform & SRE Engineer
Location: Montreal, QC
What You’ll Do
• Design, deploy, and manage scalable, containerized applications on Kubernetes clusters.
• Build and maintain CI/CD pipelines to enable reliable and repeatable deployments.
• Own and improve platform observability using tools like Prometheus, Grafana, and alerts.
• Work closely with application teams to onboard services with proper authentication and authorization (OAuth2, SSO, JWT, etc.).
• Troubleshoot production issues, improve availability, and drive root cause analysis.
• Standardize and automate routine tasks using Python scripting or Ansible.
• Support web-based deployments including APIs, UI apps, and their configurations.
• Collaborate on system design, capacity planning, and performance tuning.
Must-Have Skills
• Solid hands-on experience with Kubernetes and Docker.
• Strong understanding of containerization, deployment strategies, and orchestration.
• Deep familiarity with observability stacks: Prometheus, Grafana, alerts, and metrics.
• Working knowledge of web deployment and web component architecture.
• Experience in setting up or working with authN/authZ mechanisms (OAuth, SSO)
• Good understanding of CI/CD pipelines (Jenkins, GitLab CI, etc.).
• Comfort with Linux and basic database operations (SQL)
Nice to Have
• Proficiency in Python scripting for automation.
• Experience with Ansible or other config management tools.
• Exposure to infrastructure as code (Terraform, Helm).
• Familiarity with cloud-native concepts and services.
Who You Are
• A systems thinker who understands both the development and operations side.
• Someone who thrives in a fast-paced, cross-functional team.
• Passionate about automation, stability, and continuous improvement