Senior DevOps Engineer

Telesat.com β€” Canada Β· Posted ~21 hours ago

Senior

Skills

DevOps Azure Cloud infrastructure CI/CD Kubernetes Infrastructure automation Networking Identity management AKS Java Python TypeScript MATLAB

πŸ”“ Log in to save this job, tailor your resume & track your apply process β€” 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Take ownership of cloud infrastructure and CI/CD for a large-scale, multi-team engineering platform. You will design and manage Azure resources, including Kubernetes, networking, storage, and identity, while improving reliability, scalability, and automation. The environment spans multiple programming languages and relies on a shared monorepo, making strong platform engineering and DevOps practices essential.

Highlights

Own and evolve cloud infrastructure and CI/CD for a sophisticated large-scale platform. The role provides significant responsibility for reliability, scalability, automation, Azure architecture, Kubernetes, networking, storage, and identity across a multi-team engineering environment.

Description

About Telesat Telesat is a leading global satellite operator, providing reliable and secure satellite-delivered communications solutions worldwide to broadcast, telecommunications, corporate and government customers. Backed by a legacy of engineering excellence, reliability and industry-leading customer service, Telesat has grown to be one of the largest and most successful global satellite operators. About The Role We are seeking a Senior DevOps Engineer to own and evolve the cloud infrastructure and CI/CD platform for Telesat's Lightspeed satellite constellation. You will be responsible for the reliability, scalability, and automation of our Azure-based deployment platform, supporting a multi-technology monorepo (Java, Python, TypeScript, MATLAB) serving multiple development teams. Responsibilities Cloud Infrastructure (Azure): Design, provision, and manage Azure cloud resources β€” primarily Azure Kubernetes Service (AKS), networking, storage, and identity β€” ensuring high availability and cost efficiency for satellite mission-critical workloads. Infrastructure as Code: Author and maintain Terraform modules to provision and manage all cloud environments (dev, staging, production), following DRY principles and modular design. CI/CD Pipeline Engineering: Own and extend GitLab CI/CD pipelines (merge request pipelines, merge trains, nightly builds, deployment workflows) across a polyglot monorepo. Ensure fast, reliable feedback loops for Java (Gradle), Python (Poetry/pytest), Matlab, and TypeScript (npm workspaces/Vite) build targets. Containerization & Orchestration: Maintain Docker images, Helm charts, and Kubernetes manifests. Manage AKS cluster lifecycle including upgrades, scaling policies, and namespace isolation. Automation & Scripting: Write and maintain Bash/Python scripts for deployment automation, environment provisioning, secret rotation, and developer tooling (pre-commit hooks, quality gates). Observability: Operate and improve the monitoring stack (Grafana, Prometheus, Loki) to provide actionable dashboards, alerting, and log aggregation across all services. Security: Implement and enforce cloud security best practices β€” network policies, RBAC, secret management, vulnerability scanning, and compliance controls for satellite communication systems. Collaboration: Work closely with backend, frontend, QA, and MATLAB engineering teams to unblock deployments, improve developer experience, and maintain platform reliability. Documentation: Maintain runbooks, architecture decision records, and onboarding guides for infrastructure and deployment processes. Required Qualifications 5+ years of professional DevOps / Cloud / Platform Engineering experience Azure (AKS, Virtual Networks, Key Vault, Storage, Entra ID) β€” hands-on production experience Terraform β€” writing, reviewing, and managing modular IaC at scale GitLab CI/CD β€” advanced pipeline design (includes, templates, merge trains, multi-project pipelines) Kubernetes β€” cluster administration, Helm, troubleshooting pod/network issues Docker β€” building optimized images, multi-stage builds, registry management Linux β€” strong command-line proficiency, shell scripting Git β€” branching strategies, merge conflict resolution, monorepo workflows Nice to Have Observability tooling: Grafana, Prometheus, Loki (dashboards, alerting rules, log queries) Configuration management: Ansible Artifact management: JFrog Artifactory (Docker, Maven, PyPI registries) Experience with pre-commit frameworks and developer experience tooling Familiarity with satellite/aerospace or other regulated/mission-critical domains Experience supporting polyglot monorepos (Java + Python + TypeScript in one repository) Performance Expectations Ability to work autonomously and able to hit the ground running within the first month Must be able to work with others and be an effective communicator Work Environment Location: 160 Elgin St Suite 2100, Ottawa, Ontario, Canada, K2P 2P7 On site (downtown Ottawa) 4 days will be in office (Wednesday will be a work from home option) We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.