Observability Engineer

Zodiac Solutions Inc — Canada · Posted ~6 hours ago

Skills

Prometheus Grafana Thanos Loki GitOps Infrastructure as Code cloud object storage distributed systems

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Design and operate enterprise-scale observability infrastructure using modern monitoring, logging, and metrics technologies. You will build distributed architectures, manage deployments through GitOps and infrastructure-as-code practices, and support reliable platforms across multiple environments and cloud providers.

Highlights

Work on enterprise-scale observability infrastructure spanning development, QA, UAT, production, disaster recovery, multiple datacenters, and cloud providers. The role includes modern GitOps and infrastructure-as-code practices.

Description

• Design, deploy, and maintain enterprise-scale observability infrastructure including Prometheus, Grafana, Thanos, Loki, and modern collection agents • Manage observability deployments using GitOps principles and infrastructures code • Implement long-term metrics storage solutions with cloud object storage • Maintain and upgrade observability components across development, QA, UAT, production, and DR environments • Configure distributed observability architecture spanning multiple datacenters and cloud providers