Principal Observability Engineer

Isgsearch — Canada · Posted ~3 hours ago

Lead Contract

Skills

Observability Splunk OpenTelemetry Grafana Prometheus AWS CloudWatch Kubernetes Infrastructure as Code Scripting

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A principal engineering role focused on designing monitoring platforms, improving system reliability, automating infrastructure, and mentoring technical teams in large-scale environments.

Highlights

Principal-level role leading enterprise observability architecture, reliability improvements, automation, and engineering best practices.

Description

Job title: Principal Observability Engineer Duration: 12-month contract (extension possible) Role status: Current opening Principal tasks and responsibilities include: Lead the architecture, design, and implementation of enterprise observability solutions and monitoring frameworks.Develop and enhance dashboards, alerting, and monitoring capabilities using Splunk, Splunk Observability, OpenTelemetry, Grafana, Prometheus, and AWS CloudWatch.Design and support observability across cloud and Kubernetes-based environments, driving platform performance and reliability.Develop infrastructure automation using Infrastructure as Code (IaC), scripting, and configuration management tools.Provide technical leadership, establish observability best practices, and mentor engineering teams. Our client: A large global enterprise organization Qualifications and pre-requisites: 10+ years of experience in Platform Engineering, Site Reliability Engineering (SRE), Systems Engineering, or a related infrastructure discipline.Strong hands-on experience with Splunk and modern observability platforms, including Splunk Observability, OpenTelemetry, Grafana, Prometheus, and AWS CloudWatch.Deep expertise with AWS, Kubernetes/OpenShift, and infrastructure automation tools such as Terraform, Ansible, Python, or Bash.Proven experience designing scalable monitoring and observability solutions within large enterprise environments.Strong communication and leadership skills with the ability to collaborate across technical and business teams. Additional information or perks: Experience with additional cloud platforms (Azure or GCP), Red Hat technologies, virtualization, storage, backup, or identity management is considered an asset.Relevant certifications such as Splunk, AWS, or Kubernetes are highly desirable.isgSearch does not use artificial intelligence throughout the hiring process.