Senior Observability Engineer

Ci Financial — Canada · Posted ~2 hours ago

Mid Full-time

Skills

Dynatrace AWS CloudWatch monitoring observability dashboard development alert tuning instrumentation cloud environments AWS

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Help build and enhance enterprise observability across cloud and hybrid environments. Work hands-on with monitoring platforms, dashboards, alert tuning, instrumentation, and operational reporting while partnering with engineering and application teams to improve service visibility and reliability.

Highlights

Hands-on observability role improving visibility and monitoring across cloud and hybrid environments. Provides collaboration with Cloud Engineering, DevOps, SRE, and application teams while working with enterprise-scale monitoring technologies.

Description

Chez CI, nous considérons qu’un excellent lieu de travail est un endroit sécuritaire où chacun peut s’exprimer, où les employés ont la possibilité de s’investir dans un travail valorisant, où ils ont l’occasion de se dépasser pour progresser, où ils peuvent travailler sur des produits et des projets innovants, et où ils sont soutenus et encouragés dans leurs efforts. We are seeking a Mid-Level Observability Engineer to help build, maintain, and enhance our enterprise monitoring and observability capabilities across cloud and hybrid environments. This role is hands-on and execution-focused, supporting Dynatrace and AWS CloudWatch implementations, dashboard development, alert tuning, instrumentation, and operational reporting for critical platforms and applications. The ideal candidate will partner with Cloud Engineering, DevOps, SRE, and application teams to improve service visibility, strengthen monitoring coverage, and embed observability practices into ongoing operational and delivery workflows. Observability Platform Ownership & Architecture Design, deploy, and optimize enterprise-grade observability solutions using Dynatrace SaaS or Managed, including OneAgent, ActiveGate, full-stack monitoring, RUM, synthetic monitoring, Davis AI, distributed tracing, dashboards, and log monitoring on Grail.Define platform standards for tagging, management zones, network segmentation, alerting profiles, access control, dashboards, and telemetry governance across hybrid environments.Architect observability coverage across AWS and on-prem platforms, including containerized and serverless workloads such as EKS, ECS, Lambda, EC2, RDS, and API Gateway.Lead migration from legacy monitoring tools into Dynatrace and drive closure of enterprise monitoring gaps through structured onboarding and platform modernization. 2. Application Performance Management & Incident TriageConfigure and optimize APM instrumentation for distributed applications, APIs, microservices, databases, and business transactions.Serve as the escalation point for complex incidents, using Smartscape, Distributed Traces, Davis AI, Live Debugger, and method-level diagnostics to accelerate root cause identification and reduce MTTR.Define and maintain SLIs, SLOs, and error budgets, aligning platform telemetry to business reliability targets and engineering commitmentsLead post-incident reviews using observability evidence and drive corrective improvements in instrumentation, thresholds, dashboards, and alerting logic3. Telemetry Automation & Observability as CodeStandardize monitoring configurations using Terraform and/or Dynatrace Monaco, including alerting profiles, dashboards, SLOs, tagging rules, synthetic tests, and management zonesBuild automation for platform operations, integration workflows, reporting, and remediation using Python, Bash, or PowerShell, along with REST APIs and webhooks.5–10 years of experience in Observability, Monitoring Engineering, SRE, APM, DevOps, or Infrastructure Engineering, including several years of hands-on Dynatrace administration and architecture.Deep hands-on expertise with Dynatrace across full-stack monitoring, Davis AI, Smartscape, RUM, synthetic monitoring, distributed tracing, Grail log monitoring, DQL, management zones, Workflows/AutomationEngine, and access governance.Strong experience with AWS cloud services, especially CloudWatch, EKS, ECS, Lambda, EC2, RDS, API Gateway, networking, and modern cloud architecture patterns.Advanced knowledge of Kubernetes and cloud-native observability patterns, including instrumentation for microservices and distributed systems.Strong proficiency in observability-as-code using Terraform and/or Monaco, plus scripting in Python, Bash, or PowerShellSolid understanding of distributed application architecture, networking fundamentals, telemetry pipelines, performance engineering, and incident management.Dynatrace certification at Associate or Professional level required; higher-level certification is strongly preferred. Preferred Qualifications Experience with tools such as Nagios/SolarWinds/Prometheus/Grafana, Splunk, or ELKExperience with OpenTelemetry, Dynatrace Grail, advanced log analytics, and enterprise telemetry standardizationExperience integrating observability with ITSM or event-management platforms such as ServiceNowBackground in SRE practices such as reliability reviews, error budget management, and incident reduction programs.Integrate observability controls into CI/CD pipelines and establish telemetry quality standards for new application and infrastructure deploymentsUse Dynatrace Query Language (DQL) and Grail capabilities for advanced log analysis, event correlation, notebooks, and custom operational insights. 4. Governance, Cost Control & EnablementOwn monitoring governance practices related to telemetry quality, alert design, data retention, platform usage standards, and operational reportingManage Dynatrace usage and consumption responsibly by monitoring ingest patterns, tuning retention, and optimizing log, metric, and trace collection for value and efficiency.Build executive and engineering dashboards that communicate service health, reliability KPIs, error budgets, and infrastructure visibility to multiple audiencesMentor engineers and partner teams on observability best practices, onboarding, dashboarding, instrumentation, and platform self-sufficiency. This opportunity is for an existing vacancy with the company. The anticipated base salary range for this position is $85,000 to $125,000. Exact salary depends on several factors such as experience, skills, education, and budget. Salary range may vary based on geographic location. In addition to base salary, this position is eligible for participation in a bonus program. In addition, The Company offers a variety of benefits to eligible employees, including health insurance coverage, wellness programs, life and disability insurance, retirement savings plans, paid leave programs, education-related programs, paid holidays and vacation time, and many others. Many of these benefits are subsidized or fully paid for by the company. Financière CI est une société indépendante offrant des services de conseil en gestion de patrimoine et en gestion d’actifs à l’échelle mondiale par le biais de diverses sociétés de services financiers. Depuis 1965, nous anticipons les besoins changeants des investisseurs et y répondons de manière fiable. Nous sommes animés par la volonté d’offrir aux particuliers et aux institutions des investissements et des conseils de la plus haute qualité. Notre engagement à fournir les niveaux de rendement les plus élevés signifie que, peu importe leur poste, les employés de CI doivent être à l’aise dans un environnement trépidant qui les poussera à exploiter tout leur potentiel. Les employés qui font preuve d’un degré d’ambition élevé, d’une volonté de faire preuve de curiosité intellectuelle pour apprendre en permanence et d’une disposition à se dépasser s’épanouissent chez CI. Un environnement propice à la réussite Nous offrons un environnement de travail en présentiel, des avantages sociaux concurrentiels et un milieu de travail bienveillant, afin de permettre à nos employés de s’épanouir tant sur le plan personnel que professionnel. CE QUE NOUS OFFRONS Siège social moderne situé à distance de marche d’Union Station Remboursement de la formation Désignations professionnelles payées Régime d’épargne des employés (REE) Programme de rabais d’entreprise Avantages sociaux collectifs améliorés Programme de complément de congé parental Congés payés pour activités bénévoles Nous nous concentrons sur la création d’une main-d’œuvre diversifiée et inclusive. Si vous êtes enthousiaste à l’idée de ce poste et que vous n’êtes pas certain de répondre à toutes les exigences de qualification, nous vous encourageons à postuler pour en apprendre davantage sur l’occasion. Vous pouvez soumettre votre curriculum vitæ en toute confidentialité en cliquant sur « Postuler ». Nous ne communiquerons qu’avec les candidats sélectionnés pour une entrevue. CI Financial Corp. et toutes ses sociétés affiliées (« CI ») offrent un environnement de travail équitable et accessible. CI s’engage à prendre en compte les besoins d’accommodement pour les personnes handicapées. Si vous avez besoin de mesures d’adaptation pour postuler une offre d’emploi, ou si vous avez besoin que cette offre soit publiée dans un autre format, ou si ce besoin se fait sentir à toute autre étape du processus de recrutement, communiquez avec nous à l’adresse accessible.recruitment@ci.com, ou appelez le 416 364-1145, poste 4747.