Senior SRE / Observability Engineer

Iris Software Inc. — Canada · Posted ~4 hours ago

Senior Hybrid

Skills

Dynatrace Observability APM Infrastructure monitoring Digital experience monitoring Log analytics CI/CD integrations ITSM integrations Cloud platforms Root-cause analysis Incident prevention CI/CD ITSM Cloud

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Take ownership of an enterprise observability strategy as a senior reliability specialist. You will design, deploy, and optimize full-stack monitoring, configure advanced dashboards and alerting workflows, integrate observability with delivery and service-management systems, and turn monitoring data into actionable insights. You will also lead root-cause analysis and proactive incident prevention in a hybrid environment.

Highlights

Own an end-to-end observability strategy and serve as a technical authority for monitoring and alerting. The role offers significant ownership across APM, infrastructure, digital experience, logs, integrations, dashboards, incident prevention, and root-cause analysis.

Description

Iris's Fortune 100 direct client is looking Senior SRE/Observability Engineer (Dynatrace). Please find below Job description and share me your updated resume at Jatin.gupta@irissoftware.com. Job Title: Senior SRE/Observability Engineer (Dynatrace) Location: Toronto, ON (Hybrid, 3 days onsite in a week) Required Skills: Own the end-to-end Dynatrace observability strategy, leading design, deployment, and optimization of full-stack monitoring (APM, infrastructure, digital experience, log analytics) across environmentsServe as the technical authority for Dynatrace configuration, custom dashboards, problem/alerting workflows, and integrations (CI/CD, ITSM, cloud platforms)Partner with application, infrastructure, and business stakeholders to translate monitoring data into actionable insights, drive proactive incident prevention, and lead root-cause analysis for critical performance issues