Senior / Lead Azure DevOps Engineer

Saransh Inc Usa โ€” Canada ยท Posted ~2 weeks ago

Lead Contract Remote

Skills

Azure Dynatrace Kubernetes Splunk Grafana Azure Monitor Log Analytics Application Insights Observability JIRA

๐Ÿ”“ Log in to save this job, tailor your resume & track your apply process โ€” 7 days free, no card needed.

Log in to add to target list

Summary

Lead observability and monitoring initiatives for large-scale production systems. Improve operational reliability, manage monitoring platforms, resolve incidents, and establish best practices across cloud environments.

Highlights

Fully remote role in Canada, long-term contract, ownership of observability strategy, and work on high-availability production systems.

Description

Role: Senior / Lead Azure Devops Engineer (Observability & Monitoring) Fully Remote in Canada (CST Time zone shift) Job Type: Contract Duration: 1 year Key Notes From Client We really need an expert in Dynatrace, this is probably the main requirement here.Proven Dynatrace and Kubernetes skills are a must, plus experience with Azure cloud. Description Join the Client as a Lead DevOps Engineer where you will drive observability excellence, strengthen monitoring practices, and ensure system reliability across a complex production environment. Your expertise will directly improve operational maturity and support delivery at scale. Responsibilities Manage end-to-end monitoring, alerting, and observability for production applications using Dynatrace, Splunk, and Grafana, ensuring system reliability and performance Develop and maintain documentation outlining best practices for logging and monitoring, while conducting regular audits to ensure compliance with company policies and industry standards Triage, prioritize, and resolve medium-complexity incidents and service requests using JIRA, providing warm handoff notes and escalation support for tickets beyond Level 2 Define service level objectives for each product request type and establish average completion time benchmarks to drive continuous improvement Present metrics and escalated ticket trends regularly to leadership, identifying opportunities to shift support processes left and reduce incident recurrence Participate in cross-functional initiatives related to observability standards and be available for off-hours monitoring, on-call escalations, and pager duty coverage Requirements 8 or more years of experience working within DevOps or SRE teams, with a strong focus on incident and request management in high-availability production environments Hands-on proven expertise in Dynatrace, Splunk, and Grafana, with the ability to configure monitoring setups and administer tooling Experience with Azure logging and monitoring solutions including Log Analytics, Azure Monitor, and Application Insights Solid understanding of observability principles covering monitoring, logging, and distributed tracing across scalable, fault-tolerant systems Strong analytical and problem-solving skills with the ability to troubleshoot complex issues under pressure and communicate findings clearly to technical and non-technical stakeholders