Senior Platform Engineer

Astra North Infoteck Inc โ€” Canada ยท Posted ~3 weeks ago

Senior Full-time Hybrid

Skills

DevOps Site Reliability Engineering Dynatrace Observability Cloud platforms Automation Incident management Monitoring SRE practices AWS Azure GCP Kubernetes

๐Ÿ”“ Log in to save this job, tailor your resume & track your apply process โ€” 7 days free, no card needed.

Log in to add to target list

Summary

A technology organization is seeking an experienced platform engineer to design and operate highly available infrastructure. The role focuses on reliability engineering, monitoring solutions, automation, cloud environments, incident response, and improving enterprise application performance.

Highlights

Opportunity to work on large-scale reliable infrastructure with a strong focus on automation, observability, cloud technologies, and operational excellence.

Description

Job Role: Senior Platform Engineer โ€“ DevOps, SRE & Dynatrace Observability Location: Toronto - Hybrid (3 Days Work from Office) Job Summary We are seeking a highly skilled and experienced Platform Engineer with strong DevOps and Site Reliability Engineering (SRE) expertise to join our technology team. The ideal candidate will be responsible for designing, implementing, automating, and supporting enterprise-scale platform infrastructure while ensuring high availability, reliability, scalability, and performance of mission-critical applications. The candidate must possess strong hands-on experience in Dynatrace monitoring, dashboard creation, synthetic monitoring, observability, automation, incident management, SRE best practices, and cloud/platform engineering. This role requires a proactive engineer who can drive operational excellence through automation, monitoring, reliability engineering, and continuous improvement initiatives. Key Responsibilities Platform Engineering & Reliability Design, build, and maintain highly available, scalable, and resilient platform infrastructure. Implement modern Platform Engineering and SRE practices across enterprise applications. Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Error Budgets. Drive reliability, availability, capacity planning, performance optimization, and operational excellence initiatives. Support production environments and participate in on-call rotations when required. Observability & Monitoring Lead implementation and administration of enterprise monitoring and observability solutions. Develop and maintain Dynatrace monitoring strategies for complex distributed systems. Create and manage Dynatrace dashboards, alerts, management zones, and reporting solutions. Implement proactive monitoring for infrastructure, middleware, applications, databases, APIs, and cloud services. Configure and optimize anomaly detection, problem management, and root cause analysis capabilities. Dynatrace Expertise Hands-on experience with: Dynatrace OneAgent deployment Dynatrace SaaS/Managed environments Dynatrace Dashboards Synthetic Monitoring Real User Monitoring (RUM) Digital Experience Monitoring (DEM) Distributed Tracing Davis AI Engine Custom Metrics and Extensions Service Flow Analysis Log Monitoring and Analytics Build advanced observability solutions using Dynatrace platform capabilities. Develop synthetic tests for business-critical applications and APIs. Design executive, operational, and application performance dashboards. Preferred Qualifications Experience in BFSI (Banking, Financial Services, and Insurance) domain. Dynatrace Associate/Professional certification. Cloud Certifications (AWS, Azure, or GCP). Kubernetes Administrator (CKA) certification. ITIL Foundation certification. Knowledge of Security, Compliance, and Regulatory requirements in enterprise environments.