Observability DevOps Engineer

Vertisystem — Canada · Posted ~2 days ago

Senior Contract Remote CA$55 per hour

Skills

DevOps SRE Cloud infrastructure Terraform Kubernetes Monitoring Python AWS Datadog

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary

Improve platform reliability by managing observability tools, automating operations, monitoring distributed systems, and supporting cloud-native infrastructure.

Highlights

Fully remote contract role focused on reliability, automation, cloud infrastructure, and large-scale observability practices.

Description

Job Title: Auth0 Engineer II, Observability Location: Fully Remote Role Duration: 07+ Months Contract with Possible Extension Pay Rate: $55 CAD Per Hour on T4 Job Description: The Client Platform Observability team owns the observability tooling that monitors the CIC Platform, and we are looking for an Observability Engineer to help ensure that our Product and Platform Engineers can monitor and observe our platform while continuing to rapidly ship software that our customers love. If you have experience within the Site Reliability Engineering (SRE) field or working as a Development Operations (DevOps) engineer, and you have a passion for Observability tooling, this position will allow you to further your learning and development in these areas. We are looking for engineers who are passionate about monitoring, observing, measuring uptime and availability, and ensuring stability for our platform. Our engineers maintain and automate observability tooling for our entire platform, including metrics, logs, and traces. Responsibilities: • Proficient in running services in production environments. • Contribute to the process of designing services for high growth and high availability. • Provision, configure, and monitor cloud-native infrastructure and services. • Automate key processes. Required Qualifications: • 5+ years of platform operations engineering, SRE, or DevOps experience. • Experience with cloud infrastructure like AWS, Google Cloud, or Azure. • Experience with Datadog (preferred) or other monitoring tools. • Experience with Sentry (preferred) or other error reporting tools. • Experience managing infrastructure with Terraform. • Proficiency in Golang, Node.js, or Python • Demonstrable expertise in monitoring distributed applications at scale. • Understanding of microservice architecture and best practices. • Experience with Kubernetes. • Team player who is willing to voice their opinion. You might work on: • Troubleshooting performance issues and operational issues. • Automating operational tasks and improving scripts. • Assisting with and providing feedback for performance testing and automation. • Provisioning infrastructure and services in collaboration with product engineers. • Collaborating with other engineering teams to help enhance their observability.