Senior DevOps Engineer

Treb — Canada · Posted ~3 hours ago

Senior Full-time

Skills

AWS Linux Windows administration Networking Infrastructure management Automation Database administration Docker MySQL MariaDB

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A senior DevOps position responsible for cloud operations, infrastructure optimization, automation, troubleshooting, and improving system reliability.

Highlights

Key technical role focused on cloud infrastructure, reliability, automation, and solving complex production challenges.

Description

We are seeking an experienced Senior DevOps Engineer to join our technical operations and support team. The ideal candidate will possess deep expertise in AWS cloud services, strong system administration and troubleshooting skills across Linux and Windows environments, and a solid understanding of networking concepts and infrastructure management. Experience with LAMP stack environments (Linux, Apache, MySQL/MariaDB, PHP) for web application hosting and database integration is highly desirable. The candidate should also demonstrate proficiency in managing and optimizing relational database platforms, including Amazon Aurora, Microsoft SQL Server, MySQL, and MariaDB. This role requires a proven ability to diagnose and resolve complex, mission-critical production issues, with a strong focus on root cause analysis, automation, reliability, and continuous improvement. The Senior DevOps Engineer will serve as a key technical escalation point, collaborate closely with development, infrastructure, and security teams, mentor junior engineers, and drive operational excellence through infrastructure as code, CI/CD best practices, monitoring, and process optimization. What You’ll Do: Resolve complex, multi-service AWS technical issues (e.g., EC2, S3, VPC, RDS, Lambda). Design, implement, and maintain CI/CD pipelines. Deploy and manage containerized applications using Docker. Develop and maintain infrastructure using AWS CloudFormation / Terraform. Perform deep-dive troubleshooting and Root Cause Analysis (RCA) for critical incidents. Support core areas: networking, Linux/Windows, and cloud security best practices. Configure and manage relational database systems, including Amazon Aurora, SQL Server, MySQL, and MariaDB. Install and configure web servers including Apache and OpenLightSpeed on Ubuntu/Linux, PHP, and related modules. Collaborate with AWS engineering teams to escalate bugs and drive improvements. Communicate clearly with stakeholders during high-impact events. Advise IT teams on AWS architecture, performance, and security best practices.Mentor junior engineers and contribute to team knowledge through documentation. Provide advanced technical support for complex production issues, serve as an escalation point for critical incidents, conduct root cause analysis, and ensure the timely resolution of operational and staff support requests. Drive automation and process improvements using scripting (Python, Bash, PowerShell). Identify trends and provide feedback to improve services and reduce case volume. Collaborate with cross-functional teams to translate business needs into technical solutions. Ensure platform performance, security, and compliance with best practices and regulations. Maintain documentation and support training efforts for users and staff. What You Bring: 10+ years in Cloud support, DevOps, or system administration, with a strong focus on AWS or other public cloud platforms. In-depth expertise in core AWS services (EC2, S3, VPC, RDS, Lambda, DynamoDB) and troubleshooting complex inter-service issues. Experience in configuring and deploying web servers including Apache and OpenLightSpeed on Ubuntu/Linux, PHP, and related modules. Familiarity with LAMP stack environments (Linux, Apache, MySQL/MariaDB, PHP) for web application hosting and database integration. Proficiency in managing relational database systems, including Amazon Aurora, SQL Server, MySQL, and MariaDB. Responsibilities include installation, configuration, performance tuning, backup and recovery, and ongoing maintenance to ensure high availability and security. Advanced system administration skills (Linux and/or Windows), including performance tuning and kernel-level troubleshooting. Strong networking knowledge (TCP/IP, DNS, VPN, Load Balancing) and experience resolving cloud networking issues. Proficient in scripting (Python, Bash, PowerShell) for automation, diagnostics, and tooling. Excellent communicator, able to clearly explain technical concepts to both technical and non-technical audiences. Collaborative and customer-focused, with experience mentoring peers and working across teams. Agile development experience and familiarity with DevOps tools (GitHub, GitLab, CI/CD pipelines). Experience with production Kubernetes deployments and implementing enterprise-scale CI/CD pipelines. Excellent communication and problem-solving skills with a collaborative mindset. Why Join TRREB? Be part of a well-established, mission-driven organization with real community impact.Work in a safe, professional, and collaborative environment.Competitive compensation based on experience.Opportunity to grow your professional excellence.On-site free parkingWork one day/week from home. Ready to make your mark? Apply now and bring your skills to a team that values innovation and results.