Senior Software Engineer, Site Reliability Engineering
Jobgether — Canada · Posted ~7 hours ago
🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.
Log in to add to target listDescription
This position is listed on behalf of a partner company, who manages all applications and next steps.
Our partner is looking for a Senior Software Engineer, Site Reliability Engineering based in Canada.
This is a senior-level opportunity to help build and operate the reliable, secure, and highly scalable infrastructure behind a high-impact digital platform.
You’ll work across the technology stack, from Linux systems and cloud infrastructure to distributed applications and platform services.
The role combines hands-on engineering with architectural leadership, enabling product and infrastructure teams to deliver resilient systems at scale.
You’ll collaborate closely with engineers across product development, developer experience, and backend infrastructure.
Your work will directly influence system availability, performance, observability, and the overall engineering experience.
You’ll also help shape platform capabilities, improve operational practices, and anticipate future capacity and reliability needs.
This is an ideal environment for an experienced SRE who enjoys solving complex technical problems and creating high-leverage engineering solutions.
Accountabilities
Design, develop, and maintain software and infrastructure that improve service availability, scalability, performance, and operational efficiency.
Establish architectural direction for infrastructure and platform services while providing technical guidance and support to engineering teams.
Build and improve tools, processes, and systems for deployment, infrastructure, service, and change management.
Troubleshoot and resolve complex production issues across the software development lifecycle, with a focus on minimizing downtime and service disruption.
Develop and evolve platform capabilities that enable engineering teams to build, deploy, operate, and observe services more effectively.
Conduct capacity planning and demand forecasting to anticipate system growth, identify performance bottlenecks, and proactively address scalability challenges.
Instrument, operate, and monitor distributed microservices and cloud-based systems to maintain strong reliability and observability.
Participate in a rotating on-call schedule and contribute to incident response, service recovery, and continuous reliability improvements.
Partner with cross-functional engineering teams and stakeholders to identify opportunities, balance technical trade-offs, and deliver high-impact platform solutions.
Requirements
5+ years of experience managing infrastructure and systems, ideally within large-scale or distributed production environments.
Extensive hands-on expertise with AWS and Linux-based systems.
Strong ability to read, write, debug, and maintain production-facing software and systems.
Deep understanding of large-scale distributed systems and web technologies, including DNS, TLS, HTTP/S, TCP/IP, and related networking concepts.
Demonstrated experience operating, instrumenting, and observing distributed microservices in production cloud environments.
Strong problem-solving skills, with the ability to break down complex technical challenges and make thoughtful trade-offs based on business and engineering impact.
Experience designing resilient, scalable infrastructure and platform services with a focus on availability, performance, and reliability.
Strong communication and collaboration skills, with the ability to work effectively with technical and non-technical stakeholders across different levels of an organization.
Ability to operate independently while contributing effectively to a collaborative, cross-functional engineering environment.
A proactive mindset and strong ownership of production systems, operational excellence, and continuous improvement.
Benefits
Remote work opportunity from eligible locations in Ontario and British Columbia, Canada.
Expected total cash compensation of CAD $180,200–$233,200, depending on location, qualifications, skills, competencies, and experience.
Opportunity to work on large-scale, distributed systems with significant impact across the engineering organization.
Collaborative environment with exposure to product engineering, developer experience, backend infrastructure, and platform teams.
Opportunity to influence architectural direction and shape the evolution of engineering infrastructure and platform capabilities.
Participation in meaningful reliability, scalability, observability, and infrastructure initiatives.
Commitment to diversity, inclusion, and equal employment opportunity.
Reasonable accommodations available throughout the recruitment process for candidates who require them.
How Jobgether Works
We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements.
Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company.
The final decision and next steps (interviews, assessments) are managed by their internal team.
We appreciate your interest and wish you the best!
Why Apply Through Jobgether?
Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer.
This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR).
You may exercise your rights (access, rectification, erasure, objection) at any time.
We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information.
These tools assist our recruitment team but do not replace human judgment.
Final hiring decisions are ultimately made by humans.
If you would like more information about how your data is processed, please contact us.
We have 153,059 jobs that might be an even better fit for you
DontApply's real value goes far beyond a single job link or company name. Just upload your resume — in under a minute we'll analyze all 153,059 jobs and tell you exactly which ones you should apply to right now.
Upload My Resume