Site Reliability Engineering Manager

Sabre Corporation — Poland · Posted ~2 days ago

Lead Full-time Hybrid

Skills

Site Reliability Engineering People leadership Cloud engineering Large-scale distributed systems Observability Monitoring Alerting Incident management Capacity planning Reliability engineering Automation AI-assisted operations English Google Cloud Platform Kubernetes Terraform CI/CD Infrastructure as Code Observability platforms AI-assisted tooling Generative AI

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary

Lead a Site Reliability Engineering team responsible for the availability, scalability, security, and operational excellence of a large-scale, customer-facing cloud platform. You will develop engineers, drive reliability strategy, manage major incidents, establish service objectives, modernize cloud operations, and expand automation and AI-enabled practices to reduce operational effort and improve resilience.

Highlights

Leadership of a high-impact reliability team, broad cloud and automation scope, strong professional development support, certification and tuition reimbursement, extensive health and family benefits, paid volunteer time, and additional year-end leave.

Description

Location: Kraków (hybrid work - 3 days per week in the office) Manager, Site Reliability Engineering About The Team The Site Reliability Engineering (SRE) team for Sabre Web Services (SWS) is responsible for the reliability, scalability, security, and operational excellence of one of Sabre's most critical technology platforms. SWS provides the API gateway and connectivity infrastructure that enables airlines, travel agencies, developers, and Sabre applications to securely access Sabre services. These systems serve as the front door to major Sabre products and process a significant portion of Sabre's overall transaction traffic. The team combines software engineering and operations expertise to build highly available, fault-tolerant, cloud-native services. Engineers work extensively with Google Cloud Platform (GCP), Kubernetes, Infrastructure as Code, CI/CD automation, observability platforms, and AI-driven operational practices to improve reliability while reducing operational toil. The organization partners closely with software engineering, product, architecture, security, and customer-facing stakeholders across Sabre. What You'll Do As Manager of Site Reliability Engineering, you will lead a team of SRE engineers responsible for the availability, performance, resiliency, and continuous evolution of the Sabre Web Services platform. You will drive operational excellence across Sabre's API gateway ecosystem, ensuring customers can reliably access critical travel services at global scale. You will lead incident management, reliability engineering initiatives, cloud modernization efforts, and automation programs while developing a high-performing engineering team. Lead and develop a team of Site Reliability Engineers, providing coaching, performance management, career development, and technical leadership. Drive service reliability, availability, scalability, and operational excellence across critical production systems and customer-facing platforms. Partner with Engineering, Product, Architecture, and Infrastructure teams to establish and achieve Service Level Objectives (SLOs) and reliability targets. Lead major incident response, postmortem reviews, risk mitigation initiatives, and continuous improvement programs. Promote automation, AI-enabled operations, self-healing capabilities, and platform engineering practices to reduce operational toil and improve efficiency. Required Qualifications And Education Bachelor's degree in Computer Science, Software Engineering, Information Technology, Computer Engineering, or a related technical field required. 5+ years of experience in Software Engineering, Site Reliability Engineering, DevOps, Cloud Engineering, or a related technical discipline. 2+ years of people leadership experience, including managing engineers, setting priorities, and developing high-performing teams. Experience in maintaining or operating large-scale distributed systems in cloud environments. Understanding of observability, monitoring, alerting, incident management, capacity planning, and reliability engineering principles. Practical experience leveraging automation, analytics, AI-assisted tooling, or generative AI solutions to improve operational effectiveness and engineering productivity. Fluent English, both written and verbal Nice-to-have Qualifications Knowledge of Kubernetes, Terraform, CI/CD pipelines, Infrastructure as Code, and cloud-native architecture. Experience implementing SLOs, error budgets, reliability scorecards, and operational maturity programs. Background supporting high-availability, customer-facing SaaS or travel technology platforms. Benefits Paid time off Year-End-Break: enjoy additional fully paid days off during the last week of the yearPaid parental leave: Take up to 12 weeks off with pay after birth or adoption of a child. Sabre Global Paid Parental Leave runs concurrently with local leave policies.Paid volunteer time: take up to 4 days annually to give your time to a charitable organization of your choice Your money My Benefit platform/Multisport card: enjoy the benefit cafeteria system and use popular sport cardTax deduction: take the opportunity to claim deductible costs, reducing your income taxEmployee Capital Plans: profit from long-term saving scheme co-financed by Sabre and the State TreasuryBaby Bonus: benefit from one-time allowance on childbirth or adoption Say Thanks program: collect points on recognition program and transfer them to wide variety of gifts and services Health and wellness Luxmed extensive medical coverage: take care of yourself and your family with the extensive medical package with a broad range of additional servicesForeign travel insurance: feel safe going abroad with free Allianz insurance offered as part of our Lux Med packageEmployee Assistance Program: find help in free, confidential program with a certified counselorLife insurance: sign up for free, high coverage life insurance program Career development Professional development: enjoy access to Udemy learning platform as well as join Sabre live learning sessionsCertification and tuition reimbursementOur Communities: join one of our team member groups focused on sharing knowledge and best practices (Google Developers Group, Innovation Lab Community, Women in Technology, SOLVE!T and many more) And more Car and bike parking Fun & Relax zone in modern office: enjoy electronic tables to work, foosball, ping pong, pool table, swings, massage chairs and terraces to admire a panoramic view of Kraków. We have parents’ rooms as wellNo dress codeInnovation Lab: access Augmented Reality & Virtual Reality equipment, Robot construction kit, 3D printers and many moreAttractive Referral Bonus: earn $2500 USD for every hired referral