Site Reliability Engineer / DevOps Specialist

Sita — Jordan · Posted ~21 hours ago

Full-time

Skills

Site reliability engineering DevOps Production support System monitoring Performance optimization Incident management Continuous improvement Monitoring Performance management

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A site reliability and DevOps opportunity focused on proactively supporting critical technology products, maintaining high performance, and continuously improving operational reliability. The role involves identifying issues, optimizing systems, and helping ensure dependable services for a global transportation technology environment.

Highlights

Work on technology that supports critical global transportation infrastructure. The role emphasizes proactive reliability, high system performance, continuous improvement, professional development, and a supportive workplace culture.

Description

Overview WELCOME TO SITA At SITA, we keep airports moving, airlines flying smoothly, and borders open. Our technology and communication innovations power the success of the global air travel industry. You'll find us in 95% of international airports, working closely with over 2,500 transportation and government clients. Each partnership brings unique challenges, and we thrive on delivering fresh solutions and cutting-edge tech to keep operations running like clockwork. We don't just move the world forward-we're proud to be recognized as a Great Place to Work® by 79% of our employees and certified in most of our growing locations. Here, we feel empowered, supported, and inspired to grow. Are you ready to love your job? The adventure begins right here, with you, at SITA. About The Role & Team The Site Reliability Engineer is responsible for the proactive support of products to ensure high product performance, with a continuous focus on improvement. The role involves identifying and resolving the root causes of operational incidents, implementing solutions to enhance stability, and preventing recurrence. The Site Reliability Engineer manages the creation and maintenance of the event catalogue to trigger events and develops both manual remediation approaches and automated workflows to address alerts. Additionally, they oversee the deployment of IT services and solutions, ensuring seamless integration with minimal disruption. What You’ll Do Design, build, and maintain support systems to ensure high availability, scalability, and performance of critical infrastructure.Lead incident response and root cause analysis for system failures, including problem investigations and coordination with relevant teams.Implement and manage automation for system provisioning, deployment, self-healing, and performance monitoring to increase operational efficiency.Establish and monitor SLIs/SLOs, proactively identify performance issues, and drive continuous improvements in service reliability.Collaborate with development and operations teams to embed reliability best practices and evolve toward zero-downtime architecture.Manage and optimize an event catalog, including event definitions, thresholds, remediation actions, and relevance across products.Develop event response protocols, provide training, and ensure efficient handling of incidents across teams.Drive post-incident reviews and feedback loops to enhance event definitions and service reliability.Oversee quality and readiness of deployments, ensuring clear processes, assigned responsibilities, and minimal disruption.Maintain deployment schedules and conduct risk assessments to ensure operational stability and deployment readiness.Coordinate and execute deployment plans, manage resources, and incorporate feedback for continuous process improvement.Manage CI/CD pipelines and infrastructure as code, ensuring seamless integration between development and operations.Support and evolve DevOps practices, automating operational tasks and maintaining tools to drive ongoing efficiency. Qualifications Who You Are: Bachelor’s degree in computer science, Information Technology, Engineering, or a related field.Several years of experience in IT operations, service management, or infrastructure management, including roles such as Site Reliability Engineer, Problem Manager, or DevOps Manager.Proven experience in managing high-availability systems and ensuring operational reliability.Extensive experience in root cause analysis (RCA), incident management, and developing permanent solutions for recurring service disruptions.Hands-on experience with CI/CD pipelines, automation, system performance monitoring, and the implementation of infrastructure as code.Strong background in collaborating with cross-functional teams (development, operations, engineering, etc.) to improve operational processes and service delivery.Experience in managing deployments, risk assessments, and optimizing event and problem management processes.Familiarity with cloud technologies, containerization, and scalable architecture, including experience with zero-downtime deployment strategies. CompTIA Security+ or Certified Kubernetes Administrator (CKA). Functional Skills CollaborationStakeholder ManagementService Design CommunicationProblem SolvingIncident ManagementChange ManagementInnovation ; Technical Skills Cloud InfrastructureAutomation & AIOperations Monitoring & DiagnosticsDeploymentProgramming & Scripting Languages What We Offer We're all about diversity. We operate in 200 countries and speak 60 different languages and cultures. We're really proud of our inclusive environment. Our offices are comfortable and fun places to work, and we make sure you get to work from home too. Find out what it's like to join our team and take a step closer to your best life ever. 🏡 Flex Week: Work from home up to 2 days/week (depending on your team's needs) ⏰ Flex Day: Make your workday suit your life and plans. 🌎 Flex-Location: Take up to 30 days a year to work from any location in the world. 🌿 Employee Wellbeing: We have got you covered with our Employee Assistance Program (EAP), for you and your dependents 24/7, 365 days/year. We also offer Champion Health - a personalized platform that supports a range of wellbeing needs. 🚀 Professional Development: At SITA, we believe growth fuels innovation. Our learning ecosystem offers access to world-class platforms and programs designed to help you thrive. From LinkedIn Learning, Microsoft's Enterprise Skills Initiative, and Airport Council International -available to all employees-to specialized solutions like Pluralsight for technology upskilling, Harvard Business Publishing for people leadership, Stanford for strategic development and many others, we align learning opportunities with your Development Plan and our business priorities. Your development journey is supported every step of the way. 🙌 Competitive Benefits: Competitive benefits that make sense with both your local market and employment status. SITA is an Equal Opportunity Employer. We value a diverse workforce. In support of our Employment Equity Program, we encourage women, aboriginal people, members of visible minorities, and/or persons with disabilities to apply and self-identify in the application process. Salary / Compensation Note Hidden (-999)