Senior/Lead Site Reliability Engineer

Infoletitsolutions — Poland · Posted ~3 hours ago

Lead Full-time Hybrid Visa History ✓ 150-170 PLN/h (B2B); 18000-20500 PLN gross/month (UoP)

Skills

Kubernetes Kubernetes deployment and maintenance Infrastructure troubleshooting Root cause analysis Monitoring and observability Log analysis Performance troubleshooting Infrastructure reliability Monitoring Observability Infrastructure

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A Senior/Lead Site Reliability Engineer will join a global engineering team responsible for keeping a critical Kubernetes platform reliable, available, and performant. The role involves deploying and maintaining Kubernetes infrastructure, diagnosing complex incidents, analyzing logs and metrics, performing root-cause analysis, and implementing durable reliability improvements. The position combines hands-on production engineering with cross-team collaboration and offers a hybrid working arrangement.

Highlights

Senior/Lead opportunity focused on production-critical Kubernetes infrastructure, with strong ownership of reliability, performance, incident resolution, and long-term platform improvements. Offers a hybrid work model and competitive compensation.

Description

SUMMARY 💰 150 - 170 PLN/h (B2B) 💰 18 000 – 20 500 PLN brutto (UOP) 📍 Kraków (Hybrid - 2 days office / 3 days remote) 💼 Senior / Lead Project Join a global Site Reliability Engineering team responsible for ensuring the reliability, availability, and performance of an enterprise Kubernetes platform. You will work on production-critical infrastructure, supporting Kubernetes deployments, troubleshooting complex incidents, improving platform resilience, and driving reliability practices across a global engineering organization. You will Ensure the reliability, availability, and performance of the Kubernetes infrastructure platformSupport the deployment, configuration, and maintenance of KubernetesDiagnose and resolve infrastructure incidents, performance issues, and integration failuresPerform root cause analysis and implement long-term reliability improvementsAnalyze logs, monitoring data, and platform metrics to identify potential issuesCollaborate with engineering and infrastructure teams to improve platform resilienceParticipate in a 24/7 on-call rotation, including weekend supportCoordinate a local team of 5 engineers, including monthly schedules and on-call rotationsWork closely with the Global SRE Team Lead to coordinate the local team's contribution to global workloads Must have 10+ years of overall IT/infrastructure experience3+ years of hands-on experience with Kubernetes administrationStrong understanding of Kubernetes concepts, operations, and troubleshootingGood knowledge of containers and orchestrationSolid Unix/Linux administration skillsExperience troubleshooting production infrastructure and analyzing logs and monitoring dataUnderstanding of ITIL processes, particularly Incident, Problem, and Change ManagementKnowledge of infrastructure automation and Infrastructure as CodeStrong analytical and problem-solving skillsExcellent communication and collaboration skillsWillingness to participate in 24/7 on-call and weekend supportFluent English and Polish Nice to have Experience with Service Mesh technologiesExperience with Kubernetes platform engineeringKnowledge of advanced infrastructure automationExperience leading or coordinating an engineering/SRE teamExperience working in large-scale enterprise or regulated environments Our offer Relocation package (4500 PLN total value), paid in three installments (1500 PLN per month) if your permanent presence in the office is mandatory and you need to relocate from another city.Benefits: Extended medical care (over 2000 medical facilities in Poland, 80 in Kraków) for you and your family; Multisport Benefit card; life insurance