Kubernetes Platform Support Engineer

Thrive It Systems Ltd — Poland · Posted ~21 hours ago

Skills

Kubernetes Kubernetes infrastructure support Troubleshooting Root cause analysis Containerization Incident management Monitoring and observability ITIL processes On-call support Containers Monitoring Logging

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Join a platform engineering team responsible for keeping a critical internal Kubernetes environment reliable, available, and performant. You will troubleshoot incidents, analyze platform metrics and logs, improve resilience, support containerized workloads, contribute to automation, and participate in operational support rotations. The role suits an engineer who enjoys hands-on infrastructure operations and continuous reliability improvements.

Highlights

Work on a critical Kubernetes platform with strong focus on reliability, resilience, automation, and continuous improvement. The role offers broad exposure to infrastructure operations, incident management, observability, and collaboration with engineering teams.

Description

Responsibilities: Ensure the reliability availability and performance of the Internal Kubernetes PlatformSupport the deployment configuration and maintenance of Kubernetes infrastructureMonitor platform health and proactively identify reliability and performance issuesTroubleshoot and resolve incidents platform failures and integration issuesPerform root cause analysis and implement reliability improvementsCollaborate with engineering and infrastructure teams to improve platform stability and resilienceSupport containerization and orchestration services for internal customersParticipate in 24x7 oncall support rotations and weekend support activitiesContribute to automation initiatives to improve operational efficiencySupport infrastructure changes following ITIL processes and governance controlsAnalyze logs monitoring dashboards and platform metrics to identify issues and trendsDrive continuous improvement of Kubernetes platform operations and support processes Requirments: Handson Kubernetes administration experienceStrong understanding of Kubernetes architecture concepts and operationsExperience troubleshooting Kubernetes clusters and containerized environmentsKnowledge of containerization and orchestration technologiesStrong UnixLinux administration experienceUnderstanding of Site Reliability Engineering SRE principlesExperience with monitoring logging and observability toolsStrong analytical and troubleshooting skillsUnderstanding of ITIL processes including Incident Problem and Change ManagementExperience with Infrastructure as Code IaC concepts and automationStrong communication and stakeholder collaboration skillsAbility to work effectively across multiple teams and global environmentsExperience with Service Mesh technologies is desirableProficiency in English communication Skills Mandatory Skills : Kubernetes, Ubuntu Linux Administrator