Kubernetes Platform Engineering Lead

Tagmatix360 — Poland · Posted ~22 hours ago

Lead

Skills

Kubernetes Platform engineering Platform lifecycle management Automation Infrastructure management System performance optimization Capacity planning Monitoring and dashboards Incident detection Self-healing systems System design Operational support Team leadership Dashboards Infrastructure

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A platform engineering leadership role responsible for the end-to-end lifecycle of an internal Kubernetes platform. The position combines automation, reliability engineering, performance optimization, dashboards, self-healing capabilities, operational support, system design, and capacity planning while supporting a small engineering team.

Highlights

Lead a focused platform engineering function, improve Kubernetes reliability and upgrade processes, build automation and self-healing capabilities, and influence platform architecture and capacity planning.

Description

Roles and RESPONSIBILITIES • Provide end-to-end platform engineering and lifecycle management on Internal Kubernetes Platform• Build tools and automation to manage platform infrastructure and services lifecycle. Develop platform operations to orchestrate all automation solutions, rather than maintaining isolated pipelines• Improve reliability, quality, and time to upgrade cluster and service versions• Measure and optimize system performance and resource utilization, and plan for future capacity• Build dashboards and visualizations to users for transparency of platform status• Detect system issue and develop self-healing solution where possible• Provide operational support and engineering for multiple software development teams.• Participate in system design consulting, platform management, and capacity planning• Implement infrastructure changes at weekends or out of normal operating hours.• Support members of the small team of 5 people responsible for the above tasks and manage their monthly schedule as well as collaborate with a Global Engineering Team Lead to properly manage the local team involvement in a global workload SKILLS & EXPERIENCE REQUIRED • Administrative knowledge with hands-on experience with container and Kubernetes infrastructure as well as management and operation experience on production environment• Engineering skillset to manage infrastructure by developing automation, API and pipeline• Core competencies: Linux OS, Container, Kubernetes and K8S native system like Istio, Cilium, CSI solution etc.• Ability to script in one or more shell languages, such as Bash and to program (structured and OO) with one or more high level languages, such as Python or Golang• Ability to use orchestration tools like Jenkins and Ansible (core & AAP) and observability tools like Splunk & Datadog• Experience in infrastructure lifecycle management and in development following software lifecycle using Github• ITIL awareness• A proactive approach to spot problems, areas for improvement, and performance bottlenecks• Proficiency in English. We are a diverse company. Ability to communicate on a business level with everyone is a must