Summary
✨ AI‑Generated
Join a platform reliability team managing complex private and public cloud environments and enterprise middleware. You will containerize applications, operate and scale OpenShift and Kubernetes platforms, integrate CI/CD pipelines, monitor platform health, troubleshoot deployments and connectivity, and support large-scale upgrades and technology initiatives.
Highlights
Platform-focused SRE role working with cloud infrastructure, OpenShift and Kubernetes at scale, including containerization, CI/CD integration, monitoring, upgrades, security, performance optimization, and complex technology projects.
Description
• Manages complex platforms and frameworks like private and public cloud, middleware platforms like Business Process Management, Operational Decision Manager, and Process Federation Server.
• Supports large technology projects, hygiene activities, upgrades, rollouts at scale.
• Works closely with development teams to containerize applications, and ensuring the scalability, performance, and security of the OpenShift cluster, with expertise in Kubernetes concepts like deployments, pods, services, and CI/CD pipelines.
• Works closely with cross-functional teams to define standards, execute POCs, and integrate with CI/CD pipelines, utilizing tools such as UrbanCode Deploy, Harness, Jenkins, Artifactory, and HashiCorp products.
• Monitors platform health, identify and resolve issues related to application deployments, scaling, and network connectivity.
• Has deep understanding of Kubernetes core concepts like pods, deployments, services, ingress controllers, and scaling mechanisms.
• Ensures compliance with industry standards when deploying applications on the platform.
• Implements security best practices within the OpenShift environment, including role-based access control and image scanning.
• Utilizes DevOps tools such as Jenkins, Teamcity, GitHub, and Ansible to manage application deployments.
• Leads platforms enhancement activities - by performing automation of the processes using Python, Javascript.
• Provides guidance and support to development teams in adopting Kubernetes best practices and utilizing the Kubernetes platform effectively.
• Resolves production incidents quickly while maintaining service level agreements.
Provides on-call coverage to support major activities and events.
• Coordinates with application dev teams, vendors, and business, as necessary.
• Support release deployments and collaborate with application teams to deliver data quality and governance objectives.
Required Skills –
• IBM BPM/ODM experience
• WebSphere experience
• Scripting skills, one or more of the following (Bash, Ksh, perl, python)
• OS knowledge (Linux (CentOS RedHat), Windows)
• File transfer components and protocols
• SVN, GIT, BitBucket (version control systems)
• Kubernetes, Openshift experience
• Teamcity/Jenkins experience
• Distributed DB2 LUW