Site Reliability Engineer

Vaspire Technologies Inc — United States · Posted ~2 hours ago

Senior Full-time

Skills

Site reliability engineering Kubernetes OpenShift cloud platforms containerization CI/CD platform monitoring scalability performance security troubleshooting UrbanCode Deploy Harness Jenkins Artifactory HashiCorp

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Join a platform reliability team managing complex private and public cloud environments and enterprise middleware. You will containerize applications, operate and scale OpenShift and Kubernetes platforms, integrate CI/CD pipelines, monitor platform health, troubleshoot deployments and connectivity, and support large-scale upgrades and technology initiatives.

Highlights

Platform-focused SRE role working with cloud infrastructure, OpenShift and Kubernetes at scale, including containerization, CI/CD integration, monitoring, upgrades, security, performance optimization, and complex technology projects.

Description

• Manages complex platforms and frameworks like private and public cloud, middleware platforms like Business Process Management, Operational Decision Manager, and Process Federation Server. • Supports large technology projects, hygiene activities, upgrades, rollouts at scale. • Works closely with development teams to containerize applications, and ensuring the scalability, performance, and security of the OpenShift cluster, with expertise in Kubernetes concepts like deployments, pods, services, and CI/CD pipelines. • Works closely with cross-functional teams to define standards, execute POCs, and integrate with CI/CD pipelines, utilizing tools such as UrbanCode Deploy, Harness, Jenkins, Artifactory, and HashiCorp products. • Monitors platform health, identify and resolve issues related to application deployments, scaling, and network connectivity. • Has deep understanding of Kubernetes core concepts like pods, deployments, services, ingress controllers, and scaling mechanisms. • Ensures compliance with industry standards when deploying applications on the platform. • Implements security best practices within the OpenShift environment, including role-based access control and image scanning. • Utilizes DevOps tools such as Jenkins, Teamcity, GitHub, and Ansible to manage application deployments. • Leads platforms enhancement activities - by performing automation of the processes using Python, Javascript. • Provides guidance and support to development teams in adopting Kubernetes best practices and utilizing the Kubernetes platform effectively. • Resolves production incidents quickly while maintaining service level agreements. Provides on-call coverage to support major activities and events. • Coordinates with application dev teams, vendors, and business, as necessary. • Support release deployments and collaborate with application teams to deliver data quality and governance objectives. Required Skills – • IBM BPM/ODM experience • WebSphere experience • Scripting skills, one or more of the following (Bash, Ksh, perl, python) • OS knowledge (Linux (CentOS RedHat), Windows) • File transfer components and protocols • SVN, GIT, BitBucket (version control systems) • Kubernetes, Openshift experience • Teamcity/Jenkins experience • Distributed DB2 LUW