Lead Platform Engineer

Insight Global — United States · Posted ~2 hours ago

Lead Full-time Onsite $120000-$155000

Skills

Platform engineering Kubernetes Docker Rancher Kubernetes cluster lifecycle management AWS GCP Azure Infrastructure automation Monitoring Log management Terraform Nagios Splunk ELK

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Lead platform engineering for complex infrastructure spanning on-premises and public-cloud environments. You will directly own Kubernetes platforms and their full lifecycle, automate interactions with physical infrastructure, manage cloud environments, and operate monitoring and logging systems. The position requires a bachelor's degree or substantial equivalent experience and offers relocation assistance.

Highlights

Lead platform engineering in a fully onsite role with substantial compensation and stock options, plus relocation assistance. The role offers deep hands-on ownership of Kubernetes platforms across on-premises and cloud environments.

Description

Title: Lead Platform Engineer Location: Cleveland, OH Schedule: 5 days onsite Salary: $120,000-$155,000 with stock options (typically based on years of experience) Relocation Assistance available. MUST HAVES: Bachelor's Degree in Computer Science or a related field. In lieu of a degree, at least 9 years of experience in the role of Platform Engineer or related position6+ years of experience in platform engineering and infrastructure components [with a strong focus on containerization technologies, including Docker, Kubernetes, and Rancher.]Directly owned and operated Kubernetes platforms (not simply deployed applications into them)Hands-on experience with Kubernetes cluster lifecycle management (both on-prem and Cloud environments)Expertise in automated interaction with physical infrastructureProven track record in managing and maintaining cloud-based environments such as AWS, GCP, or AzureExpertise in monitoring and log management tools such as Nagios, Splunk, and ELKProven track record in developing and implementing backup and disaster recovery strategiesExpertise in scripting experience in languages such as Python, PowerShell, and BashExpertise in Infrastructure as Code (IaC) and configuration management tools such as Ansible, Chef, and Terraform PLUSSES: Azure Solutions Architect, Certified Kubernetes Administrator (CKA), or AWS Certified DevOps Engineer preferred3+ years of hands-on experience with containerization technologies, including Docker, and Kubernetes and Rancher for orchestration and cluster managementProven expertise in maintaining platform-level runtimes, including ingress controllers (e.g., NGINX), TLS certificate lifecycle management (e.g., cert-manager), and secrets management integrations (e.g., HashiCorp Vault)Experience deploying and managing observability tools, such as Sysdig for monitoring and CVE scanning, Fluentd for log forwarding, and Elasticsearch for log analysisFamiliarity with GitOps practices and tools, particularly ArgoCD, to support continuous deployment workflows Day to Day: Insight Global is seeking a Lead Platform Engineer to work in Cleveland, OH! The Lead Platform Engineer role is responsible for setting strategic priorities for the maintenance of platforms within the current technology environment. The role involves sharing leading best practices and providing guidance to maintain the overall integrity of the platform, including operating systems, hardware and software infrastructure that are critical to the company's user populace and application environments. They foster a culture of collaboration with cross-functional teams to ensure that all components of the platform are working efficiently, securely, and resiliently. The Lead Platform Engineer leverages a strategic mindset, offers advice, and drives the development and implementation of security policies to ensure high performance and reliability of the company's infrastructure, including both cloud and on-premises systems, in support of DevOps efforts and general business and IT objectives. Daily Responsibilities: Set strategic priorities for platform hardware, software, and network requirements in alignment with overall organizational goalsShare best practices for monitoring system performance, and lead teams responsible for system performance and troubleshooting of issuesLeverage advanced techniques and methods for optimizing DevOps tools (such as API Gateway, Teraform) to support processes as per organizational needsFoster a culture of collaboration among cross-functional teams to ensure new features and services are brought into productionLeverage a strategic mindset to oversee the execution and maintenance of automated and orchestrated fulfillment mechanisms to optimize delivery of key platform servicesLead cross-team collaboration and workstreams to ensure smooth operationsIdentify potential security threats, mitigate them proactively and set standards for a strict security complianceDrive and oversee the development and implementation of security and resiliency policies, standards, and procedures in line with organizational strategic priorities and latest industry regulationsEvaluate high-level design documentation, share feedback, and provide recommendationsInterpret results and present findings from technical research regarding user requests for new/modified systems or severe problem resolutionFoster strong relationships with vendors to create mutually beneficial opportunities for the organization