Lead DevOps Engineer

Sg Edts โ€” Indonesia ยท Posted ~2 days ago

Lead

Skills

AWS Google Cloud Platform Kubernetes CI/CD Infrastructure as Code Ansible Cloud security IAM Secrets management DevSecOps Monitoring Troubleshooting Capacity planning Performance tuning Cloud infrastructure management GCP IaC

๐Ÿ”“ Log in to save this job, tailor your resume & track your apply process โ€” 7 days free, no card needed.

Log in to add to target list

Summary

Lead the design, implementation, and operation of cloud infrastructure across AWS and GCP. You will build and optimize production Kubernetes platforms, develop secure CI/CD pipelines, automate infrastructure with Ansible, and strengthen cloud security through IAM, secrets management, and DevSecOps practices. The role also covers monitoring, troubleshooting, capacity planning, performance optimization, and close collaboration with development, QA, security, and infrastructure teams.

Highlights

Lead cloud and DevOps initiatives across AWS and GCP, with responsibility for production Kubernetes platforms, scalable infrastructure, CI/CD automation, Infrastructure as Code, security, and DevSecOps. The role provides broad technical ownership and cross-functional collaboration across the software delivery lifecycle.

Description

Responsibilities What You Will Do: Design, implement, and manage cloud infrastructure on Amazon Web Services (AWS) and Google Cloud Platform (GCP) to support business and operational requirementsBuild, manage, and optimize Kubernetes platforms for production environments, ensuring scalability, high availability, and reliabilityDevelop, maintain, and optimize CI/CD pipelines to enable fast, secure, and reliable application deploymentsImplement Infrastructure as Code (IaC) using Ansible to automate infrastructure provisioning, configuration, and managementPerform monitoring, troubleshooting, capacity planning, and performance tuning for cloud infrastructure, applications, and platformsManage cloud security, including Identity and Access Management (IAM), secrets management, and the implementation of DevSecOps best practicesCollaborate closely with Development, QA, Security, and Infrastructure teams to support the software development and delivery lifecycleDesign and ensure the implementation of High Availability (HA), Disaster Recovery (DR), backup strategies, and business continuity solutions in accordance with organizational standardsOptimize cloud resource utilization and implement cost optimization strategies to improve operational efficiencyLead incident response, problem management, and conduct Root Cause Analysis (RCA) for critical incidents to drive continuous improvementCreate and maintain technical documentation, platform standards, and DevOps best practices to ensure consistency and knowledge sharingProvide technical guidance, mentoring, and code reviews to team members, fostering engineering excellence and continuous skill development Requirements Person We Are Looking For: Bachelor's degree in Computer Science, Information Systems, Information Technology, Software Engineering, or a related field Minimum 7 years of experience as a DevOps Engineer, Site Reliability Engineer (SRE), or Cloud EngineerMinimum 2โ€“3 years of experience in a Senior or Lead role Proven experience managing production environments and high-availability infrastructuresMandatory Skills Cloud Platforms: Amazon Web Services (AWS) and Google Cloud Platform (GCP). Kubernetes (EKS, GKE, and self-managed Kubernetes). Docker and container platforms. Infrastructure as Code (IaC) using Terraform and Ansible.CI/CD tools such as GitLab CI, GitHub Actions, and Jenkins. Linux administration (Ubuntu, RHEL, Rocky Linux). Networking fundamentals, including TCP/IP, DNS, VPN, Load Balancers, Reverse Proxies, and TLS/SSL.Monitoring and observability tools such as Prometheus, Grafana, ELK Stack, and OpenSearch.Scripting using Bash and Python.Git and GitOps practices, including Argo CD.3. Technical CompetenciesDesign and manage scalable, secure, and highly available cloud infrastructure.Build, operate, and optimize production-grade Kubernetes environments.Develop, maintain, and optimize CI/CD pipelines and implement and manage Infrastructure as Code (IaC).Troubleshoot issues across cloud platforms, Kubernetes, Linux, networking, and application deploymentsImplement DevSecOps practices, including Identity and Access Management (IAM), Secrets Management, and container securityPerform monitoring, capacity planning, performance tuning, backup management, disaster recovery, and cloud cost optimizationLead the implementation and continuous improvement of DevOps practices and cloud platformsConduct code reviews and architecture reviews to ensure engineering quality and best practicesMentor and provide technical guidance to engineering teams and lead incident response, problem management, and Root Cause Analysis (RCA) for critical incidentsCollaborate effectively with Development, QA, Security, and Infrastructure teams to support the software delivery lifecycleKnowledge of Service Mesh technologies such as Istio and CiliumExperience with Chaos Engineering practices and CrossplaneFamiliarity with AI-powered developer tools such as Claude Code, GitHub Copilot, Codex CLI, and Gemini CLI.6. Cloud certifications are considered a strong advantage (AWS Certified Solutions, Professional AWS Certified DevOps Engineer โ€“ Professional Google Professional Cloud Architect Google Professional Cloud DevOps Engineer)