Senior DevOps Engineer
Involve Asia โ Malaysia ยท Posted ~22 hours ago
๐ Log in to save this job, tailor your resume & track your apply process โ 7 days free, no card needed.
Log in to add to target listDescription
About Us
Involve Asia is a performance-based marketing technology company trusted by global brands for affiliate marketing solutions.
Our platform connects advertisers with publishers across web, mobile, and social channels, enabling them to automate campaigns, optimize performance, and grow revenue.
Founded in 2014 and headquartered in Malaysia, Involve Asia has expanded its operations across Southeast Asia, including Indonesia and Thailand.
The Role
We are looking for an experienced Senior DevOps Engineer to design, operate, and continuously improve our cloud infrastructure and application delivery platform.
You will take ownership of production infrastructure across AWS and Kubernetes, drive infrastructure automation and CI/CD improvements, and work closely with engineering teams to improve the reliability, scalability, security, observability, and efficiency of our systems.
This is a hands-on senior engineering role suited to someone who is comfortable solving complex production problems, making infrastructure architecture decisions, and improving the way engineering teams build and operate services.
Responsibilities
Design, build, operate, and improve highly available and scalable infrastructure on AWS and Kubernetes.
Own Kubernetes platform architecture, including cluster configuration, upgrades, networking, workload scheduling, scaling, security, and production reliability.
Design and maintain reusable Infrastructure as Code (IaC) using Terraform, and automate system configuration, provisioning, and operational tasks using Ansible.
Develop and maintain reusable Terraform modules and Ansible roles/playbooks to ensure infrastructure is consistent, repeatable, and maintainable across environments.
Build and improve CI/CD and GitOps workflows that enable safe, repeatable, and efficient application deployments.
Improve platform reliability through effective monitoring, logging, alerting, observability, capacity planning, and performance optimization.
Lead troubleshooting and resolution of complex production infrastructure and application deployment issues.
Participate in incident response, root-cause analysis, and post-incident improvements to prevent recurrence.
Implement and continuously improve cloud and Kubernetes security, including IAM, secrets management, network security, access controls, and infrastructure hardening.
Identify opportunities to improve cloud cost efficiency, infrastructure performance, and resource utilization.
Automate repetitive operational processes using scripting and engineering tools.
Establish and promote DevOps standards, infrastructure best practices, documentation, and operational processes.
Partner closely with Software Engineering, QA, Product, and other teams to improve developer experience and production readiness.
Provide technical guidance and mentorship to engineers and contribute to infrastructure architecture and technology decisions.
Qualifications
5+ years of experience in DevOps, Site Reliability Engineering (SRE), Platform Engineering, Systems Engineering, or a similar infrastructure-focused role, with strong hands-on experience designing, operating, and troubleshooting production cloud infrastructure.
Strong production experience with AWS, including services such as EC2, VPC, IAM, S3, load balancing, auto scaling, and related cloud services.
Strong hands-on experience designing, deploying, and operating Kubernetes environments.
Solid understanding of Kubernetes architecture, networking, storage, security, resource management, scaling, and troubleshooting.
Strong hands-on experience with Terraform for Infrastructure as Code, including reusable modules, state management, and infrastructure lifecycle management.
Hands-on experience with Ansible, including developing and maintaining playbooks, roles, inventories, and configuration-management automation.
Experience integrating Terraform and Ansible into CI/CD or GitOps workflows for automated infrastructure provisioning and configuration.
Strong understanding of infrastructure automation principles, including idempotency, version control, testing, secrets handling, and environment consistency.
Experience designing and maintaining CI/CD pipelines using platforms such as Jenkins, GitHub Actions, GitLab CI, Bitbucket Pipelines, or equivalent.
Strong understanding of Docker and containerized application environments.
Experience with observability platforms such as Prometheus, Grafana, CloudWatch, ELK/OpenSearch, Datadog, or equivalent.
Strong Linux, networking, and production troubleshooting skills.
Experience automating operational tasks using Bash, Python, Go, or similar languages.
Good understanding of cloud security, IAM, secrets management, infrastructure security, and DevSecOps practices.
Experience operating production systems with an emphasis on availability, scalability, security, and operational excellence.
Strong analytical, troubleshooting, communication, and cross-functional collaboration skills.
Bachelor's degree in Computer Science, Information Technology, Engineering, or equivalent practical experience.
Nice to Have
Hands-on experience implementing GitOps using tools such as Argo CD or Flux.
Experience with Kubernetes package management and deployment tooling such as Helm.
Experience managing large-scale or multi-environment infrastructure using Terraform and Ansible.
Experience with Ansible Vault, dynamic inventories, or AWX/Ansible Automation Platform.
Experience designing or operating multi-account or multi-environment AWS architectures.
Experience with Kubernetes and cloud-native security tooling and practices.
Experience defining or working with SLIs, SLOs, availability targets, and operational metrics.
Experience with cloud cost optimization and FinOps practices.
Experience supporting high-traffic or business-critical production platforms.
Experience mentoring engineers or leading infrastructure/platform initiatives.
We have 78,512 jobs that might be an even better fit for you
DontApply's real value goes far beyond a single job link or company name. Just upload your resume โ in under a minute we'll analyze all 78,512 jobs and tell you exactly which ones you should apply to right now.
Upload My Resume