Description
About the Company
Join our technology team and help operate the AWS environments supporting software used by public transport operators across the ASEAN region.
About the Role
As our Senior AWS DevOps Engineer, you will take hands-on ownership of cloud infrastructure, automation, deployments, monitoring, security, and reliability across development, test, staging, and production environments.
You will work closely with Engineering, Support, Professional Services, customer IT teams, and third-party partners to ensure our platforms are secure, resilient, performant, and cost-effective.
This is a hands-on role for an experienced AWS/DevOps professional who enjoys solving complex infrastructure problems, automating repetitive work, and taking real ownership of production environments.
Responsibilities
AWS & Infrastructure
Design, build and maintain AWS infrastructure supporting production applications using Infrastructure as Code (Terraform, AWS CDK and/or CloudFormation).Deploy, upgrade and patch application environments across development, test, staging and production.Manage AWS services including EC2, ECS/EKS, S3, RDS, Lambda, VPC, Route 53, ELB, IAM, CloudWatch, Systems Manager and Backup.
DevOps & Automation
Build and maintain reliable CI/CD pipelines using AWS CodePipeline, CodeBuild, ShipHat, and related tooling.Automate application deployments, infrastructure provisioning and operational processes.Maintain Git-based source control and release workflows.Use scripting and automation with Bash, PowerShell and/or Python.
Reliability, Security & Operations
Implement monitoring, logging and alerting for application and infrastructure health, capacity and performance.Respond to production incidents within agreed SLAs, perform root-cause analysis and drive permanent corrective actions.Manage backup, restore and disaster recovery processes and participate in regular DR testing.Maintain security and compliance through patching, vulnerability remediation, least-privilege IAM, encryption, certificate management and security-group hygiene.Monitor AWS consumption and identify opportunities for rightsizing, reserved capacity, and architectural cost optimization.
Customer & Regional Operations
Work with Engineering, Support, Professional Services, customer IT teams and third-party vendors to coordinate deployments, integrations, change windows and issue resolution.Support production environments and customers across the ASEAN region and multiple time zones.Contribute to ITIL-based incident, change, release and configuration management.Maintain technical documentation, runbooks and as-built environment records.Participate in an on-call rotation and scheduled after-hours maintenance windows as required.Identify and deliver continuous service improvements in automation, reliability, and operational efficiency.
Qualifications
A tertiary qualification in Computer Science, Information Technology, Engineering or a related discipline is preferred, or equivalent practical experience.
Required Skills
5–8 years of relevant experience in DevOps, Cloud Engineering or Systems Engineering.At least 3 years of hands-on AWS production experience.Strong Infrastructure as Code experience, particularly Terraform and/or AWS CDK/CloudFormation.Proven experience building and maintaining CI/CD pipelines and Git-based workflows.Solid Linux administration and scripting/automation skills.Good understanding of AWS networking, IAM, and security fundamentals.Experience with Docker and preferably ECS and/or EKS.Experience supporting production environments, troubleshooting incidents, and conducting root-cause analysis.Strong communication and stakeholder-management skills, with the ability to work independently and with customer IT teams.
Preferred Skills
AWS certification such as Solutions Architect Associate, SysOps Administrator, or DevOps Engineer Professional.Experience with Python, PowerShell and/or Bash.Experience with SQL Server and/or PostgreSQL administration and tuning.Experience with observability tools such as CloudWatch, Grafana, Datadog, Nagios or similar.Experience with disaster recovery, backup/restore, and business continuity practices.Familiarity with ITIL-based managed services, incident, change and release management.Experience supporting enterprise or third-party COTS applications in AWS.Experience with Trapeze products is an advantage.