Senior Cloud DevOps Engineer

Bitazza — Thailand · Posted ~3 hours ago

Senior Full-time

Skills

AWS Infrastructure as Code CI/CD Docker Cloud infrastructure management Terraform CloudFormation ECS Kubernetes

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A technology organization is looking for a senior cloud DevOps engineer to design and operate secure, scalable infrastructure. The role focuses on automation, cloud platforms, deployment pipelines, and production reliability.

Highlights

Opportunity to improve large-scale cloud infrastructure, automate engineering workflows, and work on reliability and scalability challenges.

Description

Senior Cloud DevOps Engineer As a Senior Cloud DevOps Engineer at Bitazza, you will help build, operate, and continuously improve our cloud platform and production infrastructure. You will work closely with software engineers and Tech Leads to automate infrastructure, improve developer experience, and ensure our systems are secure, scalable, observable, and reliable. This role is hands-on and focused heavily on AWS, Infrastructure as Code, CI/CD, containerized workloads, and production reliability. Responsibilities Design, build, maintain, and improve AWS infrastructure, including services such as ECS, EC2, VPC, IAM, ALB, Route 53, CloudFront, S3, RDS, CloudWatch, and Secrets Manager.Manage infrastructure through Terraform and AWS CloudFormation, ensuring environments are consistent, maintainable, and automated.Deploy and operate containerized applications using Docker and Amazon ECS, while contributing to our future adoption of Kubernetes.Develop and maintain CI/CD pipelines, primarily using Bitbucket Pipelines, to enable fast, secure, and reliable software delivery.Manage Cloudflare services including DNS, CDN, SSL/TLS, WAF, and traffic management.Build and maintain monitoring, logging, dashboards, and alerting using Datadog and AWS CloudWatch.Improve the reliability, performance, scalability, and security of production systems, including troubleshooting incidents, performing root cause analysis, and implementing long-term improvements.Build automation and self-service tools that reduce manual operational work and improve the developer experience.Collaborate with software engineers and Tech Leads on infrastructure architecture, deployment workflows, platform improvements, and operational best practices.Participate in the on-call rotation and contribute to continuous improvement of incident response and SRE practices. Requirements 3+ years of experience in DevOps, Cloud Engineering, Platform Engineering, Site Reliability Engineering, or a similar role.Strong hands-on experience with AWS and production cloud infrastructure.Strong experience with Terraform and familiarity with AWS CloudFormation.Experience building and maintaining CI/CD pipelines using Bitbucket Pipelines or similar platforms.Strong experience with Docker and containerized applications; Amazon ECS experience is highly preferred.Strong understanding of Linux and networking, including DNS, HTTP/HTTPS, TLS, load balancing, routing, and troubleshooting.Experience with monitoring and observability tools such as Datadog, CloudWatch, Prometheus, or Grafana.Ability to automate operational tasks using Bash, Python, or similar scripting languages.Strong troubleshooting, problem-solving, communication, and collaboration skills.Good understanding of cloud security concepts, including IAM, secrets management, encryption, and least-privilege access. Preferred Qualifications Hands-on experience with Cloudflare.Experience operating highly available and scalable production systems.Familiarity with Kubernetes or experience supporting Kubernetes environments.Experience with SRE practices, incident management, post-incident reviews, and reliability improvement.Familiarity with relational databases and basic SQL.