Summary
Help operate and improve large-scale cloud infrastructure by building automation, deployment pipelines and resilient systems while supporting secure and scalable digital services.
Highlights
Work on cloud-native platforms, improve reliability through automation, collaborate with engineering teams, and drive operational excellence.
Description
Site Reliability Engineer (AWS | DevOps | Cloud Platform)
Build, automate and scale critical digital services.
We are seeking a skilled Site Reliability Engineer to join a high-performing technology team responsible for the reliability, security and continuous improvement of critical cloud platforms.
This role combines engineering, automation and operational excellence, with a strong focus on modern AWS environments, CI/CD pipelines and platform reliability.
As an SRE, you will play a key role in designing, implementing and operating application delivery pipelines while helping to simplify, secure and modernise cloud-based services.
You'll work across AWS infrastructure, Kubernetes platforms and automation tooling to drive efficiency, scalability and resilience.
What You'll Do
Design, build and support CI/CD pipelines and deployment processes.Implement automation that improves operational efficiency, reliability and security.Support and optimise AWS-hosted platforms, websites, backend APIs and cloud services.Monitor system health, investigate incidents and perform root cause analysis.Drive improvements in platform performance, availability and scalability.Collaborate with developers, security specialists and stakeholders to deliver robust solutions.Contribute to infrastructure-as-code and cloud engineering initiatives.Document processes, procedures and operational standards.
Technology Environment
You'll Work With a Modern Cloud-native Stack Including
AWS (EKS, ECS, CloudFront, WAF, ACM, ALB, RDS, CloudWatch, S3, EC2)GitHub ActionsAWS CodeBuild & CodePipelineTerraformKubernetesCDK & TypeScriptBash scriptingGit/GitHubCloudWatch and emerging observability platforms including Grafana and Prometheus.
What You'll Bring
Essential Skills & Experience
Experience developing automation through scripting and programming.Strong AWS platform knowledge and cloud operations experience.Proven experience implementing and supporting CI/CD pipelines.Experience with infrastructure as code, particularly Terraform.Knowledge of Git and modern source control practices.Experience with monitoring, alerting and operational support.Strong troubleshooting and root cause analysis capabilities.Understanding of security across code, infrastructure and delivery processes.Excellent communication and teamwork skills.A commitment to continuous learning and professional development.
Desirable
Python, Go, PHP or TypeScript development experience.Kubernetes and container platform experience.Jira/Atlassian tools.MongoDB.Automated testing tools such as Playwright.Experience with Bitrise or similar DevOps tooling.
How to apply:
If you are interested and possess the right experience, please apply now via the link to be considered.
Contact: Laura GILLES – (08) 9423 1416 – (Job reference: 271548)
Peoplebank and Leaders IT are committed to creating a diverse and inclusive workplace where everyone belongs.
We welcome applications from people of all backgrounds, identities, and experiences.
If you need adjustments to the recruitment process due to your circumstances, please let us know—we’re here to support you.