Platform Engineer

Xestro — Australia · Posted ~2 hours ago

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Description

Xestro is growing rapidly, and we’re looking for an experienced Platform Engineer to help operate and improve our production infrastructure based in Maroochydore. About Xestro Xestro is an Australian healthcare software company that helps specialist doctors manage their patients and run their practices. Our cloud-based platform is used daily in hundreds of busy clinics across Australia by thousands of healthcare providers. We’re based in Maroochydore on the Sunshine Coast, close to the Sunshine Plaza, and have recently moved into a multi-storey office to support our next phase of growth. Our platform runs on AWS, with an established Linux, Apache, PHP and MySQL environment alongside a growing Kubernetes-based platform. We’re focused on improving reliability, security and automation while modernising and scaling our infrastructure. The Role You’ll report to the Head of Engineering and take day-to-day technical direction from our Lead SRE. You’ll join our Platform team, working hands-on with Kubernetes, AWS EKS, Terraform and deployment pipelines to support a production system used every day in real healthcare settings. The work includes operating production clusters, troubleshooting infrastructure and deployment issues, managing upgrades, monitoring performance and capacity, and improving automation. You’ll work with GitLab, ArgoCD, Temporal and Datadog, including supporting Temporal workers and the infrastructure behind our background processing workflows. You’ll also contribute to the operation of our wider AWS environment as it progressively moves under the Platform team. This role suits someone who is used to being responsible for live production systems. You’re comfortable responding to incidents, making carefully planned changes and following problems through from service recovery to lasting resolution. Participation in the on-call roster is part of the role. About You You may come from platform engineering, site reliability engineering, systems administration or cloud infrastructure. What matters is substantial hands-on experience operating production environments and the judgement that comes with it. What Matters Most Is How You Approach The Work You take ownership of the reliability and health of the systems you operateYou think changes through, including their impact and how to recover if something goes wrongYou troubleshoot methodically and remain effective under pressureYou communicate clearly, document useful knowledge and escalate when neededYou work within agreed technical direction, review and change processesYou contribute practical improvements and automate recurring manual workYou’re committed to ongoing learning and secure operational practices You’ll Be a Strong Fit If You Have 5+ years of experience in infrastructure, platform engineering, SRE or a related operational roleStrong hands-on experience running and troubleshooting Kubernetes in production, ideally AWS EKSStrong Linux administration and troubleshooting skillsPractical AWS experience, including VPCs, IAM, EC2, DNS and networkingExperience managing infrastructure through TerraformExperience operating CI/CD pipelines and supporting production deploymentsExperience investigating production incidents using metrics, logs and monitoring toolsScripting skills using Bash, Python or similarExperience planning and carrying out production upgrades, maintenance and rollback procedures Highly Desirable Experience GitLab CI/CD and ArgoCDDatadog monitoring, logging, dashboards and alertingTemporal or similar workflow orchestration platforms, including worker deployments, monitoring and troubleshootingEKS Auto Mode, Karpenter and Kubernetes scalingKubernetes networking, ingress, storage and workload configurationHigh-availability systems, capacity management and performance troubleshootingAWS security practices and infrastructure access controlsSupporting containerised applications alongside established Linux and EC2 infrastructure Why Join Xestro? Job security in a long-established, growing healthcare SaaS companyMeaningful work that directly impacts clinics and patient careSupportive, forward-thinking engineering culture with experienced peersHands-on ownership of production infrastructure and opportunities to improve how it operatesOngoing learning and a strong focus on secure engineering and operationsFlexible work arrangements within an onsite-first teamSunshine Coast lifestyle with a healthy work-life balance