Senior Site Reliability Engineer

Eroad — New Zealand · Posted ~12 hours ago

Senior Full-time

Skills

Site Reliability Engineering Distributed systems Production operations Root-cause analysis Automation Monitoring Observability Infrastructure Cost optimization On-call operations Distributed Systems SaaS Infrastructure Automation

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Join a senior SRE team responsible for keeping large-scale, customer-facing distributed platforms reliable. You will design solutions, support production environments through an on-call rotation, perform root-cause analysis, automate operational work, improve monitoring and observability, optimize infrastructure costs, and mentor teammates.

Highlights

Hands-on senior SRE role with significant influence over platform reliability and architecture. Includes work on large-scale distributed systems, automation, observability, infrastructure cost optimization, incident analysis, and mentoring within an autonomous team.

Description

ABOUT THE COMPANY EROAD is a global fleet management technology company and a genuine Kiwi tech success story, listed on both the NZX and ASX and growing across New Zealand, Australia, the Philippines and the USA. We're hiring a Senior Site Reliability Engineer to help maintain and evolve the SaaS platforms that our fleet operator customers rely on every day. WHAT YOU'LL DO This is a hands-on senior role with real influence over how our platforms run and improve. Contribute to solution design for large-scale, customer-facing distributed systemsOperate and support production environments as part of an on-call rosterDrive root-cause analysis when things go wrong and turn findings into lasting fixesBuild automation that reduces manual toil and improves reliabilityImprove monitoring and observability across the platformIdentify and deliver cost optimisation opportunities across the infrastructureMentor teammates and help lift the team's overall SRE practiceWork within an autonomous, self-managed Agile platform team aligned to one of EROAD's SaaS ecosystems, reporting to the Domain Chapter Lead WHAT YOU NEED To thrive in this role you'll bring solid, hands-on production experience across the following. Deep hands-on experience operating Kubernetes (AKS or EKS) in production at scaleHands-on experience with a public cloud platform, Azure or AWSAlignment to one cloud ecosystem, Azure/Windows or AWS/Linux, is fine, you don't need bothExperience with Terraform or similar Infrastructure-as-Code toolingExperience with CI/CD tooling such as GitHub Actions, Azure DevOps, Concourse or similarPractical scripting ability in Ruby, Bash, Python or similarExperience operating and managing complex, customer-facing, multi-tier distributed production systemsExposure to monitoring, alerting or visualisation tools such as Grafana, Sumo Logic or Datadog WHY JOIN This is a chance to take real ownership of reliability and performance for platforms operating at national and global scale. Senior, hands-on individual contributor role with genuine scope to shape solution design and drive improvement, not just keep the lights onMentoring responsibilities that build your leadership profile alongside your technical depthBe part of a high-growth, innovative technology company making a real difference for customers across multiple countriesWork alongside a talented and collaborative team in an autonomous, self-managed platform teamContinuous learning support, including EAP offerings and AI tooling to help you growCompetitive salary and benefits packageA multicultural organisation that genuinely values diversity WHAT YOU NEED To thrive in this role you'll bring solid, hands-on production experience across the following. Deep hands-on experience operating Kubernetes (AKS or EKS) in production at scaleHands-on experience with a public cloud platform, Azure or AWSAlignment to one cloud ecosystem, Azure/Windows or AWS/Linux, is fine, you don't need bothExperience with Terraform or similar Infrastructure-as-Code toolingExperience with CI/CD tooling such as GitHub Actions, Azure DevOps, Concourse or similarPractical scripting ability in Ruby, Bash, Python or similarExperience operating and managing complex, customer-facing, multi-tier distributed production systemsExposure to monitoring, alerting or visualisation tools such as Grafana, Sumo Logic or Datadog WHY JOIN This is a chance to take real ownership of reliability and performance for platforms operating at national and global scale. Senior, hands-on individual contributor role with genuine scope to shape solution design and drive improvementMentoring responsibilities that build your leadership profile alongside your technical depthBe part of a high-growth, innovative technology company making a real difference for customers across multiple countriesWork alongside a talented and collaborative team in an autonomous, self-managed platform teamContinuous learning support, including EAP offerings and AI tooling to help you growCompetitive salary and benefits packageA multicultural organisation that genuinely values diversity