Site Reliability Engineer

Uniquehire β€” Australia Β· Posted ~6 hours ago

Senior Full-time

Skills

SRE Observability Monitoring CI/CD DevOps Dynatrace Logging Alerting Incident Management Monitoring tools Logging frameworks Alerting systems

πŸ”“ Log in to save this job, tailor your resume & track your apply process β€” 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A technology-focused employer is seeking a highly skilled Site Reliability Engineer to lead observability and DevOps initiatives. The role involves mentoring a team of engineers, driving SRE best practices to ensure system reliability, scalability, and performance, implementing CI/CD pipelines, and building out monitoring, logging, and alerting frameworks. The successful candidate will collaborate across functions, contribute to architecture decisions, and proactively resolve technical challenges to deliver high-quality software solutions.

Highlights

Lead role with mentoring responsibilities, opportunity to shape observability and SRE practices, and involvement in architecture and technical strategy discussions.

Description

We are looking for a highly skilled Site Reliability Engineering (SRE), Observability, and DevOps practices. This role is ideal for someone passionate about building reliable, scalable systems while mentoring and guiding a team of engineers. Key Focus: SRE | Observability | Monitoring | Dynatrace Key Responsibilities * Lead and mentor a team of developers and engineers in designing, developing, and deploying * Drive Site Reliability Engineering (SRE) practices to ensure system reliability, scalability, and performance * Implement and manage CI/CD pipelines to accelerate the software development lifecycle * Build and enhance observability frameworks including monitoring, logging, and alerting * Collaborate with cross-functional teams to deliver high-quality software solutions * Ensure application reliability through proactive monitoring and incident management * Participate in architecture discussions and contribute to technical strategy * Identify and resolve technical challenges impacting delivery * Foster a culture of continuous improvement, automation, and innovation Mandatory Skills βœ” Hands-on experience with Site Reliability Engineering (SRE) practices βœ” Expertise in Observability, Monitoring, and Application Performance Management βœ” Experience with monitoring tools such as Dynatrace, Prometheus, Grafana, or ELK Stack βœ” Hands-on experience with CI/CD tools such as Jenkins, GitLab CI, or similar βœ” Experience with Docker, Kubernetes, and containerized environments βœ” Strong troubleshooting and problem-solving skills βœ” Strong proficiency in Java 11 and Object-Oriented Programming