Site Reliability Engineer

Anson Mccade — United Kingdom · Posted ~4 hours ago

Mid Full-time Hybrid Visa History ✓ £50000-£60000 per year

Skills

Site Reliability Engineering Software engineering Infrastructure Operations Automation Continuous improvement Application availability Monitoring Resilience Incident analysis

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A Site Reliability Engineer is needed for a hybrid role supporting critical applications. You will investigate the root causes of reliability issues, engineer durable solutions, improve monitoring and resilience, and work closely with development and product teams rather than simply handling operational tickets.

Highlights

Competitive salary, hybrid working, and an engineering-led reliability role focused on automation and continuous improvement. The position offers meaningful work on critical systems and close collaboration across software, infrastructure, and product teams.

Description

Site Reliability Engineer – Gloucester Location: Gloucester Working pattern: Hybrid Security: Candidates must be eligible to achieve UKIC DV Clearance £50,000 - 60,000 p/a DOE I’m working on a Site Reliability Engineer opportunity with a growing national-security technology team in Gloucester. The role sits across software engineering, infrastructure and operations, with a strong focus on automation and continuous improvement. Rather than simply responding to incidents and closing tickets, you’ll be expected to understand why issues are happening and put the right engineering solution in place. The role You’ll be responsible for keeping critical applications available, stable and performing as they should. That will involve working closely with development and product teams, improving the way services are designed and supported, and making sure systems are properly monitored and resilient. The day-to-day work will include: Supporting and improving services used by important mission applications.Investigating incidents and resolving problems across the application and infrastructure stack.Automating repetitive operational tasks wherever possible.Improving monitoring, logging, alerting and overall service visibility.Using performance and availability data to identify areas for improvement.Advising development teams on reliability, scalability and supportability.Working with cloud platforms, containers and microservices.Contributing to SRE and DevOps standards across the wider engineering community.Introducing practical improvements that reduce manual support and improve service quality. This is not a role where you’ll spend all your time dealing with operational tickets. The expectation is that support work is kept under control so that the team can focus on automation, engineering improvements and making the underlying services more reliable. What they’re looking for You’ll ideally have a background in SRE, DevOps, platform engineering, software engineering or a similar infrastructure-focused role. Experience with some of the following would be useful: Java and web technologies such as JavaScript and HTML.Linux and Windows, including Bash and PowerShell.AWS, Azure or OpenStack.ELK, Elasticsearch or similar monitoring and logging tools.Docker, containers and microservices.Chef, Puppet or comparable deployment and configuration tools.Elasticsearch, MongoDB or similar database technologies.Application troubleshooting and incident resolution.Agile or Scrum delivery environments and tools such as Jira.ITIL terminology and service-management processes.Automated testing tools such as Selenium.Working with or improving open-source software. You don’t need to tick every single box. A strong troubleshooting background, good software engineering principles and an ability to automate and improve systems will be more important than knowing every tool listed above. The opportunity This is a chance to work on technically challenging systems where availability, resilience and security genuinely matter. You’ll join an expanding engineering community in Gloucester and work alongside experienced software, infrastructure and DevOps professionals. If your background is in SRE, DevOps, platform engineering or software-focused infrastructure and you’re looking for something more meaningful than standard operational support, get in touch with Chris Prendergast at Anson McCade or hit Apply now.