Summary
✨ AI‑Generated
A hands-on Site Reliability Engineering role focused on designing and maintaining resilient physical and virtual infrastructure. You will automate operational processes, monitor system performance, proactively resolve reliability issues, and collaborate across engineering teams to build secure, scalable platforms. The position is fully remote initially, with potential future onsite work following security clearance.
Highlights
Fully remote initially, strong compensation, hands-on technical ownership, opportunities to influence technical direction, and work on scalable, secure and reliable infrastructure supporting critical operations.
Description
The Details
Salary: £75,000 - £95,000 DOE
Location: The role is fully remote initially, but we may put you through DV clearance and, once cleared, ask you to work onsite in Cheltenham 2–3 days a week at some point in the future.
Security Clearance: Eligible for SC and/or DV Clearance
The Role
As a Site Reliability Engineer, you’ll play a key role in designing, building, and maintaining resilient systems that support critical operations.
You’ll work closely with cross-functional teams to ensure our platforms are scalable, secure, and reliable.
This is a hands-on role where you’ll have the opportunity to influence technical direction, improve system performance, and contribute to meaningful projects.
What You’ll Be Doing
Designing and maintaining reliable, scalable physical and virtual infrastructureMonitoring system performance and proactively resolving issuesAutomating processes using tools such as Ansible to improve efficiency and consistencyCollaborating with engineers and stakeholders across the businessSupporting continuous improvement of systems, tools, and practicesOperating across the full infrastructure stack, from bare metal systems through to virtualised deployments and the applications running within them.
What We’re Looking For
We recognise that no one ticks every box.
If your experience aligns with most of the below, we’d love to hear from you:
Experience in Site Reliability Engineering, DevOps, or similar rolesStrong knowledge of cloud platforms (e.g.
AWS, Azure, or GCP)Experience with infrastructure as code and configuration management tools (e.g.
Terraform, Ansible)Familiarity with CI/CD pipelines and automation practicesUnderstanding of monitoring, logging, and alerting systemsA collaborative mindset and a passion for problem-solving