Summary
✨ AI‑Generated
A skilled Site Reliability Engineer is needed to improve system performance, automate operations, and support reliable production environments. The role offers technical ownership and collaboration with engineering teams.
Highlights
High-impact SRE role with ownership of reliability initiatives, automation projects, and opportunities to influence engineering practices.
Description
Site Reliability Engineer (SRE)
Location: London/Remote-first
Working pattern: 1 day per week in the London office
Salary: £90,000 base + cash benefits + bonus
A leading financial services organisation is looking for an experienced Site Reliability Engineer to join a growing function within the organisation.
This is a hands-on opportunity for someone who enjoys working at a technical level but also wants genuine ownership of projects and the opportunity to influence how SRE and DevOps are delivered across a large enterprise environment.
The team is continuing to develop its SRE capability, with a significant pipeline of projects focused on automation, reliability and improving the way engineering teams deliver and operate services.
The role
The successful candidate will work across Azure, Kubernetes, automation and production reliability, partnering closely with engineering and delivery teams.
Key areas of responsibility will include:
Improving the reliability, performance and resilience of cloud-based servicesReducing manual intervention across deployment and release processesAutomating repetitive operational tasks and reducing engineering toilHelping develop SRE practices around SLOs, error budgets and reliabilitySupporting and improving Kubernetes environments used by engineering teamsWorking with Azure DevOps and Infrastructure as Code to improve application and environment deliveryHelping support the transition from Bicep towards TerraformAutomating certificate management across a large technology estateImproving the management and integration of development artefacts with wider enterprise platformsSupporting production and non-production environments and taking ownership of incidents when requiredUsing monitoring and observability to identify potential reliability and performance issuesWorking with Engineering Managers, Technical Leads and Delivery Managers to drive technical initiatives forwardProviding technical guidance and mentoring to less experienced engineers as the function develops
Technology skills required
The role requires strong hands-on experience with Azure and Kubernetes, alongside a solid understanding of modern DevOps and SRE practices.
Relevant experience includes:
AzureKubernetes and containerised environmentsAzure DevOps and CI/CDInfrastructure as Code – Terraform, Bicep or ARMAzure identity, secrets and access managementMonitoring and observabilityAutomation and scriptingMicroservices and cloud-native environmentsProduction support and incident managementExperience with Grafana, Azure Monitor, Log Analytics, Application Insights, PowerShell or ServiceNow would also be beneficial.
Experience with both Bicep and Terraform isn't essential.
The team is currently moving towards Terraform, so strong experience with either technology will be considered.
What you need
Technical capability is important, but the team is particularly interested in someone who demonstrates ownership and initiative.
The successful candidate will be comfortable taking responsibility for an initiative from identifying the problem through to implementing a solution and managing the delivery themselves.
Strong stakeholder management is also essential.
The role involves working closely with technical and non-technical stakeholders across the organisation, including Engineering Managers, Technical Leads and Delivery Managers.They're also looking for someone who is naturally curious about technology and enjoys finding new ways to improve engineering practices.
An interest in areas such as AI, automation and emerging technology would fit particularly well with the team's ambitions.Plenty of experience working in another financial services firm
As the SRE function continues to grow, the successful candidate will have the opportunity to help establish best practice, represent the function across the wider technology community and mentor more junior engineers.
The team is at an important stage of building out its SRE capability, meaning this isn't simply a role focused on maintaining existing systems.
There is already a substantial pipeline of work across SRE and DevOps, including deployment automation, Kubernetes, certificate management, reliability and operational improvements.
The successful candidate will have the opportunity to shape the SRE function, reduce operational toil and establish new ways of working across a large technology organisation.
The position is remote-first, with an expectation of working from the London office approximately one day per week.
The salary is £90k plus cash benefits and bonus.