Site Reliability Engineer
Lhhworldwide — United Kingdom · Posted ~2 days ago
🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.
Log in to add to target listDescription
These roles require Security Clearance (SC) and sole British citizenship due to security constraints.
We are seeking experienced Principal Site Reliability Engineers (SRE) to join a high-performing engineering team delivering resilient, cloud-native platforms for UK-based customers.
These roles blend senior technical leadership with hands-on delivery, covering both project-based work and the ongoing reliability, scalability, and security of critical services.
You'll work closely with other senior engineers in small, collaborative teams, taking ownership of platform reliability, setting best practices, and mentoring others.
The role supports critical national infrastructure, requires participation in an on-call rota, and operates within a hybrid working model across UK offices, client sites, and home.
Your role
As an integral part of a Cloud Pod, you’ll have fantastic opportunities to develop both yourself and our collective capabilities as you progress both project and foundational requirements with other like-minded SREs and Cloud Engineers.
As part of the team, you’ll be empowered to:
Build and maintain platforms with a high degree of focus on technical reuse, standardisation, and blueprints; think “GitOps” not “ClickOps”.Fully embrace modern ways of working and feel invested in our customers’ outcomes as we group around activities in agile sprints, and within error budgets.Continue to strengthen and bolster your existing capabilities in site reliability through a mix of professional training, certifications, and experiences.
Your skills and experience
Hands‑on SRE experience with Kubernetes and OpenShift, including troubleshooting key Operators (ServiceMesh, ODF, ACS, ACM, AMQ)Ability to work within complex multi‑cloud or hybrid environments, with a solid foundation in distributed systemsPractical knowledge of observability tooling such as Prometheus, Grafana, Loki, and TempoProficiency in IaC tools (Kustomize, Helm) and scripting languages (Bash, Python), with experience managing GitOps pipelines using Tekton, ArgoCD, or FluxCDStrong growth mindset with willingness to learn from senior engineers; Kubernetes certifications (CKA/CKS) and secure‑environment experience are advantageous
We have 66,600 jobs that might be an even better fit for you
DontApply's real value goes far beyond a single job link or company name. Just upload your resume — in under a minute we'll analyze all 66,600 jobs and tell you exactly which ones you should apply to right now.
Upload My Resume