Site Reliability Engineer

Radley James — United Kingdom · Posted ~21 hours ago

Senior Hybrid

Skills

site reliability engineering Python Go observability monitoring automation infrastructure engineering distributed systems performance engineering Monitoring Observability Automation Distributed Systems Infrastructure

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Work on globally distributed, ultra-low-latency infrastructure where reliability and performance are critical. As a Site Reliability Engineer, you will build scalable monitoring platforms, develop automation with Python and Go, improve production resilience, investigate complex distributed-system performance issues, and collaborate with infrastructure, networking, and engineering teams.

Highlights

Highly compensated SRE opportunity with annual bonus and exceptional benefits. The role works on globally distributed, ultra-low-latency infrastructure and combines software engineering with systems expertise to improve reliability, observability, automation, resilience, and performance.

Description

Site Reliability Engineer Location: London (Hybrid) Salary: Up to £200,000 base + Annual Bonus + Exceptional Benefits We're partnering with one of the world's leading quantitative trading firms to hire a Site Reliability Engineer who enjoys solving complex infrastructure challenges at scale. This is an opportunity to work on the systems that power a globally distributed, ultra-low latency trading environment. You'll combine software engineering with systems expertise to improve the reliability, observability, automation and performance of critical infrastructure used by engineering and trading teams worldwide. What you'll be doing • Design and build highly scalable monitoring and observability platforms • Develop automation and infrastructure tooling using Python and Go • Improve the resilience and performance of critical production systems • Investigate complex performance issues across distributed software stacks • Work closely with infrastructure, networking and engineering teams to continually optimise large-scale environments • Build and enhance deployment, configuration management and operational tooling What we're looking for • 5+ years' experience as an SRE, Platform Engineer, Infrastructure Software Engineer or similar • Strong Python and/or Go development skills • Excellent Linux systems knowledge • Experience with Kafka, RabbitMQ or other streaming technologies • Kubernetes, containers and modern CI/CD tooling • Experience working on large-scale, highly available production infrastructure • Exposure to Rust or C/C++, ClickHouse/Cassandra/Bigtable, or networking technologies is advantageous What's on offer • Opportunity to work on one of the most advanced infrastructure environments in the industry • Highly technical engineering culture • Long-term career growth • Base salary: up to £200,000 • Annual bonus: typically 30–100%+ depending on performance • Total compensation: £200,000–£350,000+ for strong performers