Site Reliability Engineer

Selby Jennings — United Kingdom · Posted ~2 hours ago

Senior Full-time Visa History ✓

Skills

Site reliability engineering Observability Deployment Incident response Low-latency systems Distributed systems Python Bash Platform infrastructure Developer tooling System scaling AI coding tools

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Take ownership of reliability for a business-critical, low-latency distributed platform. You will build observability and deployment capabilities, lead incident response, improve platform infrastructure and developer tooling, and use strong Python and Bash skills to raise reliability and engineering efficiency.

Highlights

High-impact SRE role working on business-critical, low-latency distributed systems. Strong opportunity to own observability, deployment, incident response, infrastructure, and developer productivity while operating closely with technical and front-office stakeholders.

Description

Our client, a world leading systematic multi strat hedgefund is looking for a Site Reliability Engineer to work within their ETF Trading systems. This role combines building and owning the observability of the ETF trading platform as well as coordinating across exchange connectivity quant platforms and risk team. This role sits within the front office and blends that deep technical impact with front office speed and efficiency. The optimal candidate will be someone who has previous experience in low latency trading settings setting the bar for reliability and scaling highly distributed systems within business critical systems. In addition to AI and agentic coding tools to accelerate trading work flows and strong python and bash scripting. Responsebilities: Building and owning observability, deployment, and incident response for the ETF trading platformDesigning and implementing development tooling and platform infrastructure to improve reliability and developer productivityManage systems deployments, upgrades, and migrationsCoordinate across exchange connectivity, quant platform and risk teams as new venues are onboarded.Automate exchange go-live operations, data pipelines, and desk workflows - leveraging AI toolingRequirements: Experience in production SRE or platform engineering in a latency sensitive financial environmentDeep experience in Linux systems cloud and infrastructure as codeExperience with monitoring stacks and CI/CD pipelinesProficiency with AI and agentic coding tools to accelerate workflowsStrong python and bash scripting for automation toolingCross team collaboration and communication is key