Summary
✨ AI‑Generated
A permanent hybrid C++ Site Reliability Engineer role focused on ensuring the reliability, performance, and operational excellence of critical low-latency production systems. You will work with modern C++, distributed systems, and SRE practices in an environment where speed, availability, and precision are essential.
Highlights
Permanent hybrid engineering role focused on reliability, performance, and operational excellence for demanding production systems. Combines modern C++, distributed systems, site reliability engineering, and low-latency technology in a technically intensive environment.
Description
C++ SRE Engineer
Location: London
Employment Type: Permanent
Working Pattern: Hybrid
Sector: Electronic Trading / Financial Technology
About the company
A global leader in technology-driven trading is looking for a talented C++ SRE Engineer to join its high-performing engineering organisation.
The firm designs and operates sophisticated, low-latency trading systems across global financial markets.
Technology is central to its success, with engineering teams responsible for building highly performant, resilient and scalable platforms that operate under extreme demands for speed, availability and precision.
This is an opportunity to work at the intersection of modern C++, site reliability engineering, distributed systems and electronic trading within a genuinely engineering-led environment.
The role
As a C++ SRE Engineer, you will help ensure the reliability, performance and operational excellence of critical trading infrastructure and production systems.
You will work closely with software engineers, quantitative developers, traders, infrastructure specialists and network engineers to improve system resilience, automate operational processes and resolve complex production issues.
The role will combine hands-on software development with production engineering, performance optimisation, observability and incident response.
Key responsibilities
Develop and maintain high-performance C++ tools, services and automation for critical trading platforms.Improve the reliability, scalability and operational resilience of low-latency production systems.Build monitoring, alerting and observability solutions across applications, infrastructure and networks.Investigate complex production issues involving performance, concurrency, memory, networking and distributed systems.Lead or contribute to incident response, root-cause analysis and post-incident improvement initiatives.Automate deployment, configuration, testing and operational workflows.Work with development teams to improve software design, instrumentation, release processes and production readiness.
Essential experience
Strong commercial experience developing software in modern C++.Experience working with Linux in a production engineering, SRE, systems engineering or software engineering environment.Strong understanding of systems programming, concurrency, memory management and performance optimisation.Experience supporting highly available, high-throughput or latency-sensitive production systems.Good knowledge of networking fundamentals, including TCP/IP, DNS, routing and network troubleshooting.Experience with automation and scripting, using languages such as Python, Bash or similar.Familiarity with observability principles, including metrics, logging, tracing and alerting.A structured approach to incident management, troubleshooting and root-cause analysis.Strong communication skills and the ability to collaborate with technically demanding stakeholders.
Desirable experience
Experience within electronic trading, financial technology, telecommunications, gaming, aerospace or another performance-critical environment.Knowledge of low-latency networking, kernel bypass, DPDK, InfiniBand, RDMA or user-space packet processing.Experience with Kubernetes, containers, CI/CD pipelines and infrastructure-as-code.Familiarity with configuration management and automation tools such as Ansible, Terraform or similar.
Apply now