Site Reliability Engineer (Trading Infrastructure)

Matchtalent โ€” Hong Kong Sar ยท Posted ~5 hours ago

๐Ÿ”“ Log in to save this job, tailor your resume & track your apply process โ€” 7 days free, no card needed.

Log in to add to target list

Description

We are representing a fast-paced, result-driven trading firm in Hong Kong seeking a highly capable Site Reliability Engineer to join a specialized, high-performing core infrastructure team. In this highly visible "number three" role, you will take ownership of tracking, maintaining, and optimizing system performance across a demanding trading architecture. This is an exceptional opportunity for an ambitious engineer who wants to work closely with complex data pipelines, push the limits of system observability, and gain exposure to high-throughput trading environments. Key Responsibilities System Performance & Optimization: Take ownership of tracking system health, proactively identifying bottlenecks, and optimizing infrastructure for maximum uptime and efficiency in a latency-sensitive environment.Observability & Telemetry: Design and deploy robust telemetry frameworks (e.g., Prometheus, Grafana). Ensure the trading and engineering teams have granular, real-time visibility into application metrics and infrastructure logs.CI/CD & Data Workflows: Architect and streamline Continuous Integration and Continuous Deployment (CI/CD) pipelines. Manage the underlying data workflows that support rapid, reliable software releases.Automation & Systems Engineering: Write clean, scalable code to automate infrastructure provisioning and eliminate operational toil, ensuring systems run flawlessly at scale. Qualifications & Requirements Experience: 5 to 7 years of hands-on experience in Site Reliability Engineering, DevOps, or Systems Engineering.Environment: Proven ability to thrive in a fast-paced, high-pressure, and result-driven organization. Prior exposure to financial services, fintech, or trading is a plus, but ambitious candidates from top-tier tech firms will be highly considered.Coding Proficiency: Strong automation and programming skills in Python. (Bonus points for candidates with exposure to or an interest in systems languages like Go, C++, or Java).Technical Expertise:a) Deep expertise in managing modern CI/CD pipelines and data engineering workflows. b) Extensive experience with telemetry, observability, and monitoring ecosystems. c) Strong foundation in Linux/Unix OS internals and systems administration. d) Solid understanding of core networking concepts (TCP/IP, routing) and distributed systems. Work Style: A proactive problem-solver with a strong sense of ownership. Comfortable stepping into a lean team, managing up, and executing rapidly. Those outside of these requirements, may be considered for other SRE openings.