Site Reliability Engineer (SRE)

Teksystems — Hong Kong Sar · Posted ~23 hours ago

Senior Visa History ✓

Skills

Site Reliability Engineering DevOps L3 support Incident management Kubernetes Infrastructure management Platform engineering Production troubleshooting

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A global financial services organization is seeking a proactive Site Reliability or DevOps Engineer for a hands-on L3 role. The position focuses on resolving complex production incidents, optimizing Kubernetes environments, building platform capabilities, and continuously improving reliability and operational excellence.

Highlights

Hands-on APAC-wide role with diverse project exposure, significant autonomy, L3 technical ownership, and opportunities to build and improve reliable infrastructure rather than only administer existing systems.

Description

APAC wide role at Global FS firmVariety of different project exposureDesign and hands on implementation responsabilities We are looking for proactive and hands-on Site Reliability Engineer (SRE) / DevOps Engineer to join a growing technology team. This is an L3-focused role where you'll work on ad-hoc technical requests, platform improvements, and operational excellence initiatives. This position offers significant autonomy and is best suited to someone who enjoys taking ownership, solving complex problems independently, and continuously improving platform reliability and performance. Unlike a purely operational support role, you will be expected to actively build, maintain, and enhance the underlying platform rather than simply administer existing systems. Key Responsibilities Provide L3 support for complex infrastructure and platform-related issues.Investigate and resolve production incidents and service disruptions.Build, maintain, and optimize Kubernetes environments.Develop Python scripts and automation solutions to improve operational efficiency.Implement platform enhancements and reliability improvements.Collaborate with engineering teams to improve scalability, performance, and system resilience.Support monitoring, alerting, and observability initiatives.Drive continuous improvement and automation across the technology stack.Act as an individual contributor with a high degree of ownership and accountability.