Lead Site Reliability Engineer

Awconnect — United Arab Emirates · Posted ~2 hours ago

Lead Full-time

Skills

Site Reliability Engineering DevOps cloud infrastructure Kubernetes infrastructure as code CI/CD incident response monitoring observability disaster recovery IAM security operations IaC PAM

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Take hands-on technical ownership as a Lead SRE within a regulated digital technology environment. You will manage production cloud infrastructure, Kubernetes, IaC and CI/CD, lead incident response and root-cause analysis, maintain observability and disaster recovery, and oversee access controls, security remediation, and operational compliance.

Highlights

Hands-on technical leadership over production infrastructure, reliability, security controls, and incident management in a regulated environment. The role provides broad ownership across cloud operations, Kubernetes, disaster recovery, observability, access controls, and technology risk.

Description

AW Connect is working with a Dubai-based digital asset business preparing to launch a regulated platform. We are looking for a technically strong Lead SRE / Technology Operations professional to take ownership of production infrastructure, reliability, incident response and technology operations within a regulated environment. This is a hands-on role — not a position for someone who has moved entirely into management. What You'll Do Own production cloud infrastructure, Kubernetes, IaC and CI/CDLead incident response, RCA, monitoring and observabilityOwn backup/restore, business continuity and disaster recoveryManage IAM/PAM, access controls and privileged access workflowsCoordinate with outsourced CISO, SOC and security partnersManage vulnerability remediation and penetration-test findingsMaintain security and operational controls and audit evidenceSupport wallet/custody operations and technology controls where required What We're Looking For 6+ years of hands-on SRE, DevOps, DevSecOps, platform or infrastructure engineering2+ years in fintech, banking, brokerage, exchange or another genuinely regulated/audited environmentStrong production experience with AWS or equivalent cloud infrastructureHands-on Kubernetes and Infrastructure as Code experienceStrong IAM/PAM and CI/CD knowledgeProven experience managing production incidents and conducting RCAExperience with disaster recovery, RTO/RPO and tested backup/restore processesExperience working with an outsourced CISO, SOC or managed security providerExperience with vulnerability management and penetration-test remediationComfortable owning technical documentation and audit/control evidence Digital Assets Experience Experience with crypto, digital assets, custody/wallet operations, Fireblocks, BitGo, blockchain infrastructure or KYT is a strong advantage, but not essential. Candidates from strong regulated banking, fintech, brokerage or financial infrastructure backgrounds will also be considered. The Right Profile You should be someone who has personally owned production infrastructure and been the person called when things break — not someone who has spent the last few years purely managing a team. This is an opportunity to join a small, senior technology function at launch stage, with direct exposure to the Group CTO and genuine ownership of your area. Apply with a CV that clearly demonstrates what you personally owned, implemented and operated.