Senior Cloud Platform Engineer

Plan B Ai — United Arab Emirates · Posted ~3 hours ago

Senior Full-time

Skills

Azure cloud infrastructure infrastructure as code security cloud engineering AWS GCP

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A technology firm is hiring a senior cloud engineer to design, deploy, and operate secure cloud platforms while supporting complex technical projects.

Highlights

Senior individual contributor role focused on secure cloud platforms, architecture, and end-to-end infrastructure delivery.

Description

Role SummaryWe need a senior cloud engineer who can design, build, and operate secure, scalable cloud platforms for AI and cloud-native products. This is a senior individual contributor role: you’ll take on complex technical work and see it through from design to production support with limited oversight. You’ll work across internal and client projects, sometimes within established security, access, and technology constraints. Azure is our main cloud platform, but we’re also open to candidates with strong AWS or GCP experience and solid cloud engineering fundamentals. Key Responsibilities• Take infrastructure projects from design through deployment and ongoing support, including testing and documentation. • Work with internal teams and clients to understand requirements, talk through technical options, and raise risks or dependencies when needed. • Deliver solutions within established client security, access, and approval processes when required. • Write and maintain infrastructure as code, build reusable modules, and review IaC changes. • Build and improve CI/CD pipelines. • Deploy and operate containerized applications and orchestration platforms. • Set up and maintain metrics, logs, traces, dashboards, and alerts for production systems. • Manage cloud networking, identity and access, secrets, and security controls. • Investigate and resolve production issues across infrastructure, networking, containers, identity, and applications. • Support GitOps and self-service tooling for engineering teams. • Build infrastructure for AI workloads, including model serving, scalable APIs, and vector stores. • Support disaster recovery, backups, resilience, and high availability. • Participate in incident response, root-cause analysis, and follow-up improvements. Skills & Experience• Strong hands-on experience with a major cloud provider. Azure is preferred, but AWS or GCP experience is also welcome. • Good understanding of cloud networking, including virtual networks, DNS, routing, load balancing, firewalls, and private connectivity. • Strong infrastructure-as-code experience. Terraform is preferred; experience with Bicep, Pulumi, or OpenTofu is useful. • Experience running containers and orchestration platforms. • Experience with CI/CD tools such as GitHub Actions, Azure DevOps, or GitLab CI. • Experience with observability tools such as Prometheus, Grafana, Datadog, OpenTelemetry, or cloud-native monitoring services. • Ability to automate operational work using Python, Bash, or PowerShell. • Good understanding of IAM/RBAC, secrets management, and secure configuration. • Working knowledge of cloud cost optimization, high availability, and disaster recovery. • Clear written and verbal communication skills, including the ability to explain technical issues to different audiences. • Comfort working within client security, access, or compliance requirements when needed. • Experience with AI, data, or compute-heavy workloads is an advantage. • Experience with platform engineering, internal developer platforms, GitOps, policy as code, SLOs/SLIs, multiple cloud providers, messaging systems, databases, or distributed systems would also be useful. • Familiarity with LLM-based applications, prompt engineering, RAG, and related AI concepts is a plus. What We’re Looking ForYou take ownership and can plan and deliver complex work without close supervision. You produce reliable, maintainable solutions and make practical trade-offs between security, reliability, cost, and delivery time. You troubleshoot methodically, raise risks early, and communicate clearly. You can adapt your approach when working within different tools, processes, and access restrictions. Success in This Role• You deliver reliable work with limited oversight. • You take complex projects from initial requirements through production support. • Colleagues and clients understand your decisions, risks, and progress. • The platform becomes more reliable, observable, secure, and easier to operate. • People trust your technical judgement and delivery.