Summary
✨ AI‑Generated
Join a global engineering team building advanced cloud infrastructure for AI workloads. The role offers ownership, remote flexibility, competitive compensation, and the chance to contribute to foundational systems used by demanding technology teams.
Highlights
Remote full-time role with competitive compensation, equity participation, and the opportunity to build large-scale infrastructure alongside experienced engineers.
Description
At Verda, we're building a full-stack AI cloud, covering everything from data centers and hardware to our own cloud platform that the world's leading AI teams use to do serious AI work.
We strive to make a positive mark on the world through the infrastructure we build and give leading teams a service they can truly depend on.
Headquartered in Helsinki, we operate globally with offices in London and San Francisco.
Join Verda while it’s still being built - not once it’s finished.
Why Verda
Cash and equity compensation along with local benefits.40+ nationalities, with 6 different ones on the management team.A real chance to make an impact and work alongside world class engineers, researchers, and partners across the global AI ecosystem.
Practicalities
Work mode: RemoteLevel: Mid / SeniorEmployment type: Full time and permanentCompensation: USD 110,000 to USD 130,000 base salary, depending on experience and skills, plus equity.
The final offer will reflect your experience, capabilities, and location within the range.
Your Responsibilities
Build, operate, and improve self-hosted Kubernetes platforms across deployment, automation, scaling, upgrades, and production operationsMaintain platform reliability through on-call participation, incident response, and troubleshootingManage Kubernetes infrastructure including networking, observability, storage, and platform servicesDevelop infrastructure automation using Ansible, GitOps, and CI/CD workflowsCollaborate with engineering teams to improve developer experience, deployment processes, and operational efficiencyOperate and troubleshoot Linux systems, container platforms, and distributed infrastructureImplement Kubernetes networking and security best practices, including cluster hardening and access controlsLeverage AI-assisted engineering tools to improve automation, operations, and troubleshootingContribute to platform standards, documentation, and operational best practices
Your key competencies
3-7 years operating production Kubernetes environmentsStrong experience managing self-hosted Kubernetes clusters end-to-endHands-on experience with Kubernetes operations, upgrades, scaling, and troubleshootingExperience with on-call rotations and production incident managementStrong background in infrastructure automation (Ansible), GitOps, and CI/CDSolid Linux administration and debugging skillsExperience with containerized and cloud-native infrastructureUnderstanding of Kubernetes networking, CNI plugins, and platform securityScripting or programming experience (Python, Bash, Go, or similar)Comfortable using AI-powered tooling to improve engineering workflowsStrong collaboration and communication skills
Nice to have
Experience with Rancher and CiliumFamiliarity with Kubernetes security policies and cluster hardeningExperience with GPU workloads, AI/ML infrastructure, or NVIDIA container runtimesExperience with observability tooling such as Prometheus, Grafana, and LokiFamiliarity with Infrastructure as Code and Kubernetes ecosystem tools (ArgoCD, Helm, Kustomize, operators)Knowledge of distributed systems, storage, and high-availability environments
What's Next
We're building fast and this role needs the right person behind it.
There's no artificial deadline, but when we find who we're looking for, we move.
If this sounds like your next move, apply now.
Please submit your application through our Careers page.
We don't accept applications sent by email.