DevOps Engineer

Helloera4 — United Kingdom · Posted ~2 hours ago

Mid Full-time Remote

Skills

Kubernetes Linux cloud-native platform engineering DevOps automation systems reliability GPU-accelerated workloads bare-metal servers GPU cloud-native bare-metal

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Join a remote infrastructure team building and operating secure, scalable AI compute platforms. You’ll work hands-on with Kubernetes, Linux, cloud-native technologies, automation, GPU workloads, bare-metal infrastructure, and reliability engineering while supporting technical users and enterprise workloads.

Highlights

Permanent remote role focused on modern AI infrastructure, cloud-native engineering, Kubernetes, automation, GPU workloads, and large-scale systems reliability, with occasional office visits.

Description

Era4 develops, owns and operates AI infrastructure across the UK, powered by renewable energy. Converting legacy industrial and energy sites into modern data-centre facilities, Era4 is combining brownfield regeneration opportunities with cleaner, efficient, scalable compute capacity for healthcare, research, finance, enterprise, and public-sector organisations. This is a permanent full-time remote role with occasional visits to the office. Role Summary We are looking for a DevOps Engineer to design, operate and continuously improve our Kubernetes-based AI infrastructure. This hands-on role will focus on cloud-native platform engineering, GPU-accelerated workloads, bare-metal servers, automation and systems reliability, helping to deliver a secure, scalable and high-performing AI platform for ML engineers, data scientists and enterprise customers. You will help build and operate our AI infrastructure, supporting Linux systems, Kubernetes platforms and automation initiatives. Working alongside experienced engineers and technology partners, you'll contribute to the delivery of secure, reliable and scalable services. Responsibilities Deploy and maintain bare-metal servers, Linux systems and data-centre infrastructure.Configure, support and troubleshoot Kubernetes clusters and workloads.Manage NVIDIA GPU infrastructure, monitoring performance and resource utilisation.Automate infrastructure and operational processes using IaC and DevOps tools.Support security, access management and platform governance.Collaborate with engineering teams, vendors and customers to resolve technical issues.Produce technical documentation and support infrastructure deployments.Participate in operational support and on-call activities. Required Skills & Experience Experience working with Linux servers and infrastructure environments.Exposure to bare-metal hardware, data centres or enterprise platforms.Knowledge of Kubernetes and containerised environments.Understanding of NVIDIA GPUs, AI infrastructure or high-performance computing environments.Familiarity with Infrastructure-as-Code, automation and scripting tools.Experience with technologies such as Ansible, Git, YAML, Python or CI/CD pipelines.Strong troubleshooting skills and willingness to learn new technologies.Ability to work across infrastructure, platform and operational environments. Why Join Era4 You’ll be joining a mission-driven start-up building critical national infrastructure, where operational excellence directly enables growth. This role offers high visibility with leadership, real autonomy, and the chance to shape how a next-generation company operates at scale. Diversity & Inclusion Era4 is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. Note We appreciate this is a relatively new skill set and we are open to candidates who may not tick all the boxes but are willing to learn and develop their skillset.