Senior Kubernetes & HPC Engineer

Aptino — Canada · Posted ~13 hours ago

Senior

Skills

Kubernetes Cloud infrastructure HPC Docker Helm Slurm Volcano Linux administration Python Bash NVIDIA GPU infrastructure AI/ML workloads CI/CD Monitoring Infrastructure automation AWS EKS Azure AKS GCP GKE Linux NVIDIA GPUs

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

An experienced Kubernetes and HPC engineer is sought to design, deploy, and operate scalable compute platforms supporting AI/ML and data-intensive workloads. The role involves Kubernetes, Docker, Helm, workload schedulers, Linux, scripting, GPU infrastructure, CI/CD, monitoring, automation, and major cloud platforms.

Highlights

A technically advanced role building scalable compute platforms for AI/ML, simulation, and data-intensive workloads. The position combines Kubernetes, HPC, cloud infrastructure, GPU computing, automation, and high-performance distributed systems.

Description

We are looking for an experienced Kubernetes (K8S) / HPC Engineer to design, deploy, and manage scalable compute platforms supporting AI/ML, simulation, and data-intensive workloads. 🔹 Key Skills: • 5+ years of Kubernetes / Cloud Infrastructure experience • 2+ years of HPC / large-scale compute experience • Strong hands-on experience with Kubernetes, Docker & Helm • Experience with Slurm / Volcano workload schedulers • Strong Linux administration and Python/Bash scripting • Experience with NVIDIA GPU infrastructure & AI/ML workloads • CI/CD, monitoring & infrastructure automation • Experience with AWS EKS, Azure AKS, or GCP GKE • AWS Trainium / Inferentia / Neuron ⭐ Nice to Have: • RDMA / InfiniBand • Parallel storage systems • Large-scale AI/ML or scientific computing environments • Engineering simulation/semiconductor workloads