Summary
✨ AI‑Generated
A fast-growing organization is seeking a Senior DevOps Engineer to own and evolve platform and infrastructure across cloud, edge, on-premises, and ML environments. You will operate production Kubernetes clusters, implement GitOps and Infrastructure as Code, manage on-premises appliances, automate TLS and staged deployments, and build robust observability platforms. The role is ideal for an experienced infrastructure engineer who enjoys solving complex reliability, scalability, security, and compliance challenges.
Highlights
Opportunity to own and evolve cloud, edge, on-premises, and ML infrastructure while enabling secure, reliable, and compliant AI solutions. The role offers significant ownership across Kubernetes operations, GitOps, infrastructure automation, observability, and scalable platform engineering in a fast-growing environment.
Description
Senior DevOps Engineer - Berlin, Germany
Our client is combining AI & Healthcare to revolutionise patient care globally and continuing their impressive growth across the engineering team.
As a Senior DevOps Engineer, you will own and evolve their platform and infrastructure across cloud, edge, and ML environments, enabling their product and ML teams to deliver secure, reliable, compliant AI solutions.
Key Skills & Responsibilities
Production Kubernetes Operations: Own and operate Kubernetes clusters (AWS EKS, bare-metal K3s, on-prem K3s appliances), ensuring stability and scalability.GitOps & Infrastructure as Code: Use GitOps workflows (FluxCD) and IaC tools (CloudFormation, Ansible) to manage and automate infrastructure changes.Edge & On-Prem Appliance Management: Handle lifecycle of on-prem gateway appliances including VM image builds, TLS automation, and staged rollouts.Monitoring & Observability: Build and maintain Prometheus, Grafana, Loki, Tempo, and OpenTelemetry stacks to provide deep visibility across cloud and edge environments.Security, Compliance & Secrets Management: Drive security hardening, access control, audit logging, and encrypted secrets management (e.g.
SOPS/KMS).ML Infrastructure & Data Pipelines: Support GPU clusters, training job orchestration, and reliable ML data processing pipelines.CI/CD & On-Call: Design and maintain robust CI/CD pipelines and participate in on-call rotations, including incident detection, triage, and post-incident reviews.
Benefits with Our Client
Genuinely build and own infrastructure, directly seeing how your work improves lives Open, collaborative culture with team events in BerlinOpportunity for remote working and 'workaction' supported flexibilityStock options to become a co-creator of our client's success.30 vacation days plus your birthday off, Germany Transport Ticket, Urban Sports Club.Flexible working hours and hybrid working setup.
Interested?
If this aligns with your experience and ambition, apply now to learn more!