Site Reliability Engineer

Roc Search — Germany · Posted ~2 hours ago

Senior Contract

Skills

Kubernetes Linux CI/CD Jenkins ArgoCD Git Infrastructure as Code Prometheus Grafana Elasticsearch Terraform

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

An experienced SRE is required to operate production platforms, manage Kubernetes infrastructure, build automation pipelines, and improve monitoring and reliability practices.

Highlights

Long-term engineering engagement focused on production reliability, automation, observability, and secure enterprise platforms.

Description

Site Reliability Engineer (SRE) – 24x7 Operational Support Germany Long-term contract – 2–3 years Ü2 Security Clearance Required / Must Be Eligible Site Reliability / DevOps / Platform Engineering We are supporting a major organisation in Germany that is looking to engage experienced Site Reliability Engineers for a long-term 2–3 year contract programme. This is a hands-on engineering position supporting business-critical platforms, with a strong focus on Kubernetes, Linux, observability, secure logging, automation and production reliability. Key Responsibilities Operate and support Kubernetes-based production environments, including Helm configurations and container orchestration.Maintain and develop CI/CD and automation using Jenkins, ArgoCD, Git and Infrastructure-as-Code.Manage observability platforms including Prometheus, Thanos and Grafana, covering alerting, PromQL and monitoring.Configure and optimise Elasticsearch / OpenSearch, Logstash and Kibana environments.Develop automation and operational tooling using Python, Bash and/or Go.Troubleshoot complex Linux, platform, networking, data and performance issues.Participate in 24x7 on-call support, including incident response and Major Incident Management.Execute scheduled maintenance and contribute to continuous improvements across platform stability and operational procedures.Maintain secure operational processes, documentation and access controls. Technical Skills Strong experience across several of the following: Linux | Kubernetes | Prometheus | Elasticsearch | OpenSearch | Grafana | Thanos | Logstash | Kibana | Terraform | Helm | ArgoCD | Jenkins | Git | Python | Bash | Go | CI/CD | IaC | Networking | REST APIs Security Clearance / Eligibility – Please Read Before Applying Due to the secure nature of this programme, there are strict eligibility requirements. Candidates must: Hold citizenship of a country that is a member of both the EU and NATO.Not hold dual citizenship.Be eligible to undergo and obtain German Ü2 (Erweiterte Sicherheitsüberprüfung) security clearance.Currently reside in Germany and be able to provide a registered German residential address.Be able to work under a German employment contract.Be able to provide a complete and verifiable residential and employment history as required by the clearance process.Be prepared to undergo the required security/background vetting.Be willing to participate in shift-based 24x7 operational/on-call support, including weekends and public holidays. Candidates who do not meet the citizenship and German residency requirements unfortunately cannot be considered for this particular programme. Desirable Certifications Elastic Certified EngineerLPIC-2Certified Kubernetes Administrator (CKA)Relevant AWS / Cloud / DevOps certifications The Opportunity This is an excellent opportunity for an experienced SRE, DevOps Engineer or Platform Engineer looking for the stability of a 2–3 year engagement while working on large-scale, secure and business-critical infrastructure. Apply with your CV or get in touch directly to discuss the programme in more detail.