Site Reliability Engineer

Deltaclass Technology Solutions Limited — Germany · Posted ~2 hours ago

Skills

Kubernetes Helm CI/CD Jenkins ArgoCD Python Bash Go Infrastructure as Code Prometheus PromQL Thanos Grafana Elasticsearch Logstash Kibana Incident response Troubleshooting Access control

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A technically demanding Site Reliability Engineer role focused on building and operating reliable containerized platforms. You will manage Kubernetes infrastructure, CI/CD, automation, observability, centralized logging, security controls, and production incident response while improving scalability and operational resilience.

Highlights

Highly technical SRE role spanning cloud-native platform engineering, observability, automation, logging, security, and incident response. Offers broad ownership across modern infrastructure technologies and exposure to complex production environments.

Description

Key Responsibilities: Platform Engineering & DevOps: Manage Kubernetes and container orchestration, including Helm chart configurations and CI/CD pipelines (Jenkins, ArgoCD). Develop automation scripts (Python, Bash, Go) and deploy Infrastructure-as-Code (IaC) solutions.Observability, Monitoring & Visualisation: Maintain Prometheus solutions (scrape configurations, alert rules, PromQL queries), administer Thanos and Grafana.Elastic Stack Operations & Log Management: Configure and optimise Elasticsearch clusters, Logstash pipelines, and Kibana dashboards for secure, scalable log processing.Incident Response, Troubleshooting & Collaboration: Participate in 24x7 on-call rotations for rapid incident response, troubleshoot platform, data and performance issues, and engage in Major Incident Management (MIM).Secure Operations & Compliance: Ensure system operations meet security and data protection requirements, maintain secure documentation, and manage access control policies. Qualifications, Requirements, and Skills Strong grasp of Linux concepts, preferably in Kubernetes environments.Solid understanding of networking fundamentals and REST APIs.Proficiency in Python, Go, or Bash.Proficiency in Git-based configuration management workflows.Familiarity with CI/CD tools like Helm, Jenkins, or ArgoCD.Experience with Elasticsearch and/or OpenSearch.Fluent English communication skills.Willingness to work shift-based 24x7 on-call support, including weekends and holidays.Must possess Ü2 security clearance.Citizenship required: Member state of EU and NATO. No dual citizenship outside these countries.Must reside in Germany and hold a German labor contract.Preferred Certifications: Elastic Certified Engineer, LPIC Level 2, Kubernetes Administrator.