Site Reliability Engineer

Lancesoft Europe — Germany · Posted ~1 day ago

Mid Contract Remote Visa History ✓

Skills

Prometheus Grafana Loki OpenTelemetry Jaeger SIEM Observability

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A site reliability engineering role focused on building telemetry systems, monitoring distributed infrastructure, implementing tracing and logging solutions, and improving operational visibility.

Highlights

Remote reliability engineering contract focused on observability, telemetry, monitoring, and modern cloud operations.

Description

Observability & SRE Engineer Mode: 1 year Fixed term contract Remote but candidate based in Germany are preferred German Profecincy minimum of C1 is mandatory Primary Skills: Prometheus, Grafana, Loki, OpenTelemetry SDKs, Jaeger tracing, SIEM log export. Security Clearance & Vetting Level: Public Sector Clearance + NdK (Nachweis der Kundigkeit) Position Overview The Observability & SRE Engineer builds and operates central telemetry stacks to provide visibility across distributed cloud infrastructure. This role implements metric collection, log aggregation, distributed tracing, and standalone alerting tailored for Air-Gap operations. Technical Qualifications & Skills • Must Have: o Deep expertise in Prometheus (PromQL, ServiceMonitor, Federation, Remote Write) and Grafana. o Hands-on experience with Loki log aggregation and AlertManager routing. o Proficiency with OpenTelemetry (Collectors, SDKs, OTLP) and Jaeger distributed tracing. o Knowledge of SIEM integrations and security event logging. o Scripting skills in Go, Python, Shell, and YAML.