Observability Engineer - Data Platform (m/f/n)

Enovos Luxembourg Sa — Luxembourg · Posted ~1 day ago

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Description

We are looking for an Observability Engineer — Data Platform to help us build reliable, scalable and reusable observability capabilities for our data platform and critical data pipelines. This is a hands-on engineering role for someone with a strong Observability, Platform Engineering, SRE, DevOps or Cloud Engineering background. The role starts with data pipeline observability, but the ambition is broader: to contribute to observability standards that can progressively support product teams and wider IT capabilities. You will help Enovos move towards an OpenTelemetry-based observability approach, improve alerting reliability, and make observability easier to adopt for engineering and product teams. Your tasks • You design and implement observability standards for data pipelines, data products and platform services. • You support the progressive adoption of OpenTelemetry for logs, metrics and traces. • You will work with existing tools such as ELK / OpenSearch, Fluent Bit, Grafana and Prometheus. • You define reusable patterns for instrumentation, dashboards, alerting and runbooks. • You improve detection and diagnosis of data pipeline incidents, delays, failures and SLA breaches. • You correlate telemetry with meaningful context such as pipeline, run ID, dataset, data product, owner and environment. • You build actionable alerts and reduce noise from non-relevant or redundant alerts. • You support incident response, root cause analysis and post-incident improvements. • You automate observability onboarding using CI/CD, Infrastructure as Code and reusable templates where relevant. • You enable self-service observability for product and engineering teams. • You explore pragmatic AI-assisted use cases for anomaly detection, alert correlation or incident analysis. Your profile • You have at least 5 years of experience in observability, SRE, platform engineering, DevOps, cloud engineering or a similar role. • Strong hands-on experience with production systems. • Good understanding of modern observability: logs, metrics, traces, alerting, dashboards and incident response. • Practical experience with OpenTelemetry, or strong knowledge and motivation to lead its adoption. • Experience with tools such as Grafana, Prometheus, ELK / OpenSearch, Fluent Bit or similar. • Experience with cloud environments, preferably AWS. • Experience with Infrastructure as Code, preferably Terraform. • Ability to automate, standardise and document engineering practices. • Interest in data pipelines and data products, even if your core background is not Data Engineering. • A pragmatic mindset focused on production reliability, operational value and continuous improvement. • Good communication skills in English. French and/or German are considered an advantage.