Observability & SRE Engineer
W3Global โ Germany ยท Posted ~2 hours ago
๐ Log in to save this job, tailor your resume & track your apply process โ 7 days free, no card needed.
Log in to add to target listDescription
Position Overview
The Observability & SRE Engineer builds and operates central telemetry stacks to provide visibility across distributed cloud infrastructure.
This role implements metric collection, log aggregation, distributed tracing, and standalone alerting tailored for Air-Gap operations.
Key Responsibilities
Monitoring Stack Setup: Deploy and maintain Prometheus, Grafana, Loki, and Jaeger stacks as code.
Dashboards & Alerting: Develop PromQL/LogQL monitoring dashboards; configure AlertManager routing and inhibition for autonomous Air-Gap operations.
Distributed Tracing: Implement OpenTelemetry Collectors and SDKs for auto-instrumentation and OTLP trace propagation.
SIEM Integration: Configure secure log export and security event management (CEF/Syslog) for external SIEM platforms.
Must Have
Technical Qualifications & Skills
Deep expertise in Prometheus (PromQL, ServiceMonitor, Federation, Remote Write) and Grafana.Hands-on experience with Loki log aggregation and AlertManager routing.Proficiency with OpenTelemetry (Collectors, SDKs, OTLP) and Jaeger distributed tracing.Knowledge of SIEM integrations and security event logging.Scripting skills in Go, Python, Shell, and YAML.
We have 137,514 jobs that might be an even better fit for you
DontApply's real value goes far beyond a single job link or company name. Just upload your resume โ in under a minute we'll analyze all 137,514 jobs and tell you exactly which ones you should apply to right now.
Upload My Resume