Senior Site Reliability Engineer

Salt — Netherlands · Posted ~13 hours ago

Senior Contract Hybrid Visa History ✓ €90-€110 per hour

Skills

Site Reliability Engineering distributed systems production operations incident response root cause analysis observability monitoring alerting logging tracing CI/CD cloud migration cloud

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A high-impact contract opportunity for a senior Site Reliability Engineer to join a high-performing platform team operating a mission-critical event-processing environment at massive scale. You will own production reliability, lead incident response and root-cause analysis, strengthen observability, improve CI/CD, and contribute to a major cloud migration while ensuring scalability, performance, and operational excellence.

Highlights

Work on a mission-critical, large-scale distributed platform handling billions of events daily, with strong ownership of reliability, operational excellence, observability, and cloud migration. The role offers a high hourly rate and a flexible hybrid arrangement.

Description

Salt is currently hiring a Site Reliability Engineer for a client of ours in Amsterdam. Senior Site Reliability Engineer (SRE) – Amsterdam Hourly rate: €90 - €110 (kvk/zzp needed) Duration: 3-6 Month Contract Start: ASAP Hybrid: 2 days a week onsite My client is looking for an experienced Site Reliability Engineer to join a high-performing platform team responsible for operating and evolving a mission-critical event processing platform that handles billions of events every day. This is an exciting opportunity to work on large-scale distributed systems, drive operational excellence, and support a major cloud migration initiative while ensuring the reliability, scalability, and performance of a business-critical platform. What You'll Be Doing Own end-to-end reliability of production services.Lead incident response, root cause analysis, post-mortems, and remediation activities.Improve observability through monitoring, alerting, logging, tracing, and dashboards.Build and enhance CI/CD pipelines and Infrastructure-as-Code solutions.Drive automation initiatives to reduce operational toil and improve platform efficiency.Support performance testing, capacity planning, and scalability initiatives.Contribute to the migration of a large-scale event streaming platform to a cloud-native architecture.Collaborate with software engineers and platform teams to improve reliability, resilience, and operational maturity.Participate in a shared on-call rotation and help maintain high service availability. What We're Looking For Proven experience as a Site Reliability Engineer, SRE, Platform Engineer, or DevOps Engineer in high-scale production environments.Strong hands-on experience with:API ManagementKubernetesAWSTerraform, Helm, GitOps, or similar Infrastructure-as-Code toolsCI/CD pipelines and deployment automationExperience with modern observability tooling such as Prometheus, Grafana, OpenTelemetry, ELK/EFK, Datadog, or similar.Strong background in incident management, production troubleshooting, and reliability engineering.Experience operating and scaling distributed systems handling high transaction or event volumes.Excellent communication skills and the ability to work effectively across engineering teams.