Summary
✨ AI‑Generated
A platform engineering role focused on improving reliability, scalability, and automation for critical distributed systems used at large scale.
Highlights
Build reliable large-scale platforms through automation, cloud engineering, and distributed systems optimization.
Description
We are looking for a Site Reliability Engineer to join a platform team responsible for operating and evolving a large-scale NoSQL database environment that supports business-critical services used by millions of users worldwide.
This role sits at the intersection of Software Engineering, Site Reliability Engineering, and Platform Engineering.
You will work on reliability, scalability, automation, and modernization of distributed database systems, while helping reduce operational toil through software development and infrastructure automation.
The team owns the database platform end-to-end and is responsible for ensuring high availability, performance, capacity optimization, and operational excellence across both cloud and hybrid environments.
Key Responsibilities
Operate, maintain, and optimize large-scale NoSQL database environments.Improve platform reliability through SLI/SLO-driven engineering practices.Develop automation and tooling to reduce operational workload and manual processes.Design and implement scalable solutions for distributed database platforms.Support database migrations, upgrades, and modernization initiatives.Contribute to capacity planning, performance tuning, and cost optimization.Build, maintain, and improve infrastructure using Infrastructure as Code principles.Participate in incident response, troubleshooting, root cause analysis, and postmortems.Collaborate closely with software engineers, platform teams, and stakeholders across the organization.Continuously evaluate emerging technologies and propose improvements to the platform landscape.
Required Skills & Experience
Strong experience with Cassandra and DynamoExperience operating distributed database systems in production environments.Hands-on AWS experience.Strong Terraform knowledge.Programming experience in Python and JavaUnderstanding of SLI, SLO, SLA, and reliability engineering principles.Experience with automation, observability, monitoring, and alerting.Knowledge of Elasticsearch or DynamoDB is highly beneficial.Strong communication skills and a collaborative mindset.Details
Start: AsapDuration: 12 months (+)Location: Amsterdam, hybride
Let op: vacaturefraude
Helaas komt vacaturefraude steeds vaker voor.
We waarschuwen je voor mogelijke misleiding:
* Wij zullen nooit via WhatsApp of in een videogesprek vragen om jouw persoonlijke gegevens (zoals een kopie van je ID, bankgegevens of BSN).
* Twijfel je over de echtheid van een vacature of contactpersoon? Neem dan altijd rechtstreeks contact met ons op via de officiële contactgegevens op onze website.
Important: job fraud
Unfortunately, job fraud is becoming more common.
Beware of such scams:
* We will never ask for personal information (such as a copy of your ID, bank details, or social security number) via WhatsApp or during a video call.
* If you're unsure whether a vacancy or contact person is legitimate, please reach out to us directly using the official contact details on our website.