Summary
✨ AI‑Generated
A senior DevOps/SRE opportunity focused on building reliable cloud-native environments and modern software delivery practices. You will improve monitoring and incident management, work with KPIs, SLIs and SLOs, and collaborate across engineering and business teams. Strong German communication skills at B2 level or higher are required.
Highlights
Senior reliability-focused engineering role covering modern cloud platforms, delivery practices, monitoring, and incident management. Offers the opportunity to improve engineering-wide reliability practices and collaborate with engineers, architects, and business stakeholders in complex environments.
Description
Tech Stack
DevOpsSite Reliability Engineering (SRE)Cloud PlatformsCI/CDMonitoring & ObservabilityIncident ManagementKPI, SLI & SLOReliability Engineering
Requirements
Several years of experience in DevOps, Platform Engineering, or Site Reliability Engineering (SRE).Strong understanding of cloud-native environments and modern software delivery practices.Experience implementing and improving DevOps and reliability practices within engineering teams.Hands-on experience with incident management, monitoring, and service reliability.Good understanding of KPIs, SLIs, and SLOs and their role in maintaining reliable services.Experience working in complex or enterprise-scale environments.Ability to collaborate effectively with engineers, architects, and business stakeholders.Fluent German (B2 level or higher).Strong communication and problem-solving skills.
Nice To Have
Experience with large-scale DevOps transformation initiatives.Knowledge of Error Budget practices and reliability governance.Experience supporting platform or cloud modernization programs.Experience defining engineering standards and best practices.Cloud, DevOps, or SRE-related certifications.Experienced in using AI tools in day-to-day workflow
Project Description
We are looking for a Senior DevOps Engineer to join a strategic transformation initiative for a leading German organization.
In this role, you will help engineering teams improve reliability, operational excellence, and software delivery processes.
You will work closely with architects, engineers, and business stakeholders to implement DevOps and SRE best practices, establish reliability standards, and support the adoption of cloud-native ways of working.
This is an opportunity to contribute to a high-impact project where reliability, automation, and continuous improvement are at the heart of the engineering culture.
The position is fully remote within the EU, with occasional travel to Nuremberg, Germany (approximately 5-6 days per year).
Main Responsibilities
Support the implementation and evolution of DevOps and SRE practices across engineering teams.Define, monitor, and improve KPIs, SLIs, and SLOs.Contribute to the development of reliability standards and operational excellence processes.Improve incident management workflows and support root cause analysis activities.Collaborate with engineering teams to increase system reliability, scalability, and performance.Promote automation and modern software delivery practices.Support cloud-native adoption and continuous improvement initiatives.Work closely with technical and business stakeholders to ensure reliability objectives are met.Share knowledge and promote DevOps and SRE best practices across teams.
Why join?
Work on a strategic project for a leading German organization.Influence reliability and operational excellence across multiple engineering teams.Fully remote work within the EU.Long-term project with strong extension potential.Opportunity to collaborate with experienced DevOps, SRE, and cloud professionals in an international environment.