Summary
Build and operate the infrastructure behind enterprise-scale AI workloads as an SRE. You will improve cloud reliability and scalability, establish SLOs and observability, lead incident response and root-cause analysis, optimize databases and deployments, automate operations, and partner with engineering teams to improve system reliability and developer experience.
Highlights
Build infrastructure for enterprise-scale AI workloads, with strong ownership of reliability, scalability, observability, automation, incident response, and developer experience. The role offers close collaboration with product engineering teams.
Description
Build the infrastructure behind enterprise AI.
At Lio, we're building the AI workforce for procurement.
As a Site Reliability Engineer, you'll ensure our platform remains fast, scalable, and reliable as we grow.
You'll work closely with our product engineering teams to improve infrastructure, automate operations, and build systems that support enterprise-scale AI workloads.
What You'll Do
Build and operate reliable cloud infrastructure for production workloadsImprove performance, scalability, and reliability of backend servicesDefine SLOs, monitoring, alerting, and observability across the platformDrive incident response, root cause analysis, and postmortemsOptimize databases, deployments, and CI/CD pipelinesAutomate infrastructure and operational processesPartner closely with engineering teams to improve developer experience and system reliability
What we're looking for
Experience operating production workloads on a major cloud platformGood Python skills and experience optimizing backend servicesStrong understanding of monitoring, observability, and incident managementKnowledge of distributed systems, asynchronous processing, and scalable architecturesExperience with databases at scale (MongoDB is a plus)Familiarity with CI/CD pipelines (GitHub Actions preferred)A passion for automation, reliability, and building systems that scale
Why Lio?
Build the infrastructure powering one of Europe's fastest-growing AI startups.
Work on high-scale, production-critical systems.
Own reliability, performance, and developer tooling.
Competitive compensation, meaningful equity, and exceptional teammates.
100% on-site in our Munich office, where we build together every day.