Lead Site Reliability Engineer

Mobiskill — France · Posted ~1 day ago

Lead Full-time

Skills

Site Reliability Engineering Kubernetes Distributed databases Infrastructure automation Observability Performance engineering Cloud infrastructure Security Compliance FinOps Autoscaling

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Lead a small, high-performing SRE team while staying approximately 90% hands-on and defining the long-term infrastructure vision for systems operating at extraordinary scale. You will design and scale distributed databases, automate infrastructure and operational workflows, improve Kubernetes autoscaling and efficiency, and drive reliability, observability, security, compliance, performance, and FinOps initiatives.

Highlights

Lead an elite SRE team while remaining deeply hands-on, shaping long-term infrastructure strategy for systems operating at massive global scale. The role offers substantial ownership across reliability, observability, automation, security, performance, and cost optimization.

Description

Founded by the team behind one of the most successful location-sharing apps in the world, this startup is already handling massive global traffic (millions of active users across 20+ countries) and is preparing for its next phase of explosive growth. Job Context: The scale is already elevated (peaks at 100k+ requests/second), and the next milestone is critical. The challenge is to lead a small, high-performing SRE team while remaining 90% hands-on to define and execute the long-term infrastructure vision. Responsibilities: Define and execute the SRE strategy for the coming years (performance, reliability, observability, FinOps, security and compliance).Design, maintain, and scale distributed databases across multiple clusters (e.g., high-throughput monitoring for 1M+ series per second).Eradicate manual tasks by automating everything from the ground up (monitoring dashboards, database installations, cluster management).Optimize Kubernetes autoscaling to achieve near 90% CPU efficiency without compromising service quality.Act as a Software Engineer: build CI/CD pipelines, write code, and bring new tools to production to help the company scale faster.Partner closely with the backend engineering team to ensure smooth, reliable code deployments.Act as the primary SRE point of contact for top management. Stack: GCP, Kubernetes, Datadog, ScyllaDB, Redpanda, Rust, Go. Experience: You have a proven track record of surviving and thriving in an explosive scale-up/high-traffic environment. You are 80-90% hands-on but capable of driving high-level strategy. You have a "low ego", a pragmatic mindset, and ideally a strong background in software development. Conditions: Around €140k base salary, depending on the profile, plus equity (BSPCE).Located in central Paris (100% on-site in a beautiful Parisian-style office).Full relocation package available (visa sponsorship, 1st month Airbnb, relocation agency, French lessons).