Senior Site Reliability Engineer

Oxylabs Io — Lithuania · Posted ~2 hours ago

Senior Full-time

Skills

Kubernetes Docker Swarm Ansible infrastructure management observability Infrastructure as Code CI/CD incident response on-call operations IaC

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A senior SRE will own and evolve production infrastructure across hundreds of servers and numerous services. Responsibilities include Kubernetes migration, observability, IaC, CI/CD reliability, incident leadership, post-mortems, and platform tooling, with shared ownership alongside software engineers.

Highlights

Ownership of production infrastructure, challenging large-scale reliability work, migration to Kubernetes, strong collaboration with developers, and opportunities to improve engineering tooling.

Description

We’re a team of 500+ professionals who develop cutting-edge proxy and web data scraping solutions for thousands of the world’s best known businesses, including Fortune 500 companies. What’s in store for you: You’ll be solving complex challenges and maintaining our own infrastructure. In this role, you’ll: Own and evolve Webshare's production infrastructure - lead the migration from Docker Swarm to Kubernetes (or hybrid K8s + Ansible)Maintain high availability across hundreds of servers and :50 servicesDrive observability in cooperation with development teamEstablish and enforce IaC practices, CI/CD pipeline reliability, and change management processesParticipate in the on-call rotation alongside backend developersRespond to and lead incident resolution, run post-mortems and drive systematic remediationContribute platform tooling that improves developer experience and reduces infrastructure toilKeep backend engineers informed and capable — no silos, shared infrastructure ownership Your skills & experience: Have built and operated highly available infrastructure at comparable scale — hundreds of servers, dozens of services, real production loadHands-on K8s in self-hosted / bare-metal environmentsConfident with Infrastructure as CodeHave owned CI/CD pipelines end-to-end (GitLab CI or equivalent)Have been on call in a production environmentProactive— surfaces problems before being asked, keeps the team informed without promptingScripting and development skills Nice to have: Led at least one major infrastructure migration - planned, executed, and stabilised itPython and/or Go familiarity - backend is Python, edge services are GoExposure to proxy, networking-heavy infrastructurePrevious experience in a small team where developers shared infrastructure responsibilityFamiliarity with edge clusters or split compute/edge architectures Salary & Benefits: Gross salary: 5800 eur/month - 7400 eur/gross + quarterly KPI-based performance bonus. Keep in mind that we are open to discussing a different salary based on your skills and experienceGrowth & Learning: 40+ internal learning options, external conferences, mentorship, and year-round knowledge-sharingHealth & Well-being: Private health insurance, gym allowance and a wellness appCelebration & Community: Team events, an overseas workation and plenty of ways to mark milestones together We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.