Site Reliability Engineer

Kiloverse — Lithuania · Posted ~1 day ago

Mid Full-time

Skills

SRE cloud infrastructure observability reliability engineering cloud monitoring DevOps

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Join an engineering team responsible for keeping business-critical platforms reliable, scalable, and observable. You will improve infrastructure performance and operational excellence.

Highlights

Own reliability and performance of critical systems with exposure to high-impact digital platforms.

Description

About The Team Kiloverse is the operational force that helps ventures grow — building the systems, teams, and infrastructure that turn ideas into scalable businesses. Behind every fast-growing product, there’s a group of people who think like owners, act fast, and execute with precision. Here, you’ll find people who want to break big in their careers — working side by side with entrepreneurs and founders to build ventures that redefine the health, wellness, beauty, or travel industries. Our PayTech team runs the money layer behind all of it. We build and operate both centralized and decentralized business-critical systems: subscription-billing solution, payments gateway SaaS that routes payments across multiple payment providers, accounting solution that keeps every transaction traceable and others. In this domain, a minute of downtime means lost revenue. That is why we’re looking for a Site Reliability Engineer to own the reliability, performance and observability of these systems under high load. Core Components And Technologies Used Kubernetes (AWS EKS), AWS, MySQL, ElastiCache, Message brokers (AWS SQS and others), DataDog, Grafana, Terraform, Data Lake (GCP BigQuery). Get ready to Decide with the team what “working well” means for each service, measure it and set clear targets.Build dashboards and alerts that show problems early, from technical issues to business signals like a sudden drop in successful payments.Design systems that keep running when something breaks, like a server, a database or an external payment provider.Be part of the on-call rotation, help resolve incidents, and lead reviews afterwards so the same problem doesn’t happen twice.Keep our MySQL databases and caches healthy, backed up and ready to recover.Make sure messages between services are delivered and processed, and that nothing gets stuck or lost in queues.Prepare our systems for traffic peaks with capacity planning and load testing.Help keep our payment systems secure and compliant together with the security team.Work closely with developers and share reliability know-how. Ready to apply? Jump to the application form and start your journey with us. Apply Now We expect you to Proven experience (3+ years) in SRE, DevOps, platform or backend engineering, running production systems under high load.Production experience with Kubernetes, preferably EKS: deployments, autoscaling, resource tuning and troubleshooting.Good hands-on AWS knowledge, especially EKS, RDS, ElastiCache, SQS, IAM and VPC networking.Proficiency with Infrastructure as Code in Terraform or OpenTofu.Deep practical observability experience, ideally with DataDog, and a working understanding of SLO-based alerting.Experience operating MySQL in production: query performance, replication, failover and backups.A good understanding of distributed systems and asynchronous messaging: idempotency, retries, ordering, backpressure and exactly-once versus at-least-once trade-offs.Hands-on incident management and on-call experience.Coding or scripting skills in at least one language, such as Go, Python or Bash.Experience with GitLab CI/CD.An ownership mindset, clear communication and a collaborative approach. Nice to have Experience in payments, fintech or subscription billing.Experience working in PCI DSS-regulated environments.Familiarity with GCP BigQuery and data pipeline reliability.Experience with load testing (k6, Gatling) or chaos engineering. Salary Gross salary range is 5000-5900 EUR/month. Location We have several great workspaces to choose from, including our headquarters in Vilnius and our hubs in Klaipėda and Riga. For our Vilnius office, we follow a 4 days onsite, 1 day remote hybrid model. For roles based outside Vilnius, the work arrangement -onsite, hybrid, or remote-depends on the specific position. Speaking Of Perks Own your wellness Stay sharp and well in your own way – health insurance (after probation), on-site physiotherapist, office gym, and fitness classes keep you moving. Plus 7 extra days off, 3 for weddings, hybrid work, and the freedom to work 2 months from anywhere in the world. Boost your daily life Every breakthrough needs a spark, and our daily setup is designed to support it. Enjoy a pet-friendly office, fresh breakfasts, hot lunches, and fully stocked drawers and fridges for a quick snack or even a team-cooked meal together. Rooftop and cellar events keep things lively, while a public transport card for Vilnius and access to Perks.lt benefits help make everyday life a little easier. Shape your growth path Builders never stop learning. Coaching, mentorship, training, conferences, online courses, subscriptions, books – you’ve got a €1100 yearly budget to fuel personal or team growth. You pick the path – we’ll back you all the way. Drive global impact This is where ambition scales. Build impact across 30+ brands worldwide. Launch bold projects. Even step in as a co-founder – we reward those who push further. And if you bring in more great people, you get €1500 for every successful referral. additional conditions apply based on your residence location. Ready to apply? Jump to the application form and start your journey with us. Apply Now