Senior Site Reliability Engineer

Bairesdev — United States · Posted ~1 week ago

Senior Full-time Remote

Skills

Site Reliability Engineering Infrastructure Engineering Kubernetes Multi-cluster Management CI/CD Infrastructure as Code Terraform Helm Observability Prometheus Grafana Datadog OpenTelemetry SLOs and SLIs Error Budgets Incident Management Cloud Security IAM Hardening English IAM

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary

Join a globally distributed engineering organization as a Senior Site Reliability Engineer responsible for keeping critical systems reliable, observable, and resilient at scale. You will manage multi-cluster Kubernetes environments, build CI/CD and infrastructure-as-code solutions, implement modern observability, define SLOs and SLIs, manage error budgets, lead post-incident reviews, and strengthen cloud security and IAM practices. The role is fully remote, offers flexible scheduling, and is designed for an experienced reliability or infrastructure engineer with advanced English communication skills.

Highlights

Fully remote work from anywhere, flexible hours, compensation available in USD or local currency, home-office hardware and software support, paid parental leave, vacations and holidays, and strong mentorship and career-development opportunities. The role offers significant technical ownership over reliability, observability, resilience, and infrastructure at scale within a globally distributed engineering environment.

Description

At BairesDev®, we've been leading the way in technology projects for over 15 years. We deliver cutting-edge solutions to giants like Google and the most innovative startups in Silicon Valley. Our diverse 4,000+ team, composed of the world's Top 1% of tech talent, works remotely on roles that drive significant impact worldwide. When you apply for this position, you're taking the first step in a process that goes beyond the ordinary. We aim to align your passions and skills with our vacancies, setting you on a path to exceptional career development and success. Senior Site Reliability Engineer (SRE) at BairesDev In this role, you'll ensure systems stay reliable, observable, and resilient at scale, combining deep infrastructure expertise with a data-driven approach to reliability. Working across multi-cluster environments and modern observability stacks, you'll define what reliability actually means for critical systems and build the practices that keep them there. This is your opportunity to work where engineering rigor meets operational excellence, directly shaping the uptime and performance that users and businesses depend on. What You'll Do: - Manage and scale Kubernetes environments across multiple clusters. - Build and maintain CI/CD pipelines and Infrastructure as Code. - Implement and maintain observability across systems for proactive issue detection. - Define and track reliability standards, driving continuous improvement through incident learnings. What we are looking for: - 5+ years of experience in Site Reliability Engineering or infrastructure engineering. - Strong expertise in Kubernetes, including operators, autoscaling, and multi-cluster management. - Experience with CI/CD pipeline engineering. - Proficiency in Infrastructure as Code using Terraform and Helm. - Hands-on experience with observability stacks such as Prometheus, Grafana, Datadog, or OpenTelemetry. - Experience defining SLOs/SLIs, managing error budgets, and leading post-incident reviews. - Background in cloud security tooling and IAM hardening. - Advanced proficiency in English. How we do make your work (and your life) easier: - 100% remote work (from anywhere). - Excellent compensation in USD or your local currency if preferred - Hardware and software setup for you to work from home. - Flexible hours: create your own schedule. - Paid parental leaves, vacations, and holidays. - Innovative and multicultural work environment: collaborate and learn from the global Top 1% of talent. - Supportive environment with mentorship, promotions, skill development, and diverse growth opportunities. Apply now and become part of a global team where your unique talents can truly thrive!