Site Reliability Engineer, Infrastructure Platforms

Gitlab Com — United Kingdom · Posted ~2 hours ago

Senior Remote

Skills

Site reliability engineering Infrastructure platforms Cloud infrastructure DevOps Automation Observability Incident management AI tools

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A remote site reliability engineering role supporting infrastructure platforms at significant scale. You will help improve reliability, operational efficiency, and developer productivity while collaborating in a globally distributed engineering environment with strong emphasis on continuous learning.

Highlights

Join a globally distributed engineering organization with a strong knowledge-sharing culture, high autonomy, and an emphasis on operational excellence, innovation, and continuous improvement.

Description

GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. An Overview Of This Role Site Reliability Engineers keep GitLab's user-facing services and production systems running reliably at scale. They combine software engineering with operational excellence, applying sound engineering principles, automation, and continuous improvement to build, operate, and evolve our production infrastructure. This is a single application for Site Reliability Engineering opportunities across Infrastructure Platforms. Rather than asking you to choose the right team or level upfront, we evaluate your skills holistically and match you to the opportunity that best aligns with your experience and our hiring needs. We hire Site Reliability Engineers from Intermediate through Senior Staff across multiple Infrastructure Platforms teams. We don't expect every candidate to have experience with every technology in our environment. We're looking for engineers with strong technical fundamentals, a growth mindset, and the ability to learn quickly. We'll support you in becoming successful with GitLab's tools, systems, and ways of working. Please note: This position is open to candidates based in the United Kingdom only. Candidates based in the United States or Canada can apply to this posting: Site Reliability Engineer, Infrastructure Platforms — AMER (Intermediate to Senior Staff) How Our SRE Hiring Works Because this is a single application for SRE roles across Infrastructure Platforms, our process is built to evaluate you once and match you well, rather than interviewing separately for every team. Recruiter Screen: A conversation about your background, what you're looking for, and the level and teams that fit, so we can point your process in the right direction.Core Technical: The shared assessment every SRE candidate takes, regardless of eventual team. A low-stress, collaborative discussion covering system architecture and incident review.Hiring Manager Interview: A conversation about ownership, judgment, execution, collaboration, and growth, the non-technical signals that make an SRE effective at GitLab.Peer Technical: Team-specific depth, run by SREs from the team you're most likely to join, focused on the problems that team actually works on.Skip-Level Interview: A conversation with a senior leader on values alignment, and how you'll work across teams. After your interviews, we consider your performance alongside our current hiring needs to confirm the level and team where you'll do your best work. Interview results are a major factor, and final placement also reflects our active hiring priorities at the time. We’ll calibrate your level throughout the interview process based on the scope and impact of your experience. Intermediate: You independently deliver meaningful reliability improvements within a defined area.Senior: You own complex reliability work end to end and raise the effectiveness of your team.Staff: You shape reliability across multiple teams, solving systemic problems and creating approaches others can reuse.Senior Staff: You set technical direction across a broader Infrastructure area and influence reliability strategy at organizational scale. What You'll Do Keep user-facing services and production systems reliable, scalable, and efficientBuild automation and tooling that reduces toil and replaces manual work with repeatable, infrastructure-as-code-driven workflowsOperate and troubleshoot production systems on Kubernetes, including deployments, rollouts, and scalingWrite and maintain infrastructure as code, and ship changes safely through CI/CD and GitOpsParticipate in on-call, triage alerts, follow and improve runbooks, and escalate appropriatelyContribute to the observability stack, using metrics, logs, and SLOs to detect symptoms early rather than just outagesTake part in incident response and post-incident reviews, turning learnings into changes in automation and processDocument runbooks, architecture decisions, and reviews so your findings become repeatable practices What You'll Bring Experience keeping production systems reliable, combining an operations mindset with real software engineering practiceExperience building net-new infrastructure tooling and automation, not just configuring existing tools. For example, Terraform modules, Kubernetes operators or controllers, or production automation and services written from scratchThe ability to read, debug, and reason about code. Most of our teams work in Go; some work in Ruby. You can discuss a piece of code's behavior, performance, and failure modesExperience with infrastructure as code, and with Kubernetes and its ecosystem, at a depth appropriate to your levelHands-on experience with at least one major cloud provider (GCP or AWS)Familiarity with observability practices, including metrics, logging, alerting, and SLOs or SLIs, and using data to inform operational decisionsComfort participating in on-call and incident response, with a structured approach to troubleshooting under pressureStrong written communication and the ability to operate as a manager-of-one in an async, distributed environmentA track record of using automation, and increasingly AI, to reduce toil and improve how you and your team workAlignment with GitLab's values and a commitment to working in accordance with them About The Team Infrastructure Platforms is responsible for the availability, reliability, performance, and scalability of GitLab’s user-facing services, most notably GitLab.com. The organization spans teams across Production Engineering, GitLab Dedicated, GitLab Delivery, and Developer Experience, covering everything from the production fleet and networking platform to observability, incident response, deployment infrastructure, tenant scale, and our single-tenant Dedicated offering. We are a globally distributed, remote-first organization that works asynchronously, favors automation over toil, and uses monitoring, metrics, and clear ownership to continuously improve the reliability of GitLab at scale. For more on how we work, see the Infrastructure Handbook Page. How GitLab Supports Full-Time Employees Benefits to support your health, finances, and well-beingFlexible Paid Time Off Team Member Resource GroupsEquity Compensation & Employee Stock Purchase PlanGrowth and Development FundParental Leave Please note that we welcome interest from candidates with varying levels of experience; many successful candidates do not meet every single requirement. Additionally, studies have shown that people from underrepresented groups are less likely to apply to a job unless they meet every single qualification. If you're excited about this role, please apply and allow our recruiters to assess your application. Country Hiring Guidelines: GitLab hires new team members in countries around the world. All of our roles are remote, however some roles may carry specific location-based eligibility requirements. Our Talent Acquisition team can help answer any questions about location after starting the recruiting process. Privacy Policy: Please review our Recruitment Privacy Policy. Your privacy is important to us. GitLab is proud to be an equal opportunity workplace and is an affirmative action employer. GitLab’s policies and practices relating to recruitment, employment, career development and advancement, promotion, and retirement are based solely on merit, regardless of race, color, religion, ancestry, sex (including pregnancy, lactation, sexual orientation, gender identity, or gender expression), national origin, age, citizenship, marital status, mental or physical disability, genetic information (including family medical history), discharge status from the military, protected veteran status (which includes disabled veterans, recently separated veterans, active duty wartime or campaign badge veterans, and Armed Forces service medal veterans), or any other basis protected by law. GitLab will not tolerate discrimination or harassment based on any of these characteristics. See also GitLab’s EEO Policy and EEO is the Law. If you have a disability or special need that requires accommodation, please let us know during the recruiting process.