Site Reliability Engineer
Obsidiansecurity — United Kingdom · Posted ~22 hours ago
🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.
Log in to add to target listDescription
Founded in 2017, Obsidian Security was created to close a critical gap: securing the SaaS applications where modern business happens—platforms like Microsoft 365, Salesforce, and hundreds more.
Backed by top investors including Greylock, Norwest Venture Partners, and IVP, we’ve built a complete SaaS security platform to reduce risk, detect and respond to threats, and prevent breaches at the source.
Our team includes leaders who helped define the categories of endpoint and identity security at CrowdStrike, Okta, Cylance, and Carbon Black.
Now, we’re transforming how SaaS is secured—in the era of agentic AI.
Today, Obsidian is trusted by global enterprises like Snowflake, T-Mobile, and Pure Storage.
We protect more than 200 organizations across North America, Europe, the Middle East, Southeast Asia, Australia, and New Zealand—including many of the world’s largest Fortune 1000 and Global 2000 companies.
With strong global momentum, a growing partner ecosystem including SentinelOne, Databricks, and Google Cloud, and a major fundraise on the horizon, we’re scaling quickly toward long-term growth and IPO readiness.
Join us as we define the future of SaaS security!
Site Reliability Engineer (UK)
At Obsidian, our Site Reliability Engineers ensure the reliability, scalability, and operational excellence of a complex multi-tenant SaaS platform serving enterprise and financial customers.
As an SRE, you will work closely with DevOps, Platform Engineering, and product teams to improve system observability, incident response, and service resilience across the platform.
This is a hands-on engineering role focused on building operational excellence through monitoring, automation, debugging, and continuous improvement.
You will help ensure that issues are detected and addressed quickly while contributing to systems that improve platform reliability at scale.
Key Responsibilities
Reliability Engineering: Improve the reliability, availability, and resiliency of Obsidian’s production systems and distributed servicesDetection & Observability: Build and maintain monitoring, alerting, dashboards, and observability tooling to enhance system visibility and reduce operational noiseIncident Response & Operations: Support incident response, on-call operations, troubleshooting, and postmortem processes to drive operational excellenceCollaboration: Partner with engineering teams to implement SLI/SLO practices, operational standards, and reliability-focused workflowsExecution: Automate infrastructure operations, deployment workflows, and platform tooling across Kubernetes, cloud infrastructure, and data pipelines
Required Qualifications
2–5 years of experience in Site Reliability Engineering, DevOps, Production Engineering, or related rolesExperience operating and supporting production systems in AWS and/or GCPFamiliarity with Kubernetes and Helm in cloud-native environmentsExperience with observability and monitoring tools such as Prometheus, Grafana, Datadog, or similar platformsExposure to CI/CD systems such as GitLab CI/CD, GitHub Actions, ArgoCD, or equivalentStrong troubleshooting and debugging skills across distributed systems and microservicesExperience writing automation or infrastructure tooling using scripting or programming languagesStrong systems thinking and a collaborative engineering mindset
Preferred Qualifications
AI Agent development experienceExperience supporting SaaS platforms in production environmentsFamiliarity with incident management and postmortem practicesExposure to infrastructure-as-code and GitOps workflowsUnderstanding of SLI/SLO concepts and operational metricsExperience with enterprise-scale monitoring or customer-facing production systems
Why This Role
Work on reliability challenges across a large-scale distributed SaaS platformBuild and improve observability and operational tooling used across engineeringGain hands-on experience with cloud infrastructure, Kubernetes, and production systemsHelp safeguard critical services for enterprise and financial customers
What Success Looks Like
Production issues are detected and resolved quicklyMonitoring and alerting provide clear, actionable operational insightsReliability metrics and operational practices improve over timeEngineering teams can effectively troubleshoot and self-serve observabilityAutomation reduces operational toil and improves platform stability
Employee BenefitsOur competitive benefits packages are designed to support our employees' well-being, both at work and at home.
Our US based employees enjoy:Competitive compensation with equity and 401kComprehensive healthcare with dental and vision coverageFlexible paid time off and paid holiday time off 12 weeks of new parent or family leavePersonal and professional development resourcesFor more details on our US benefits, or for information on our international benefits, please see here.Pay TransparancyPlease note that the base pay range is a guideline and for candidates who receive an offer, the base pay will vary based on factors such as work location, as well as the knowledge, skills and experience of the candidate.
In addition to a competitive base salary, this position is eligible for equity awards and may be eligible for sales commission or incentive compensation based on the role or function within the company.At Obsidian, we are proud to be an equal-opportunity employer.
We value diversity and hire for talent, passion, and compassion.
In compliance with federal law, all persons hired will be required to submit satisfactory proof of identity and legal authorization.
If you have a need that requires accommodation, please contact accommodations@obsidiansecurity.comInformation collected and processed as part of any job applications you choose to submit is subject to Obsidian’s Applicant Privacy Policy.
Base Salary Range: £85,000 GBP - £103,000 GBP
We have 70,403 jobs that might be an even better fit for you
DontApply's real value goes far beyond a single job link or company name. Just upload your resume — in under a minute we'll analyze all 70,403 jobs and tell you exactly which ones you should apply to right now.
Upload My Resume