Site Reliability Engineer

Quik Hire Staffing — Germany · Posted ~3 hours ago

Full-time Remote

Skills

Site Reliability Engineering High-availability systems Infrastructure engineering SLOs SLIs Error budgets Incident response Root cause analysis Automation Infrastructure as Code Monitoring Scripting SLO SLI

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A fully remote, full-time Site Reliability Engineer position focused on designing and maintaining scalable infrastructure for critical services. You will establish reliability objectives, manage incidents, automate operations with scripting and infrastructure-as-code, monitor system health, perform root cause analysis, and optimize resources for efficiency.

Highlights

Fully remote, full-time SRE role focused on scalable infrastructure, reliability engineering, automation, monitoring, and operational efficiency. The position offers broad ownership of critical systems and collaboration across engineering functions.

Description

Role: Site Reliability Engineer - LInE (Remote)Location: Remote (Work from Anywhere) Role Overview: We are hiring for one of our clients, seeking a Site Reliability Engineer (LInE) to work on a Full-Time basis. This role requires expertise in maintaining and optimizing high-availability systems to ensure seamless operations for end users. The position involves collaborating with cross-functional teams to implement and monitor infrastructure solutions. Key Responsibilities: • Design, implement, and maintain scalable infrastructure to support critical services and applications. • Develop and enforce SLOs, SLIs, and error budgets to ensure system reliability and performance. • Troubleshoot and resolve incidents, performing root cause analysis and implementing preventative measures. • Automate operational tasks using scripting languages and infrastructure-as-code tools. • Monitor system health, analyze metrics, and optimize resource utilization for cost efficiency. Required Skills & Qualifications: • Proficiency in Linux system administration and troubleshooting. • Experience with cloud platforms such as AWS, GCP, or Azure. • Knowledge of containerization and orchestration tools like Docker and Kubernetes. • Familiarity with monitoring and observability tools such as Prometheus, Grafana, or Datadog. • Strong scripting skills in Python, Bash, or Go for automation and tooling. More About the Opportunity: This role offers a unique opportunity to work with a global leader in the Software Development industry, contributing to the delivery of high-performance, fault-tolerant systems. The position involves direct impact on user-facing services and infrastructure decisions. Equal Opportunity Employer: We hire based on skills and expertise. All qualified candidates are welcome regardless of background, experience, or prior employment history. Applications are reviewed solely on demonstrated technical ability and qualifications. Apply Now!