Site Reliability Engineer

Quik Hire Staffing — United Arab Emirates · Posted ~3 hours ago

Mid Full-time Remote

Skills

Site reliability engineering High-availability systems SLOs SLIs Error budgets Infrastructure automation Infrastructure as code Monitoring Incident response Root cause analysis Scripting Infrastructure as Code SLO SLI

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Work remotely from anywhere as a Site Reliability Engineer responsible for keeping critical systems highly available, scalable, and efficient. You will design infrastructure, establish reliability objectives, automate operational processes, monitor system health, investigate incidents, and drive preventative improvements. The role suits an engineer who enjoys solving complex operational challenges and improving system resilience.

Highlights

Fully remote, full-time SRE opportunity open to candidates working from anywhere. The role focuses on scalable infrastructure, reliability engineering, automation, observability, incident management, and performance and cost optimization.

Description

Role: Site Reliability Engineer - LInE (Remote)Location: Remote (Work from Anywhere) Role Overview: We are hiring for one of our clients, seeking a Site Reliability Engineer (LInE) to work on a Full-Time basis. This role requires expertise in maintaining and optimizing high-availability systems to ensure seamless operations for end users. The position involves collaborating with cross-functional teams to implement and monitor infrastructure solutions. Key Responsibilities: • Design, implement, and maintain scalable infrastructure to support critical services and applications. • Develop and enforce SLOs, SLIs, and error budgets to ensure system reliability and performance. • Troubleshoot and resolve incidents, performing root cause analysis and implementing preventative measures. • Automate operational tasks using scripting languages and infrastructure-as-code tools. • Monitor system health, analyze metrics, and optimize resource utilization for cost efficiency. Required Skills & Qualifications: • Proficiency in Linux system administration and troubleshooting. • Experience with cloud platforms such as AWS, GCP, or Azure. • Knowledge of containerization and orchestration tools like Docker and Kubernetes. • Familiarity with monitoring and observability tools such as Prometheus, Grafana, or Datadog. • Strong scripting skills in Python, Bash, or Go for automation and tooling. More About the Opportunity: This role offers a unique opportunity to work with a global leader in the Software Development industry, contributing to the delivery of high-performance, fault-tolerant systems. The position involves direct impact on user-facing services and infrastructure decisions. Equal Opportunity Employer: We hire based on skills and expertise. All qualified candidates are welcome regardless of background, experience, or prior employment history. Applications are reviewed solely on demonstrated technical ability and qualifications. Apply Now!