Site Reliability Engineer

Venquis Ltd — Poland · Posted ~1 hour ago

Mid Full-time

Skills

SRE DevOps Cloud infrastructure Monitoring Incident response Automation Scalability Security Cloud

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A reliability engineering role responsible for maintaining highly available services, improving operational efficiency, and supporting scalable production environments.

Highlights

Opportunity to improve reliability of large-scale services, automate operations, and collaborate with development teams on critical systems.

Description

Site Reliability Engineer Job Description About the Role As a Site Reliability Engineer, your overarching responsibility is to ensure we meet our customers’ Service Level Agreements, and that we respond to incidents in a timely and professional manner. You will proactively monitor production environments to ensure scalability, availability, and security, improving the alignment of service performance to customers and co-workers by reducing or eliminating manual and repetitive tasks, and removing bottlenecks and inefficiencies from services. You will create, deliver, and manage business critical services that are used 24/7 by customers and co-workers. You will work closely with Development and DevOps teams to give them the tools they need and support the application release process, and you will be involved in designing, building, and scaling our global product platform. We welcome engineers from development, DevOps, SRE, or similar backgrounds who want to grow their career in Site Reliability Engineering. Responsibilities Spend an equal amount of time building software to automate manual work and providing operational support to the products you cover, balancing feature development speed and reliability against service-level objectivesLead incident response, diagnosis, and follow-up on system outages or alertsPerform and assist in root cause analysis and blameless post-mortems, enabling incidents to be understood and avoided in futurePropose improvements to infrastructure and productImprove the reliability, quality, and time-to-market of our software solutionsProvide out-of-hours support based on an on-call rotaSkills and Experience Experience with AzureExperience with KubernetesProficiency in a programming language such as Python or GoA track record of writing code you care about, including unit tests, integration tests, static analysis, and resilience testsExperience with database technologies such as MySQLExperience with Infrastructure as Code tools such as TerraformA strong drive to engineer solutions using best practicesA security-first design philosophyAn appetite for learning new skills and pushing yourself furtherA mindset suited to complex, polymorphic problem-solving