Summary
Join a team responsible for keeping large-scale production systems reliable and scalable. Build cloud-native infrastructure, automate operations, improve observability, contribute to CI/CD pipelines, and collaborate across engineering teams while solving complex operational challenges.
Highlights
Hybrid role focused on reliability engineering, cloud infrastructure, automation, modern DevOps practices, and production systems with opportunities to work across diverse cloud-native technologies.
Description
We are Looking for SRE Engineer Based in Poland which is Hybrid
Roles and RESPONSIBILITIES
Maintain and support production systems, ensuring high availability, reliability, and scalability.
Assist in implementing SRE best practices including monitoring, alerting, SLOs/SLIs, and incident response.
Deploy, configure, and manage containerized applications using Docker and Kubernetes.
Contribute to CI/CD pipeline development and configuration to streamline development and deployment processes.
Automate repetitive tasks and processes using scripting languages such as Python, Bash, or Go.
Collaborate with development, QA, and operations teams to address issues and drive improvements across the stack.
Participate in on-call support rotation and help resolve incidents in a timely manner.
Document procedures, configurations, and post-incident reviews to foster a learning culture.
SKILLS & EXPERIENCE REQUIRED
Bachelor’s degree in Computer Science, Engineering, or a related discipline, or equivalent practical experience.
At least 3-5 years of hands-on experience in SRE, DevOps, or Production Support roles.
INTERNAL
Working knowledge of containerization technologies (Docker, Kubernetesl,Istio, Helm) and cloud platforms (AWS, Azure, or GCP).
Experience with monitoring and logging tools (e.g., Prometheus, Grafana, ELK, Zipkin, Jeager, Datadog, etc.).
Familiarity with CI/CD tools such as Jenkins, GitLab CI, or similar.
Familiar with kubernetes gitops practice and toolings, such as Argo CD, Flux CD or Tekton;
Scripting/programming proficiency in Python, Bash, Go, or comparable languages.
Strong troubleshooting skills and willingness to dive into complex problems.
Nice to have:
Experience with infrastructure-as-code (Terraform, Ansible, or similar tools)
Exposure to microservices and distributed systems
Awareness of ITIL, incident/change management
Experience working in 24x7 or high-availability production environments.