Summary
✨ AI‑Generated
A recruitment-focused firm seeks a production support engineer to own incident workflows, perform root cause analysis, monitor services, and drive CI/CD deployments. The role blends on-site presence with remote flexibility.
Highlights
Hybrid role across Polish cities, managing incidents, RCA, monitoring, CI/CD deployments, and proactive stability maintenance.
Description
As a recruitment company, DCG understands that every business is powered by experienced professionals.
Our management style and partnership approach enable us to meet your needs and provide continuous support.
Due to our ongoing growth and the large number of recruitment projects we undertake for our partners, we are currently looking for:
Production Support Engineer
Hybrid work model from Warsaw, Lodz, Gdynia or Gdansk (3 days per week in office)
Salary ranges: 120-125PLN net+VAT/h
Responsibilities:
Own incident and problem management processes, including RCA, resolution, and preventive actionsMonitor production service health, detect anomalies, and respond before issues become incidentsExecute and oversee deployments through CI/CD pipelines and coordinate rollback activities when requiredTroubleshoot failed deployments and support release execution processesMaintain stability and alignment of Pre-Production and Production environmentsDocument operational procedures, known issues, and resolution steps in runbooks and knowledge repositoriesCollaborate with development and platform teams to triage incidents and address operational requirementsIdentify recurring operational issues and propose automation or tooling improvementsImprove observability through dashboards, alerts, and log analysisContribute to service continuity activities and disaster recovery exercises
Requirements:
Proven experience in IT operations, application support (2nd/3rd line), or a similar production-facing role with 5+ years of experienceDemonstrated ownership of incidents end-to-end, from alert handling through RCA and preventionPractical experience working within an ITIL framework, including incident, problem, and change management, with 2+ years of experienceExperience working in Agile delivery environments alongside development teamsExcellent English communication skills with the ability to explain technical issues to both technical and non-technical stakeholdersHands-on experience with Splunk, Apica, and Sysdig for log analysis, monitoring, and alertingStrong understanding of Prometheus and Grafana for dashboard analysis and alert tuningPractical experience operating services on Kubernetes and Linux CLI, including pod health checks, log analysis, and service restartsExperience executing and troubleshooting Jenkins deployment pipelinesProficiency with Git version control systemsExperience querying and troubleshooting relational databases including Oracle and DB2Working knowledge of Spring/Hibernate applications, Kafka message flows, and XML/JSON payload analysisNice to have:
Experience with Helm deploymentsJava/J2EE development backgroundOperational experience with IBM DataStageScripting skills in Bash or Python for automationExperience using Ansible for controlled configuration changes
Offer:
Private medical careCo-financing for the sports cardConstant support of dedicated consultantEmployee referral program