DevOps Engineer

Purequad — Romania · Posted ~22 hours ago

Mid Full-time

Skills

DevOps automation cloud operations system reliability CI/CD Cloud Automation

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A DevOps engineering role focused on improving platform stability, automation, operational excellence, and collaboration across software delivery teams.

Highlights

Work on reliability, automation, and continuous improvement while collaborating with diverse engineering teams.

Description

Mission As a DevOps Engineer for, you will be responsible for application reliability, operational excellence, automation, and continuous improvement. Beyond operating the platform, you will proactively identify risks, eliminate operational toil, increase automation, improve resilience, and drive end-to-end solutions that enhance the stability and efficiency of the service. You will collaborate closely with developers, QA, Product Owners, DBAs, Network, Security, and SRE teams. About the team Business DevOps team consists of DevOps engineers embedded across three product squads supporting different initiatives within the Business channels. DevOps engineers work closely within their squads alongside developers, business analysts, QA engineers, and Product Owners, while aligning with DevOps standards, engineering best practices, and operational guidelines defined across the organisation. The focus is twofold: driving the continuous evolution of the platform through the adoption of modern technologies, delivery of new capabilities, automation initiatives, and technical improvements, while ensuring the reliability, stability, security, performance, and operational excellence of existing production services. Your day to day Application Reliability & Troubleshooting • Investigate incidents using logs, metrics, traces, and runtime diagnostics to identify and resolve issues. • Troubleshoot and optimise applications across Test, Acceptance, and Production environments. • Drive corrective and preventive actions to improve reliability, resilience, and service stability. • Contribute to SLOs, availability, capacity, performance, and operational KPIs. • Escalate complex issues with clear diagnostics, RCA findings, and recommendations. Infrastructure, Web Servers & Platforms • Configure, maintain, and troubleshoot web servers and supporting components (JBoss, Apache, NGINX). • Perform deployments, upgrades, and releases across non-production and production environments for both legacy and containerised applications. • Support Oracle-related configurations and contribute to overall platform performance and stability. • Validate end-to-end service flows across application, database, network, and infrastructure layers. Automation & CI/CD • Design and maintain automation solutions using Azure DevOps, Ansible, Bash, and Python. • Build and improve CI/CD pipelines to ensure reliable, governed, and standardised software delivery. • Reduce operational toil by automating manual activities and simplifying operational processes. • Identify opportunities for self-service capabilities, process improvements, and AI-assisted automation. Observability & Monitoring • Maintain and enhance observability capabilities covering logs, metrics, alerts, traces, and health checks. • Develop dashboards and alerting solutions that provide visibility into platform health and performance. • Drive proactive monitoring through anomaly detection, trend analysis, and alert optimisation. Environment & Platform Operations • Manage and support production and non-production environments, ensuring consistency and operational readiness. • Perform OS patching, maintenance, and security hardening activities. • Maintain runbooks, technical documentation, and operational procedures. • Proactively identify and mitigate reliability, security, and capacity risks. Collaboration & Support • Collaborate with engineering, infrastructure, and business teams to ensure reliable service delivery. • Maintain ServiceNow records with appropriate documentation and traceability. • Contribute to Agile ceremonies, user stories, and sprint activities. • Share knowledge, support onboarding activities, and promote engineering best practices. What you bring to the team • Solid understanding of DevOps practices, including automation, deployment workflows, operational reliability, and excellent technical documentation skills. • Understanding of both monolithic and microservices architectures, including how they differ from an operational, deployment, and reliability perspective • Webservers expertise is mandatory, with handson experience configuring, maintaining, and troubleshooting platforms like JBoss, Apache, and NGINX • Experience with running applications in Kubernetes • Scripting skills (Ansible, Bash, Python) • Solid UNIX/Linux expertise with strong troubleshooting abilities across system, application, and infrastructure layers • Handson CI/CD experience, particularly with Azure DevOps, covering build, test, release, and environment governance. Github Actions is a plus • Experience with modern observability stacks, including logging and monitoring tools such as ELK, Grafana, Kibana, Prometheus • Familiarity with SRE concepts: SLI, SLO, Error budget • Solid analytical and problemsolving skills, able to quickly diagnose issues and propose effective solutions. • Familiarity with Scrum and ITIL, comfortable operating within Agile delivery and operational governance frameworks.