DevOps Engineer

Agap2Netherlands — Netherlands · Posted ~4 hours ago

Visa History ✓

Skills

DevOps Cloud Infrastructure CI/CD Automation System Administration Cloud

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A DevOps engineering position supporting multiple technology projects, improving deployment processes, and working with modern infrastructure and automation practices.

Highlights

International consulting environment with exposure to diverse industries, knowledge sharing, and continuous learning opportunities.

Description

About us Agap2 is a European IT and Engineering Consultancy, active in 14 countries with over 9.500 employees. The company was founded in Portugal and since 2014 we are based in the Netherlands with our Amsterdam office to provide our expertise to the Dutch market. Our employees are defined by the so-called “Agapian spirit”, which is summarized in 5 keywords: family, union, friendship, team spirit, and learning Did you identify with what defines us? Then read on to find out more. The position As a DevOps Engineer at agap2, you take on the challenge within an international environment with a local focus. You will have the opportunity to work on various projects over the years, based on your experience and interests. Projects within industries such as Banking, Fintech, Telco, Retail, Sports, and Healthcare. While working at Agap2, you are expected to share your knowledge while also having the opportunity to learn new technologies. We believe that sharing knowledge is the key to personal development. Our talent management program is tailor-made to give you the opportunity to develop and grow in your career, the way you envision it. Because we work with various customers, you are also expected to be able to adapt to different environments. Key responsibilities Design, implement, and maintain observability platforms leveraging Grafana OSS/Enterprise and related technologies, including Prometheus/Mimir, Loki, Tempo, and OpenTelemetry.Define, collect, monitor, and report key performance and reliability KPIs across infrastructure, applications, and user-facing services.Develop and maintain Grafana dashboards, alerts, and monitoring reports that provide actionable insights to engineering and operations teams.Collaborate with DevOps, SRE, and application development teams to instrument services and applications for metrics, logs, and distributed traces.Integrate observability platforms with incident management, alerting, and notification systems.Optimize monitoring infrastructure for data retention, storage utilization, query performance, scalability, and cost efficiency.Establish and promote best practices and governance for monitoring, alerting, observability, and root cause analysis.Define and monitor SLIs and SLOs and help teams establish meaningful reliability targets.Troubleshoot observability infrastructure and resolve issues affecting monitoring, logging, tracing, or alert delivery. What You bring to the position Proven hands-on experience with the Grafana observability stack, including: Grafana, Prometheus and/or Mimir, Loki, TempoStrong knowledge and practical experience with OpenTelemetry for application and infrastructure instrumentation and telemetry collection.Experience defining and monitoring KPIs for:Infrastructure performance, including CPU, memory, network, and storageApplication performance and availabilityService-level indicators (SLIs) and objectives (SLOs)Hands-on experience working with containerized environments, particularly Docker and Kubernetes.Solid understanding of networking, distributed systems, and application performance monitoring (APM).Experience designing effective dashboards, alerts, and monitoring strategies for production environments.Strong troubleshooting and analytical skills, with the ability to investigate complex performance and reliability issues.Ability to collaborate effectively with DevOps, SRE, application development, and operations teams. Nice to have: Experience with enterprise observability and monitoring platforms such as Datadog, Splunk, Elastic, or New Relic.Familiarity with incident management processes and platforms, particularly ServiceNow.Experience supporting observability platforms in large-scale or highly distributed production environments.Knowledge of cloud platforms and cloud-native monitoring architectures.Experience establishing observability standards, governance, and best practices across engineering organizations. What we offer Competitive Salary & Benefits: Including a laptop, mobile phone, full travel allowance, and pension plan.Work-Life Balance: Vacation days, flexible working hours in a hybrid mode.Professional Growth: A tailored talent management program with a financial budget to help you achieve your professional goalsDynamic Culture: Diverse team activities from tech meetups to sports and social events. Is there still room for fun? Of course! The combination of fun and development is essential in our team. A small impression: our team consists of more than 10 nationalities; we play paddle and regularly organize tech meetups, happy hour drinks, and pub quizzes. At agap2, we defend equality and value diversity. We create a safe, diverse environment where opportunities are equal for all employees. All candidates with skills matching the position are welcome to join us!