Staff Software Engineer - Platform Observability

Adyen — Netherlands · Posted ~1 day ago

Senior Full-time Visa History ✓

Skills

Software engineering Platform observability Distributed systems System reliability Scalable systems Technical problem solving Observability Cloud infrastructure Platform engineering

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary

A Staff Software Engineer will help shape a platform observability function responsible for reliability across a large-scale financial technology platform. The role involves building systems that provide comprehensive operational data, supporting reliability, and solving challenging distributed-systems problems.

Highlights

High-impact staff-level engineering role focused on platform reliability and observability at significant scale, with opportunities to solve complex technical challenges and shape critical infrastructure.

Description

This is Adyen Adyen provides payments, data, and financial products in a single solution for customers like Meta, Uber, H&M, and Microsoft - making us the financial technology platform of choice. At Adyen, everything we do is engineered for ambition. For our teams, we create an environment with opportunities for our people to succeed, backed by the culture and support to ensure they are enabled to truly own their careers. We are motivated individuals who tackle unique technical challenges at scale and solve them as a team. Together, we deliver innovative and ethical solutions that help businesses achieve their ambitions faster. Staff Software Engineer - Platform Observability Adyen provides a global, unified platform with a wide variety of financial services. Observability plays a key role providing comprehensive data to a wide variety of use cases, ensuring the platform’s reliability and supporting our operations to run smoothly. As part of the Observability team you will build and maintain products and services that enable engineers at Adyen to understand how their services are behaving in real-time, reliably diagnose issues and streamline data discovery through a common observability ecosystem. Between logs, metrics and traces, our platform collects and processes billions of events per day. What You Will Do As a Staff Software Engineer working on the observability of the platform, you will play a key role in shaping how teams work with telemetry data, driving strategic decisions, identifying and advocating for best practices for building, running and maintaining observable components. This Includes Define and lead Logging, Metrics, Tracing & Alerting strategies Solve scaling bottlenecks in critical services in our telemetry data pipelinesCreate advanced tooling to accelerate root-cause analysis, reduce Mean Time to Resolution (MTTR), and eliminate alert fatigueGuide and mentor software engineering and SRE teams on monitoring best practices and instrumenting code.Identify and tackle bottlenecks and single points of failures in the architecture aiming to improve the reliability of the observability platform and the overall user experience.Driving features and products idealization, implementation and adoption.Collaborate on planning and refinement proactively providing input whilst striving for engineering alignment, quality and reducing complexity and dependencies.Respond to alerts and duty responsibilities, like providing support to other engineers and troubleshooting production issues What You Bring You have 5+ years of relevant work experience with highly distributed systemsExpertise in designing and implementing APIs and data pipelines for high-throughput, real-time data ingestionExperience improving software reliability between different categories including availability, performance, latency, efficiency, capacity, SLOs and incident management.Proficiency developing and maintaining software and frameworks written in Go and/or JavaAdvanced understanding of system design and evaluating its tradeoffsStrong stakeholder management, technical communication, and mentorship capability.Highly motivated to learn and continuously develop yourselfExperience with containerisation and orchestration technologies (Docker, Kubernetes) and infrastructure as code tools(e.g- Terraform)Observability Stack Expertise- You have hands-on experience operating core telemetry data stores at scale e.g. Elasticsearch/Opensearch/VictoriaLogs/Clickhouse for logging, Prometheus/ VictoriaMetrics for metrics and Grafana Tempo for distributed tracing, Grafana LGTM stack, OpenTelemetry, Alertmanager, Clickhouse Nice To Haves Experience with highly available/fault tolerant, replicated data storage systems, large scale data processing systems is a strong plusInfrastructure and Platform Experience Contributions to open-source observability projects Our Diversity, Equity and Inclusion commitments Our unique approach is a product of our diverse perspectives. This diversity of backgrounds and cultures is essential in helping us maintain our momentum. Our business and technical challenges are unique, and we need as many different voices as possible to join us in solving them - voices like yours. No matter who you are or where you’re from, we welcome you to be your true self at Adyen. Studies show that women and members of underrepresented communities apply for jobs only if they meet 100% of the qualifications. Does this sound like you? If so, Adyen encourages you to reconsider and apply. We look forward to your application! What’s next? Ensuring a smooth and enjoyable candidate experience is critical for us. We aim to get back to you regarding your application within 5 business days. Our interview process tends to take about 4 weeks to complete, but may fluctuate depending on the role. Learn more about our hiring process here. Don’t be afraid to let us know if you need more flexibility. This role is based out of our Amsterdam office. We are an office-first company and value in-person collaboration; we do not offer remote-only roles.