Senior DevOps / Site Reliability Engineer

The Pictet Group — Luxembourg · Posted ~22 hours ago

Senior Contract

Skills

DevOps site reliability engineering monitoring observability application reliability performance optimization SRE

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Join an established technology organization as a Senior DevOps/SRE professional for a 12-month fixed-term assignment. You will strengthen the reliability and performance of strategic application processes by developing advanced monitoring and observability capabilities and supporting resilient, high-quality operations.

Highlights

A 12-month fixed-term opportunity within a sophisticated technology environment, with a strong focus on reliability, performance, monitoring, and observability for critical applications.

Description

Your team The Pictet Group is one of the world’s leading independent wealth and asset managers. Founded in 1805 and headquartered in Geneva, Switzerland, the Group is represented at 30 offices in financial centres worldwide, and currently employs over 5200 people. Pictet has been present since 1989 in Luxembourg, where it employs over 700 people. Pictet Techdivision specializes in designing and integrating cutting-edge software applications, including advanced portfolio management systems, sophisticated trading platforms, and comprehensive banking and corporate solutions. As a key contributor to the Group’s strategic advancements, the division plays a vital role in driving transformative innovations that enhance our services and deliver exceptional value to our clients. We are seeking a DevOps / Site Reliability Engineer to strengthen the reliability and performance of the bank’s strategic application processes through an advanced monitoring and observability framework across both on-premises and cloud environments. Monitoring is designed and implemented in close cooperation with software engineering teams in order to track end-to-end application chains. Industrialize application and business‑process deployments by designing, implementing, and maintaining CI/CD pipelines, ensuring secure production releases and a stable IT production environment.. Participation in a periodic 24/7 on-call shift is required. Your role Provide support to operations teams. Lead post-incident reviews, including Root Cause Analysis and post-mortems. Define and track preventive action plans.Collaborate with business and technical teams to align system reliability with operational needs and produce clear operational documentation.Industrialize and standardize the deployment of applications and business processes through automated CI/CD pipelines.Design, implement and maintain monitoring and observability solutions for business processes, including dashboards, reliability indicators, alerting and end-to-end flow traceability, with advanced use of the Grafana Observability Stack across on-premises systems and cloud services.Analyze business-process performance and recommend improvements.Integrate new processing tasks into the production schedule and/or optimize existing ones.Configure and implement data transfers with external partners, ensuring reliability, traceability and end-to-end monitoring.Contribute to the cloud strategy, including deployments, scalability, resilience, disaster recovery and business continuity planning.Automate recurring tasks through scripts, workflows and tooling in collaboration with development and operations teamsEnsure compliance with regulatory and data-protection requirements. Your profile Master's degree / Engineering degree in Computer science Fluent in french and english Minimum 7 years of experience in a similar role, ideally within a banking environment or fund administration Strong expertise in monitoring and observability tools, as well as distributed tracing concepts (Grafana, Dynatrace, Jaeger).Good knowledge of cloud environments from major providers (AWS, Azure, or equivalent), with a sound understanding of IaaS, PaaS and SaaS service models.Solid knowledge enabling effective use of Kubernetes and Docker infrastructures, including operations, scaling and monitoring.Proven hands-on experience with scripting and automation (Shell, Python, PowerShell, orchestration tools).Working knowledge of systems and networking (Linux/Unix, Windows Server, core networking fundamentals) to support diagnosis and troubleshooting, without acting as a day-to-day system administrator.Familiarity with DevOps concepts and Site Reliability Engineering practices, including reliability indicators, capacity management and scalability.Good command of SQL and ability to work with data in Oracle, PostgreSQL and Microsoft SQL Server (MSSQL) databases, as well as sound knowledge of Prometheus for metrics collection and analysis.Excellent analytical and problem-solving skills, ability to remain composed during critical incidents, strong communication skills (including in English), and the ability to explain technical topics clearly to non-specialists.Strong team spirit and ability to collaborate with a wide range of stakeholders, including business teams, IT teams and external partners.Good knowledge of ITIL processes (Incident, Problem, Change and Configuration Management), with the ability to apply them pragmatically in a critical IT production environment.Understanding of business processes related to fund administration / transfer agency and the associated regulatory framework. Note We will not accept any CVs via agencies Diversity & Inclusion Pictet is an equal opportunity employer and is committed to creating a diverse environment. We respect all individuals and seek their inclusion in the workplace.