Description
The Quant DevOps Engineer responsible for the reliability, scalability, security, and cost efficiency of large scale, data intensive cloud workloads.
As a leading global reinsurer, SCOR offers its clients a diversified and innovative range of reinsurance and insurance solutions and services to control and manage risk.
Applying "The Art & Science of Risk," SCOR uses its industry-recognized expertise and cutting-edge financial solutions to serve its clients and contribute to the welfare and resilience of society in around 160 countries worldwide.
Working at SCOR means engaging with some of the best minds in the industry - actuaries, data scientists, underwriters, risk modelers, engineers, and many others - as we work together to find solutions to pressing challenges facing societies.
As an international company, our common culture is defined by "The SCOR Way." Serving both to build momentum that drives the Group forward and as a compass to guide our actions and choices, The SCOR Way is anchored by five core values, reflecting the input of employees at all levels of the Group.
We care about clients, people, and societies.
We perform with integrity.
We act with courage.
We encourage open minds.
And we thrive through collaboration.
SCOR supports inclusion and the diversity of talents, and all positions are open to people with disabilities.
The Quant DevOps Engineer o wns and operates the application engineering platform, cloud infrastructure, and data execution environment that application and modeling workloads depend on.
The role is accountable for the end‑to‑end reliability, scalability, security, performance, and cost efficiency of the platform, supporting large‑scale, data‑intensive and compute‑heavy workloads such as parallel Databricks clusters and high‑memory systems.
This position requires broad and deep technical expertise, strong operational judgment, and the ability to act as a central technical interface between engineering teams, IT operations, security/SecOps, and data users.
It is a senior ownership role with significant impact on delivery speed, platform stability, and infrastructure cost.
Platform & Infrastructure Ownership
End-to-end ownership and delivery accountability for the application execution platform and production model runs (Databricks, Servers).Day-to-day operational responsibility: availability, incident handling, runtime management, and delivery continuity for business-critical runs (incl.
peak periods like YE).Manage Azure cloud infrastructure, including: Virtual machines, storage, identity and access management (RBAC)Networking components such as firewalls, peering, and cross‑subscription connectivityLead standardization with central teams across observability, security controls, platform services, and "golden paths," while keeping delivery runningEnsure platform reliability, scalability, and long‑term sustainability
CI/CD & Automation
Design, maintain, and continuously improve CI/CD pipelines using Azure DevOps and related toolingBuild and evolve automation using scripting and build toolsOptimize pipeline performance, reliability, and parallel execution to support large‑scale workloads
Container & Runtime Management
Own Docker image creation, lifecycle management, and governanceOptimize build processes, caching strategies, and container securitySupport containerized execution environments for compute‑heavy workloads and services
Infrastructure as Code & Configuration Management
Maintain and evolve Infrastructure as Code using TerraformOperate and improve configuration management systems (e.g.
SaltStack)Reduce configuration drift and improve reproducibility across environments
Observability, Security & Secrets
Own and operate observability platforms (e.g.
ELK, Prometheus, Grafana)Ensure meaningful metrics, logs, dashboards, and alerting are in placeManage secrets and platform security tooling (e.g.
Wiz, Snyk)Collaborate closely with Security and SecOps teams on controls, findings, and improvements
Data Platform & Capacity Planning
Configure/support Databricks usage for application workloads; manage workspace-level configuration/permissions as delegated; partner with central Databricks lead for global administration and optimization.Support large‑scale data and compute workloadsLead capacity planning for: Highly parallel Databricks clusters (e.g.
up to ~10 × 80‑node clusters)Memory‑intensive systems (multi‑terabyte RAM)Data pipelines producing terabytes of dataBalance performance, reliability, and cost across platform decisions
Operations & Incident Response
Act as senior escalation point for platform and infrastructure incidentsParticipate in a limited on‑call rotationInvestigate incidents and execute or coordinate remediationPerform manual interventions when automation is insufficientDrive post‑incident reviews and platform improvements
Cross‑Team & Organizational Coordination
Serve as primary technical contact for: IT OperationsSecurity and SecOpsArchitecture and governance bodiesService management processesCoordinate platform‑related work across teamsSupport customer‑facing technical discussions related to platform capabilities and constraints
Experience
Background in Insurance, Finance, or Scientific / High‑Performance Computing environmentsStrong experience in platform or DevOps engineering within production environmentsSolid expertise in cloud infrastructure, preferably Microsoft AzureHands‑on experience with CI/CD, Infrastructure as Code, and container platformsProven experience operating data‑intensive and compute‑heavy systems
Technical Competencies
Strong troubleshooting and operational mindsetAbility to manage and balance competing constraints: CostPerformanceSecurityReliabilityDeep understanding of platform stability, scalability, and automation
Professional Competencies
Senior‑level autonomy and decision‑making capabilityOwnership mindset; accountable for outcomes rather than tasksAbility to operate as a trusted senior technical interface across teams