Site Reliability Engineer

Fulcrum Digital — Norway · Posted ~2 hours ago

Mid Contract

Skills

Site reliability engineering DevOps Production environment management Application performance monitoring Incident response Automation Monitoring Alerting ITSM Troubleshooting SRE Application Performance Monitoring CI/CD

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A Site Reliability Engineer is sought for a 12-month contract with potential extension. The role covers production environment management, performance optimization, incident response, automated deployments, standardized monitoring and alerting, and continuous improvement throughout the service lifecycle. Strong problem-solving and cross-stack troubleshooting are central to the position.

Highlights

A 12-month SRE/DevOps contract with potential extension, focused on production reliability, automation, monitoring, incident reduction, deployment support, and continuous improvement across the full service lifecycle.

Description

SRE Devops 12 months Contract with potential extension Oslo- Norway The Role Plan, manage, and oversee all aspects of a Production EnvironmentDefine strategies for Application Performance Monitoring Optimization in the production environmentRespond to Incidents and improvise platform based on feedback and measure the reduction of incidents over time.Support deployment of code into multiple lower environments. Support current processes with an emphasis on automating everything as soon as possible.Design, develop and standardize Monitoring and alerting mechanisms for the supported applications.Take a holistic approach to problem solving by connecting the dots during a production event through the various technology stack that makes up the platform, to optimize mean time to recover.Engage in and improve the whole lifecycle of services—from inception and design, through deployment, operation and refinement.Analyze ITSM activities of the platform and provide a feedback loop to development teams on operational gaps or resiliency concerns.Support services before they go live through activities such as system design consulting, capacity planning and launch reviews.Support the application CI/CD pipeline for promoting software into higher environments through validation and operational gating, and lead in DevOps automation and best practices.Maintain services once they are live by measuring and monitoring availability, latency and overall system health.Scale systems sustainably through mechanisms like automation and evolving systems by pushing for changes that improve reliability and velocity.Work with a global team spread across tech hubs in multiple geographies and time zones.Ability to share knowledge and explain processes and procedures to others.Share knowledge and mentor junior resourcesAble to perform on-call duties on a rotational basis.Occasional off-hours work required.Candidate should have an inclination for Training and should be a good trainer and ready to mentor others Requirements Skills – Must Have LinuxShell ScriptingITIL / ITSMSQL - Basic / Good to haveApplication TroubleshootingAny Monitoring tool (Preferred Splunk/Dynatrace)Jenkins - CI/CDGroovy Scripting/YamlGit basic/bit bucketGood To Have Even Framework architectureAnsible/Chef (Basic)