NOC Engineer - AWS

Netrolynx Ai — United Kingdom · Posted ~4 hours ago

Mid Full-time Remote

Skills

AWS Kubernetes cloud platforms network operations incident troubleshooting 24/7 operations cloud infrastructure

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A remote operations engineering role focused on maintaining reliable cloud infrastructure. The position involves monitoring systems, resolving incidents, supporting Kubernetes-based environments, and working in a rotating 24/7 operations schedule.

Highlights

Fully remote role working on resilient cloud platforms supporting essential services, with exposure to AWS, Kubernetes, and large-scale operational environments.

Description

About The Company Spectrum IT Recruitment is a leading technology recruitment agency specializing in connecting talented IT professionals with innovative organizations across various industries. With a strong reputation for excellence and a commitment to understanding client needs, Spectrum IT Recruitment prides itself on delivering tailored staffing solutions that drive business success. The company values integrity, professionalism, and a deep understanding of the evolving tech landscape, ensuring that both clients and candidates receive exceptional service and support throughout the recruitment process. About The Role The NOC Engineer with AWS Kubernetes expertise is a critical role within Spectrum IT Recruitment's client organization, which is dedicated to building resilient cloud platforms that support essential national services. This position is fully remote, based in the UK, and involves working on a 24/7 shift pattern, including days and nights on a 28-day rota. The successful candidate will be part of an engineering-led team focused on maintaining high availability, automation, and continuous improvement of cloud infrastructure. This role offers a unique opportunity to work on complex, large-scale production environments, contributing to the prevention of incidents through proactive system enhancements and automation. The role extends beyond traditional network operations, emphasizing reliability engineering, automation, and operational excellence. The engineer will collaborate with various teams, including Software, Platform, Cloud, and Security Engineers, to ensure the stability and resilience of cloud services. The position is ideal for individuals passionate about cloud infrastructure, automation, and solving complex technical challenges in a dynamic environment. Qualifications The ideal candidate will have experience in a Production Engineering, Cloud Operations, or NOC environment, with a strong background in the following areas: Linux systems administration and troubleshootingAWS cloud infrastructure management and supportKubernetes and Docker container orchestrationProduction support and incident management experienceScripting skills in Python, Bash, or GoExperience with monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk, or CloudWatchFundamental networking knowledge including DNS, TCP/IP, and load balancingA passion for automation, continuous improvement, and operational excellence Additional beneficial experience includes Infrastructure as Code (Terraform), Site Reliability Engineering (SLIs, SLOs), or working within regulated environments. However, these are not mandatory. Responsibilities The NOC Engineer will be responsible for maintaining and supporting highly available production platforms in AWS, including: Monitoring and ensuring the health and performance of cloud-based production environmentsResponding to and managing incidents in a 24/7 operational setting to minimize downtimeInvestigating complex technical issues, diagnosing root causes, and restoring services efficientlyDeveloping automation scripts and tools to reduce manual operational tasks and enhance platform resilienceBuilding and refining monitoring, alerting, and observability solutions across cloud environmentsCollaborating with cross-functional teams to improve system reliability and operational processesParticipating in post-incident reviews to identify lessons learned and implement continuous service improvementsSupporting containerized workloads using Kubernetes and Docker, ensuring optimal deployment and management Benefits Joining Spectrum IT Recruitment's client offers a comprehensive benefits package, including competitive salary, performance-based bonuses, and excellent benefits that support your well-being and professional growth. The role provides the flexibility of a fully remote working environment, allowing for a healthy work-life balance. You will have the opportunity to work within a forward-thinking, engineering-led organization that values innovation, continuous learning, and operational excellence. The company fosters a collaborative culture, encourages professional development, and provides opportunities to work on cutting-edge cloud technologies and automation initiatives. Equal Opportunity Spectrum IT Recruitment is committed to promoting diversity and inclusion within the workplace. We are an equal opportunity employer and welcome applications from all qualified candidates regardless of race, gender, age, disability, religion, or sexual orientation. We believe that a diverse workforce enhances our ability to serve our clients effectively and fosters a positive, innovative work environment. All employment decisions are made based on merit, qualifications, and business needs.