Senior Cloud DevOps Engineer

Reqroute — United States · Posted ~2 hours ago

Senior Full-time Onsite

Skills

AWS Kubernetes Terraform GitOps CI/CD Linux Argo CD Helm GitHub Actions

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A senior cloud infrastructure role responsible for designing and operating scalable cloud environments. Requires deep expertise in cloud platforms, infrastructure as code, Kubernetes, and deployment automation.

Highlights

Work on secure and scalable cloud infrastructure with modern DevOps practices, automation, and container orchestration technologies.

Description

Senior Cloud DevOps Engineer Location: Onsite – Lowell, MA or Troy, MI Position Overview We are seeking a highly experienced Senior Cloud DevOps Engineer to design, implement, and support scalable, secure, and resilient cloud infrastructure. The role will focus heavily on AWS, Kubernetes, Terraform, GitOps, and CI/CD, with responsibility for infrastructure automation, cloud services, reliability, and continuous improvement. The ideal candidate will have strong hands-on experience with AWS, Kubernetes, Terraform, Argo CD/Helm, Linux, DataDog, and GitHub/GitHub Actions, along with deep knowledge of cloud networking and infrastructure automation. Key Responsibilities Design, implement, and manage cloud infrastructure using AWS, Kubernetes, and Terraform.Develop reusable Terraform modules, manage remote state, and implement infrastructure policy-as-code.Implement GitOps practices using Argo CD or Flux for declarative, version-controlled infrastructure and application deployments.Manage Kubernetes platforms, including EKS, multi-cluster lifecycle management, upgrades, scaling, and workload optimization.Troubleshoot complex Kubernetes and cloud networking issues involving CNI plugins, ingress controllers, service mesh, and load balancing.Design, build, and maintain reliable CI/CD pipelines.Automate infrastructure provisioning and configuration using Terraform and tools such as Ansible, Puppet, or Chef.Develop and maintain infrastructure-as-code and configuration-as-code across environments.Monitor, troubleshoot, and optimize infrastructure for performance, availability, reliability, and cost.Configure and manage monitoring and observability solutions such as DataDog, Prometheus, Grafana, and ELK.Integrate third-party tools, plugins, and internal automation scripts.Create and maintain technical documentation, including runbooks, architecture diagrams, and operational procedures.Participate in on-call rotations and improve incident response and troubleshooting processes.Collaborate with application, engineering, security, and product teams to ensure infrastructure meets business and technical requirements.Participate in Agile ceremonies and work with tools such as Jira and Confluence.Required Qualifications Education Bachelor's degree in Computer Science, Computer Engineering, Software Engineering, Electrical Engineering, or equivalent experience.Experience 7+ years of experience in DevOps, SRE, Platform Engineering, or Infrastructure Engineering.5+ years of hands-on experience with cloud platforms, with strong expertise in AWS.Advanced Linux expertise, including systemd, cgroups, namespaces, OS tuning, troubleshooting, and production hardening.Strong hands-on knowledge of Kubernetes and cloud networking beyond the fundamentals.Experience managing Kubernetes/EKS clusters and Docker/containerized workloads.Strong expertise in Terraform and Infrastructure as Code (IaC).Proficiency in Python and Bash scripting.Experience with Git, GitHub, GitHub Actions, Jenkins, and artifact repositories.Strong understanding of networking concepts including DNS, DHCP, VPN, LDAP, VPCs, security groups, and IAM policies.Experience with monitoring and logging platforms such as DataDog, Prometheus, Grafana, ELK, or equivalent.Experience working with Agile development teams and tools such as Jira and Confluence.Must-Have Skills Candidates should have strong hands-on experience with: AWSKubernetesTerraformArgo CD / GitOpsHelmLinuxDataDogGitHub / GitHub ActionsCloud networking and AWS VPC/IAMPython and BashNice-to-Have Skills KafkaAnsibleRabbitMQ / EMQX / other messaging brokersExperience with Industrial IoT (IIoT) environmentsHybrid cloud and on-premises infrastructureHashiCorp Vault / AWS Secrets ManagerKubernetes management using Helm chartsService mesh technologiesExperience mentoring engineers or leading DevOps initiativesIdeal Candidate Profile The ideal candidate is a senior-level Cloud/DevOps/SRE/Platform Engineer with deep hands-on experience building and operating production environments using AWS + Kubernetes + Terraform + GitOps. Strong expertise in Linux, cloud networking, Kubernetes troubleshooting, infrastructure automation, CI/CD, and observability/DataDog is essential. Work Arrangement This is a 100% onsite position. Candidates must be able to work onsite at either: 📍 Lowell, MA 📍 Troy, MI