Summary
✨ AI‑Generated
Design, automate, and operate production infrastructure at scale with a strong focus on reliability and safe deployments. You will build AWS infrastructure with Terraform, operate Kubernetes clusters and containerized workloads, develop CI/CD automation, improve observability, and participate in incident response. The role involves close collaboration with software engineering and security teams across multiple environments.
Highlights
A hybrid DevOps role focused on highly available cloud infrastructure, Kubernetes platform operations, automation, observability, and reliability. The position provides broad ownership across development, staging, and production environments and emphasizes safer deployments and resilient systems.
Description
About The Role
The DevOps Engineer will design, automate, and operate the infrastructure that runs production services at scale.
The role spans AWS cloud architecture, Kubernetes platform operations, CI/CD, observability, and incident response, with a focus on making deployments safer and systems more resilient.
This role matters because reliability is a product requirement.
The engineer will partner with software developers and security teams to improve release velocity, reduce operational risk, and establish repeatable infrastructure practices across development, staging, and production environments.
The position is based in Seattle, WA with a hybrid work schedule.
Key Responsibilities
Build and maintain highly available AWS infrastructure using Terraform, including networking, IAM, compute, storage, and managed database servicesOperate Kubernetes clusters and containerized workloads, improving resource efficiency, deployment safety, and platform reliabilityDevelop and maintain CI/CD pipelines with tools such as GitHub Actions, GitLab CI, or Jenkins, including automated testing, security checks, progressive delivery, and rollback proceduresImplement observability across services using Prometheus, Grafana, Datadog, or equivalent tooling for metrics, logs, traces, dashboards, and actionable alertingAutomate operational workflows with Python, Go, or Bash to eliminate manual procedures and improve consistency across environmentsParticipate in incident response, root-cause analysis, and on-call rotations; document follow-up actions and drive reliability improvements to completion
What We Are Looking For
3–8 years of experience in DevOps, SRE, cloud infrastructure, or platform engineering, including responsibility for production systemsStrong hands-on experience with AWS services and infrastructure as code, particularly Terraform or an equivalent toolProduction experience operating Kubernetes, Docker, Helm, and Linux-based systemsProficiency building CI/CD pipelines and integrating automated testing, artifact management, secrets handling, and deployment controlsSolid understanding of networking, DNS, TLS, IAM, distributed systems, and common application security practicesBachelor’s degree in computer science, engineering, or a related technical field, or equivalent practical experienceBonus: experience with Go, Python, service mesh technologies, GitOps tools such as Argo CD, and compliance frameworks such as SOC 2 or ISO 27001