DevOps Engineer

Evlo Ai — United States · Posted ~11 hours ago

Skills

AWS or GCP Kubernetes Terraform CI/CD infrastructure automation observability incident response cloud infrastructure AWS GCP GitHub Actions GitLab CI ArgoCD Prometheus Grafana OpenTelemetry

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Own the infrastructure behind high-traffic production systems as a hands-on DevOps Engineer. You will design and maintain cloud infrastructure using AWS or GCP and Terraform, operate Kubernetes clusters, improve CI/CD pipelines, build observability systems, and participate in incident response. The role emphasizes reliable deployments, fast recovery, automation, and predictable infrastructure costs.

Highlights

Hands-on DevOps role owning production infrastructure, with substantial work in Kubernetes, Terraform, CI/CD, observability, incident response, cloud platforms, and infrastructure cost optimization.

Description

About The Role The role owns the infrastructure that keeps production running — Kubernetes clusters, CI/CD pipelines, observability stacks, and cloud environments serving high-traffic systems around the clock. The team works closely with application engineers to make deployments boring, incident response fast, and infrastructure costs predictable. This is a hands-on role: expect to spend the majority of your time in terminals, Terraform repos, and incident channels, not in slide decks. Key Responsibilities Design, build, and maintain scalable infrastructure on AWS or GCP using Terraform — VPCs, EKS/GKE clusters, IAM, networking, and autoscaling policiesOwn and improve CI/CD pipelines (GitHub Actions, GitLab CI, or ArgoCD) to enable safe, frequent deploys with automated rollback pathsBuild and operate observability stacks — Prometheus, Grafana, OpenTelemetry, and structured alerting — to reduce mean time to detection and resolutionLead and participate in incident response: run postmortems, identify root causes, and drive systemic fixes rather than patchesHarden production systems: implement security best practices, least-privilege access, secrets management (Vault), and compliance-ready audit trailsContainerize and orchestrate services with Docker and Kubernetes, including Helm chart management and workload right-sizingPartner with development teams to improve service reliability — defining SLOs, building runbooks, and eliminating operational toil through automation What We Are Looking For 3–6 years of experience in DevOps, SRE, or infrastructure engineering, including on-call ownership for production systemsDeep hands-on experience with at least one major cloud (AWS, GCP, or Azure), including networking and IAMProduction experience with Kubernetes and containerized workloads at meaningful scaleStrong infrastructure-as-code skills with Terraform (or equivalent) and version-controlled operational practicesProficiency in at least one scripting or systems language: Python, Go, or BashBachelor's degree in Computer Science, Engineering, or equivalent practical experience