DevOps Engineer

Evlo Ai — United States · Posted ~2 hours ago

Mid

Skills

AWS Terraform Kubernetes CI/CD Cloud infrastructure Infrastructure automation Observability Production operations GitHub Actions GitLab CI Jenkins Containerization VPC IAM ECS EKS RDS S3 CloudWatch Helm

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A DevOps Engineer is sought to build and operate the infrastructure behind reliable, secure, and scalable production services. The role covers AWS infrastructure as code, Kubernetes operations, CI/CD automation, observability, deployment workflows, and continuous improvements to availability, latency, and incident response.

Highlights

Hands-on DevOps role focused on reliable, secure, scalable production infrastructure, with broad ownership across AWS, Kubernetes, CI/CD, observability, and deployment automation.

Description

About The Role The DevOps Engineer builds and operates the infrastructure that keeps production services reliable, secure, and scalable. The role spans AWS cloud environments, Kubernetes clusters, CI/CD systems, observability, and the automation required to deploy and run services with minimal manual intervention. Working with software engineers, security, and platform teams, the role improves deployment velocity and operational resilience across the stack. Success means faster, safer releases, clear ownership of production systems, and measurable improvements in availability, latency, and incident response. Key Responsibilities Design and maintain AWS infrastructure using Terraform, including VPCs, IAM, ECS or EKS, RDS, S3, and CloudWatchBuild and improve CI/CD pipelines with GitHub Actions, GitLab CI, or Jenkins for automated testing, deployment, rollback, and environment promotionOperate Kubernetes workloads in production, including cluster configuration, Helm releases, autoscaling, networking, and resource managementImplement observability with Prometheus, Grafana, ELK or OpenSearch, and distributed tracing to identify performance and reliability issuesAutomate operational workflows with Python, Go, or Bash, reducing repetitive work across provisioning, releases, incident response, and access managementDefine reliability practices including health checks, alerting, runbooks, backup validation, disaster recovery procedures, and capacity planningParticipate in on-call rotations, lead technical incident response, and drive post-incident remediation through measurable engineering improvements What We Are Looking For 3–8 years of experience in DevOps, site reliability engineering, cloud infrastructure, or platform engineeringHands-on experience operating production workloads in AWS and managing infrastructure as code with Terraform or an equivalent toolStrong Kubernetes experience, including deployments, services, ingress, Helm, autoscaling, and troubleshooting containerized applicationsProficiency with Linux administration, networking fundamentals, Git, and at least one scripting or programming language such as Python, Go, or BashExperience building CI/CD pipelines and applying automated testing, secrets management, security controls, and safe deployment strategiesWorking knowledge of observability, incident management, high availability, disaster recovery, and common reliability metrics such as SLOs, SLIs, and error budgetsBachelor’s degree in computer science, information technology, engineering, or a related field; equivalent practical experience is accepted. Bonus: experience with Argo CD, service meshes, AWS certifications, compliance frameworks, or multi-region infrastructure