Summary
✨ AI‑Generated
A technology company is seeking a DevOps engineer to automate infrastructure, manage cloud platforms, improve reliability, and support production operations.
Highlights
Build scalable cloud infrastructure and automation systems supporting high-availability production environments.
Description
About The Role
The role owns the infrastructure that keeps production systems running: CI/CD pipelines, cloud architecture, observability, and incident response.
The team operates a Kubernetes-based platform on AWS serving high-traffic services where downtime is measured in dollars, not minutes.
You will work alongside backend engineers and SREs to automate everything from build pipelines to disaster recovery, turning manual operations into reliable, self-healing systems.
Key Responsibilities
Build and maintain CI/CD pipelines using GitHub Actions, Terraform, and ArgoCD to enable safe, frequent deployments across microservicesOperate and scale Kubernetes clusters on AWS (EKS), managing node pools, autoscaling policies, and multi-AZ resilienceDesign observability stacks with Prometheus, Grafana, and OpenTelemetry — dashboards, SLOs, and alerting that catch issues before customers doLead incident response for production outages: triage, mitigation, root cause analysis, and blameless postmortems with actionable follow-upsHarden infrastructure security: IAM least-privilege policies, secrets management with Vault, network segmentation, and audit loggingOptimize cloud costs across compute, storage, and networking — rightsize workloads and eliminate waste without sacrificing reliabilityAutomate repetitive operational work with Python and Bash, progressively eliminating toil across the platform
What We Are Looking For
3–6 years of experience in DevOps, SRE, or infrastructure engineering, including at least 2 years running production KubernetesDeep hands-on expertise with AWS (EKS, EC2, RDS, VPC, IAM) and infrastructure-as-code with TerraformStrong scripting and automation skills in Python, Bash, or GoPractical experience with CI/CD tooling (GitHub Actions, GitLab CI, or ArgoCD) and GitOps workflowsSolid grasp of networking fundamentals: DNS, TLS, load balancing, and service mesh conceptsBachelor's degree in Computer Science, Engineering, or equivalent practical experienceBonus: Experience with Istio or Envoy, chaos engineering practices, compliance environments (SOC 2, PCI), or on-call leadership at scale