Summary
✨ AI‑Generated
An experienced DevOps Engineer is sought to design, automate, and operate scalable and secure cloud infrastructure. You will work hands-on with AWS, Kubernetes, Terraform, CI/CD, observability, and production systems while partnering with engineering and security teams to improve reliability and delivery speed.
Highlights
Hands-on opportunity to own scalable, secure, and highly reliable cloud infrastructure from design through production. The role emphasizes automation, deployment velocity, infrastructure reliability, system performance, and developer productivity.
Description
Location: New York City, NY
Experience: 5+ years
Compensation: $150,000–$250,000 per year
Employment Type: Full-Time
We are looking for an experienced DevOps Engineer to build, automate, and operate scalable, secure, and highly reliable cloud infrastructure.
The ideal candidate will have strong hands-on experience with AWS, Kubernetes, Terraform, CI/CD, observability, and production infrastructure.
You will work closely with software engineers, security, and product teams to improve deployment velocity, infrastructure reliability, system performance, and developer productivity.
This is a hands-on engineering role requiring ownership of infrastructure and automation from design through production.
Requirements
Key Responsibilities
Design, build, and maintain scalable and secure cloud infrastructure on AWS.
Develop and manage Infrastructure as Code using Terraform and related tooling.
Build, maintain, and optimize CI/CD pipelines for automated software delivery.
Deploy and operate containerized applications using Docker and Kubernetes, including EKS.
Implement GitOps and automated deployment practices to improve release reliability.
Establish monitoring, logging, alerting, and observability across production systems.
Define and improve reliability practices, including incident response, root-cause analysis, and disaster recovery.
Automate repetitive infrastructure and operational processes using Python, Bash, or similar scripting languages.
Improve infrastructure security through IAM, secrets management, network controls, and secure deployment practices.
Monitor infrastructure performance and cloud costs and identify opportunities for optimization.
Partner with engineering teams to improve developer tooling and deployment workflows.
Participate in architecture and infrastructure design discussions.
Create technical documentation, runbooks, and operational standards.
Mentor engineers and contribute to DevOps and infrastructure best practices.
Must-Have Skills
5+ years of hands-on experience in DevOps, SRE, Cloud Infrastructure, or Platform Engineering.
Strong production experience with AWS.
Strong hands-on experience with Kubernetes, preferably Amazon EKS.
Strong experience with Terraform / Infrastructure as Code.
Experience designing and maintaining CI/CD pipelines.
Strong understanding of Docker and containerized applications.
Experience with Linux systems administration and networking fundamentals.
Experience with monitoring and observability tools such as Prometheus, Grafana, Datadog, CloudWatch, or similar.
Strong scripting/programming skills in Python, Bash, Go, or similar.
Experience with Git and modern software development workflows.
Strong troubleshooting, problem-solving, and incident-management skills.
Good-to-Have Skills
Experience with Terraform Cloud, Terragrunt, or Pulumi.
Experience with Helm, Argo CD, Flux, or other GitOps tooling.
Experience with AWS services including EKS, ECS, EC2, S3, RDS, IAM, VPC, and CloudWatch.
Experience with Kafka or other distributed messaging systems.
Experience implementing SRE practices, SLIs, SLOs, and SLAs.
Experience with security, compliance, IAM, and secrets-management systems.
Experience optimizing AWS infrastructure and cloud spend.
Experience with service meshes or distributed systems.
Experience working in high-growth startups or SaaS/product companies.