DevOps Engineer

Thedeveloperportugal — United States · Posted ~1 day ago

Onsite

Skills

Cloud infrastructure Infrastructure as Code (IaC) Terraform CI/CD Cloud automation Monitoring and observability Security best practices Multi-cloud environments High availability Fault tolerance Least-privilege security Containerization Containers Monitoring Observability

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary

A DevOps Engineer is sought to help operate and evolve secure, highly available cloud infrastructure for a mission-critical technology platform. The role focuses on automating multi-cloud environments with Infrastructure as Code, building scalable CI/CD workflows, improving monitoring and observability, applying strong security and least-privilege practices, managing containerized delivery environments, and optimizing infrastructure costs. The position is based on the U.S. East Coast.

Highlights

Work on a mission-critical platform with a strong focus on scalability, security, reliability, and high availability. The role offers hands-on ownership of cloud infrastructure, automation, observability, and modern CI/CD practices, with opportunities to optimize cloud costs and improve software delivery.

Description

Your Mission As a DevOps Engineer, you will play a crucial role in ensuring the scalability, security, and reliability of our AI-driven financial crime prevention platform. You will automate cloud infrastructure, implement monitoring and observability solutions, and build secure, scalable CI/CD pipelines. Your work will directly contribute to maintaining high availability for a platform that fights financial crime 24/7. This role is based on the East Coast, U.S. and requires expertise in cloud infrastructure, automation, security best practices, and continuous integration/deployment (CI/CD). Your Responsibilities Provision, manage, and scale multi-cloud environments using Infrastructure as Code (IaC) (e.g., Terraform).Maintain high availability (HA), fault tolerance, and least-privilege security practices, while optimizing cloud costs.Design and maintain developer-friendly CI/CD workflows, container templates, and reusable artifacts for seamless software delivery.Implement real-time monitoring, alerting, and observability solutions (e.g., Elastic Stack, Prometheus, Grafana, CloudWatch) to proactively detect and resolve issues.Implement and enforce cloud security best practices, identify and mitigate vulnerabilities, and ensure compliance with data protection regulations.Provide technical guidance to clients running Hawk’s platform in their own VPC environments, supporting onboarding and integration.Develop structured documentation for cloud architectures, best practices, and deployment processes, ensuring seamless team collaboration. Your Profile 5+ years of experience in DevOps, Site Reliability Engineering (SRE), or Cloud Engineering roles.Bachelor’s degree in Computer Science, Engineering, or a related field (or equivalent experience).Strong expertise in Kubernetes, containerized applications, and cloud-native technologies.Hands-on experience with AWS or GCP and their core services.Proficiency with Terraform and Infrastructure as Code (IaC) methodologies.Experience with CI/CD tools such as GitLab CI, GitHub Actions, or similar.Strong knowledge of observability and monitoring tools (e.g., Elastic Stack, Prometheus, Grafana, CloudWatch).Solid understanding of cloud security principles, least-privilege access, and automated security policies.Ability to diagnose complex technical challenges and provide scalable, secure solutions.Strong communication and collaboration skills; able to work effectively in a remote, cross-functional environment.Comfortable in a fast-paced, hands-on role, with a willingness to get your hands dirty and embrace feedback for continuous improvement. Preferred Qualifications Experience in cybersecurity, penetration testing, and cloud compliance.Familiarity with Java Spring Boot & Apache Kafka.Experience in 24/7 uptime environments with on-call rotations.Knowledge of big data systems (PostgreSQL, S3/Azure Blob Storage, Elasticsearch).