Summary
✨ AI‑Generated
An experienced cloud infrastructure engineer is needed to design, operate, and improve secure cloud environments for a large-scale software platform. The role involves automation, infrastructure management, and reliability improvements.
Highlights
Work on secure and scalable cloud infrastructure with ownership of production environments and opportunities to improve automation and reliability practices.
Description
Overview
Role:
The Cloud Infrastructure Engineer is an experienced individual contributor responsible for designing, implementing, and operating secure, scalable, and highly available AWS infrastructure supporting a production-grade, multi-account B2B SaaS platform in the financial services industry.
This role is critical to maintaining platform resilience, security, automation maturity, and operational excellence across production and non-production environments.
The engineer operates with moderate autonomy, owning defined areas of the AWS environment while contributing to infrastructure standards, reliability engineering, and security best practices.
Responsibilities
Design, build, and maintain AWS infrastructure across production and non-production environmentsSupport application deployments and underlying platform services required for reliable operationImplement and manage infrastructure using Infrastructure as Code (Terraform)Configure and maintain AWS services including compute, storage, networking, and managed servicesMonitor infrastructure health, availability, and performance; respond to incidents and perform root cause analysisPartner with application engineering and security teams to support system reliability and scalabilityImplement and maintain backup, disaster recovery, and high-availability solutionsSupport automation related to infrastructure provisioning and application deploymentsManage IAM roles, permissions, and access controls in alignment with security policiesContribute to cloud cost monitoring and optimization effortsMaintain documentation, runbooks, and operational proceduresParticipate in on-call or incident response rotations as required
Core AWS & Cloud Technologies - Compute & Containers
Amazon ECS (Elastic Container Service)Amazon EC2Amazon ECR (Elastic Container Registry)AWS Lambda (event-driven workloads)
Networking & Edge
Amazon VPC (multi-AZ design, subnet segmentation)Internet Gateway (IGW)VPC PeeringElastic Load Balancing (ALB)AWS WAFAmazon Route 53AWS Data Transfer optimization
Databases & Caching
Amazon RDS (PostgreSQL) and PSQLAmazon DocumentDB (MongoDB compatibility)Amazon ElastiCache (Redis)
Messaging & Streaming
Amazon MSK (Managed Streaming for Apache Kafka)
Security & Compliance
AWS IAMAWS KMSAWS Secrets ManagerAWS ConfigAWS CloudTrailAmazon GuardDutyAmazon InspectorAWS WAFAuth0
Monitoring & Observability
Amazon CloudWatch (metrics, logs, alarms, events)
CI/CD & DevOps
CircleCIBitBucketGithubInfrastructure as Code (Terraform)
Storage & Data Lifecycle
Amazon S3Amazon S3 Glacier Deep Archive
Managed Platform Services
AWS Transfer FamilyAmazon WorkSpaces (administrative environments)
Skills and Experience
Strong hands-on experience operating production AWS environments in a multi-account architecture.Deep experience with containerized platforms (ECS) and distributed system design.Proven experience managing relational and document databases in high-availability configurations.Strong understanding of secure VPC architecture, ingress/egress control, and network isolation strategies.Experience operating Kafka-based streaming architectures (MSK).Expertise implementing infrastructure as code using Terraform.Experience building and maintaining CI/CD pipelines for containerized workloads.Strong knowledge of IAM design, encryption practices, and secrets management.Experience implementing monitoring, alerting, and log aggregation for production systems.Ability to troubleshoot performance, networking, scaling, and availability issues across distributed systems.Experience supporting regulated or compliance-sensitive environments (e.g., financial services).Demonstrated ability to leverage AI-assisted development tools (e.g., Cursor, ChatGPT) to accelerate infrastructure automation, troubleshooting, documentation, and code development while maintaining security and compliance standards.Strong judgment in the responsible use of AI tools to enhance productivity, improve code quality, and streamline operational workflows within regulated environments.Strong collaboration, documentation, and communication skills.
Minimum Qualifications
Bachelor’s degree in Computer Science, Engineering, or related field, or equivalent practical experience.4–6 years of experience in cloud infrastructure, platform engineering, or systems engineering roles.AWS certification (e.g., AWS Certified Solutions Architect – Associate, SysOps Administrator, or equivalent) preferred.
Integrity and Ethics
All StarCompliance employees are expected to commit to a high standard of personal integrity and carry out their responsibilities in an ethical manner