Description
About Us
Our client is a software development company delivering effective digital solutions to clients around the world, with a strong focus on the U.S.
market.
Headquartered in sunny San Diego, California, we’ve built a team of over 300 talented professionals across multiple countries.
Fueled by courage and curiosity, we go beyond providing services; we dive deep into our clients’ businesses to uncover real needs and tackle key challenges, combining human insight with technology.
Our success comes from a deep commitment to results and long-term partnerships built on trust.
About the Role
We continue our strategic partnership with a long-standing client in the game development space – a company operating a rapidly scaling, cross-platform social gaming ecosystem used by millions of players worldwide.
The product brings together casual games and real-time social interaction within a single, seamlessly integrated platform.
It combines competitive gameplay with chat, matchmaking, and community-driven features, creating a highly engaging and retentionfocused user experience.
Growth is fully organic, driven by product excellence rather than paid acquisition or advertising.
We are looking for an experienced Senior DevOps Engineer to join our infrastructure team working on a large-scale Terraform-based infrastructure for a gaming platform.
The infrastructure manages multiple environments (dev, stage, prod) across AWS, supporting Kubernetes clusters, analytics pipelines, and high-traffic gaming services.
What We Offer
• Work on a large-scale, production infrastructure serving millions of users
• Opportunity to solve complex performance and scalability challenges
• Modern tech stack with best practices
• Impact on critical infrastructure decisions
• Collaborative team environment
• Professional growth opportunities
Technology Stack
Core Infrastructure
• Terraform v1.10.5 (Infrastructure as Code)
• AWS Services: EKS, RDS (PostgreSQL/Aurora), Lambda, S3, VPC, ALB/NLB, CloudWatch, IAM, KMS, SSM
• Kubernetes: EKS clusters with Helm charts, Envoy ingress, service mesh, Karpenter for node autoscaling
• Container Runtime: Containerd for container management
• Monitoring: DataDog (APM, logs, metrics, dashboards)
• CI/CD: AWS CodeBuild, CodePipeline, GitHub Actions
Analytics & Data Platform
• Snowflake: Data warehouse with multiple databases and schemas
• MWAA (Apache Airflow): Data pipeline orchestration
• S3: Data lake and analytics storage
• Redis Cloud: Caching layer (migration to AWS ElastiCache planned)
Additional Services
• CloudFlare: CDN and DNS management
• PagerDuty: Incident management
• VPN: OpenVPN infrastructure across 2 regions
Project Structure
The infrastructure is organized into:
• Environments: dev, stage, prod, analytics, etc.
• Modules: 50+ reusable Terraform modules for common patterns
• Globals: Shared configurations and remote state management
• Helm Charts: Kubernetes application deployments
• Utils: Custom tooling (tftool) for Terraform operations
Key Responsibilities
• Maintain and evolve Terraform-based infrastructure across multiple environments (dev, stage, prod)
• Manage Kubernetes clusters, workloads, and container orchestration
• Develop and optimize analytics platform infrastructure (MWAA, Snowflake, data pipelines)
• Implement security best practices, access controls, and compliance requirements
• Optimize system performance, latency, and resource utilization
• Design and implement monitoring, observability, and alerting solutions
• Troubleshoot production issues and ensure system reliability
Requirements
• 5+ years of DevOps/Infrastructure engineering experience
• Strong Terraform expertise - Advanced knowledge of Terraform modules, state management, remote state
• AWS proficiency - Deep understanding of EKS, RDS, Lambda, VPC, IAM, networking
• Kubernetes experience - EKS cluster management, Helm, service mesh, ingress controllers
• Monitoring & Observability - DataDog, CloudWatch, APM, metrics, dashboards
• Linux/Unix systems - Shell scripting, system administration
• Git/GitHub - Version control, CI/CD workflows
• English proficiency - Technical communication
Nice to Have
• Snowflake experience - Data warehouse management, MFA setup, API authentication
• Apache Airflow/MWAA - Data pipeline orchestration
• Karpenter - Kubernetes node autoscaling and spot instance management
• Go programming - For custom tooling (tftool utilities)
• Network optimization - Global Accelerator, CDN, latency optimization
• Security - IAM best practices, OIDC, MFA, secrets management
• Cost optimization - AWS cost analysis and optimization, spot instance strategies
Why Choose Us
• We’re a company where curiosity fuels everything - from bold ideas to personal growth.
• We believe common sense drives smart decisions.
It keeps us focused and helps move fast, without the noise.
• We trust each other to deliver.
Real ownership, no micromanagement.
• You’ll be heard here, even when you challenge the status quo.
Because courage matters.
Who dares — wins.
Join us!