Description
About Company
PortPro builds DrayOS, the leading transportation management system for drayage and intermodal trucking.
Carriers across the US and Europe run their dispatch, driver pay, billing, and tracking on our platform every day,
so reliability, security, and cost efficiency directly affect our customers' businesses.
About Role
You will join a small DevOps team that owns the AWS platform end to end: infrastructure as code, CI/CD, databases,
observability, security compliance, and cloud cost.
This is a hands-on role.
You will triage production incidents, ship
Terraform changes, work with developers to fix performance problems in their code, and find ways to run the
platform for less money.
Key Responsibilities
Infrastructure and Platform
• Build and maintain infrastructure in Terraform/OpenTofu across a multi-account AWS Organization
(production, DR, dev, test, shared services, security)
• Run containerized workloads on ECS Fargate across multiple tenant deployments in the US and EU
• Manage ALB, CloudFront, WAF, VPC networking, NAT, and VPC endpoints
• Maintain backup and disaster recovery (AWS Backup, cross-region and cross-account)
CI/CD and Releases
• Maintain GitHub Actions pipelines that build, scan, and deploy services to ECS
• Support the daily release process across multiple application repositories
• Add security gates to pipelines (image scanning, license compliance, dependency checks)
Databases and Data Stores
• Operate Aurora PostgreSQL, MongoDB Atlas, TimescaleDB, Redis, and Amazon MQ (RabbitMQ)
• Triage database alerts such as high CPU, lock contention, connection exhaustion, and slow queries, and trace the
load back to the service and code path that caused it
• Plan capacity, indexing, maintenance, and data archival with development teams
Observability and Incident Response
• Own CloudWatch alarms, dashboards, and Logs Insights, plus New Relic APM
• Participate in the on-call rotation and write clear post-incident reports
• Keep alerting useful by tuning noisy alarms and closing detection gaps
Cost optimization(Fin0ps)
• Track AWS spend with Cost Explorer and CUR and investigate cost spikes
• Right-size Fargate and RDS, manage Savings Plans and Reserved Instances, and remove waste
• Present savings opportunities with clear annual dollar impact
Security and Compliance
• Support SOC 1 and SOC 2 audits by gathering evidence and keeping controls effective
• Manage IAM Identity Center (SSO), least-privilege roles, and access-key hygiene
• Track and remediate vulnerabilities using Amazon Inspector, GuardDuty, and Trivy
Required Qualifications
• 3+ years of hands-on AWS experience in production (ECS or EKS, RDS, VPC, IAM, CloudWatch, S3)
• Strong Terraform skills, including modules, remote state, and safely changing existing codebases
• Experience building and maintaining CI/CD pipelines (GitHub Actions preferred)
• Solid Linux and Docker fundamentals; comfortable scripting in Bash and Python
• Working knowledge of PostgreSQL: reading query plans and identifying blocking sessions
• Real production incident experience: you have debugged an outage under pressure and found the root cause
• Clear written English for incident notes, pull requests, and reports
Preferred Qualifications
• MongoDB Atlas, TimescaleDB, Redis, or RabbitMQ operations
• AWS cost optimization experience (Savings Plans, Reserved Instances, right-sizing, CUR/Athena)
• SOC 2, SOC 1, or ISO 27001 audit experience
• Node.js or Python application background
• New Relic, OpenTelemetry, or Prometheus/Grafana
• AWS certifications (Solutions Architect, DevOps Engineer Professional)
• Comfortable using AI coding tools as part of daily work
How we work
• All infrastructure changes go through pull requests, not the AWS console
• Production access is read-only by default; write access is scoped and audited
• We measure impact: uptime, time to resolve incidents, and dollars saved
• Small team, high ownership, and your work reaches production quickly
What we offer
• Paid time off and holidays
• Learning budget and paid AWS certification exams
• Work on a high-traffic, multi-region SaaS platform
Hiring process
1.
Application review
2.
Intro call with HR (30 min)
3.
Technical interview with the DevOps team (60 min): AWS, Terraform, and a live troubleshooting scenario
4.
Practical exercise: a small Terraform and CI task (2-3 hours)
5.
Final interview with engineering leadership
When applying, include a short note about an incident you debugged or a cost saving you delivered.