Description
About Arca
Arca is a talent agency connecting exceptional Filipino professionals with ambitious U.S.
companies.
For this opportunity, Arca’s client is looking for a DevOps Engineer who can make cloud infrastructure and software delivery reliable, secure, and easier to operate.
The opportunity
You will work with developers and other engineering teammates to improve how software is built, released, monitored, and recovered when something fails.
Your work will span existing production environments and the infrastructure improvements needed as the product grows.
This is a hands-on role for someone who can follow a production problem through to its cause, make a thoughtful change, and verify the result.
You care about dependable systems, understandable automation, and helping the team deliver with confidence.
What you will own
• Build and improve the path from a reviewed code change to a healthy production release, including test gates, configuration, deployment verification, and rollback.
• Provision and maintain cloud resources through version-controlled infrastructure code, with reusable patterns and clear change reviews.
• Keep development, staging, and production environments consistent enough for teams to find problems before users do.
• Operate containerized services and their supporting networking, identity, storage, and managed infrastructure.
• Create useful dashboards and alerts; connect system symptoms to the logs, metrics, and traces needed to diagnose them.
• Investigate infrastructure and deployment failures, restore service with the team, and turn incident findings into lasting improvements.
• Apply practical security controls to infrastructure and delivery workflows, including limited access, protected secrets, patching, and appropriate scanning.
• Improve resilience through tested backups, recovery procedures, deployment safeguards, and sensible capacity planning.
• Identify avoidable cloud spending and performance bottlenecks, and explain the reliability and cost tradeoffs of proposed changes.
• Automate repetitive operational work and help developers use environments and release tools confidently.
• Keep runbooks, architecture notes, and handoff documentation current so critical knowledge is shared.
• Communicate changes, risks, progress, and incidents clearly to engineering teammates and other stakeholders.
Must-have requirements
• At least 3 years of professional experience in DevOps, cloud infrastructure, platform engineering, or site reliability, including direct responsibility for production systems.
• Strong hands-on experience with at least one major cloud platform: AWS, Microsoft Azure, or Google Cloud.
• Experience building and maintaining CI/CD pipelines with tools such as GitHub Actions, GitLab CI/CD, Azure DevOps Pipelines, or Jenkins.
• Practical infrastructure-as-code experience with Terraform, OpenTofu, CloudFormation, AWS CDK, Bicep, Pulumi, or an equivalent tool, using version control and reviewable changes.
• Experience building and operating Docker-based workloads and a container orchestration or managed container platform such as Kubernetes, ECS, or an equivalent service.
• Working knowledge of Linux administration, networking, DNS, TLS, load balancing, and troubleshooting application connectivity.
• Ability to automate operational tasks in Bash, Python, PowerShell, Go, or another suitable language.
• Experience using metrics, logs, alerts, and monitoring tools to investigate production issues and improve reliability.
• Practical understanding of access controls, least privilege, secrets management, patching, and secure deployment practices.
• Experience working with developers on releases, environment configuration, incident recovery, and clear operational documentation.
• Strong written and spoken English, sound judgment during incidents, and the ability to work independently with a distributed team.
Nice-to-have
• Deeper Kubernetes operations, Helm, or GitOps experience with Argo CD or Flux.
• Experience modernizing an existing production environment, migrating infrastructure, or improving an unreliable deployment process.
• Experience with distributed tracing, OpenTelemetry, service-level objectives, and reducing noisy alerts.
• Experience integrating dependency, container, infrastructure, or secret scanning into delivery pipelines.
• Hands-on work with database operations, backup restoration, disaster recovery exercises, or capacity and cloud-cost optimization.
• Experience supporting SaaS products, multiple environments or accounts, or a U.S.-based product team.
• Relevant cloud, Terraform, or Kubernetes certifications, or equivalent practical expertise.
• Experience using AI-assisted engineering tools while reviewing generated changes and validating their safety before deployment.
Who you are
You take responsibility for the systems you change and consider how those changes affect customers and teammates.
You investigate before guessing, make tradeoffs explicit, and keep others informed when reliability is at stake.
You value simple, maintainable solutions and leave useful documentation for the next person who needs to operate them.
Work arrangement
Full-time, remote opportunity for professionals based in the Philippines, intended to be 40 hours per week.
The specific cloud stack, working-hour overlap, and any on-call expectations will be discussed for the client engagement.
Why this role matters
Reliable infrastructure gives the product team room to build.
Your work will help reduce avoidable interruptions, make releases more predictable, and keep the platform ready for growth.
How to apply
Apply here: https://talent.arca.ph/apply/devops-engineer-linkedin