DevOps / SRE Engineer – AI Cloud
Trulyyy — Singapore · Posted ~1 hour ago
🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.
Log in to add to target listDescription
We are hiring for a fast-growing global technology company headquartered in Singapore, currently expanding its AI Cloud and GPU infrastructure platform.
What you’ll do:
Support day-to-day operations of Linux and Kubernetes environmentsMaintain CMDB / asset management processes and perform routine platform health checksOperate internal platforms such as OpenBao / Vault, Sonatype and audit systemsMaintain monitoring and observability platforms including Prometheus, Grafana, OpenSearch and Alert managerSupport CVE remediation, patching, access control, account audits and security compliance activitiesSupport K8s-based AI workloads including monitoring, RBAC, backup and resource managementAutomate operational tasks using Python / Shell
What We’re Looking For
3+ years of experience in DevOps, SRE, Cloud Operations, or System EngineeringHands-on experience with Linux and Kubernetes, including deployment, scaling, monitoring, and basic troubleshootingExperience with Prometheus / Grafana or similar monitoring and observability toolsGood scripting skills in Python and/or ShellBasic understanding of CI/CD, networking, RBAC, and infrastructure securityStrong troubleshooting skills and a good sense of ownership and operational disciplineExperience with OpenBao / Vault, CMDB, CVE remediation, or SOC2 is a plus
We have 141,734 jobs that might be an even better fit for you
DontApply's real value goes far beyond a single job link or company name. Just upload your resume — in under a minute we'll analyze all 141,734 jobs and tell you exactly which ones you should apply to right now.
Upload My Resume