Summary
✨ AI‑Generated
An experienced DevOps role responsible for designing and maintaining scalable cloud infrastructure and reliable delivery pipelines. You will automate infrastructure, manage containers, improve observability and security, troubleshoot production systems, and collaborate with engineering stakeholders to optimize performance and availability.
Highlights
Experienced DevOps opportunity covering multi-cloud infrastructure, automation, Kubernetes, infrastructure as code, observability, security, and production reliability, with strong exposure to Agile collaboration.
Description
Requirements:
Minimum 5 years of experience as DevOps Engineer, Site Reliability Engineer (SRE), Cloud Engineer, Platform Engineer, or related role.Strong understanding of Software Development Life Cycle (SDLC) and DevOps practices.Hands-on experience with CI/CD pipelines and automation tools.Experience managing cloud platforms such as AWS, Azure, or Google Cloud Platform.Strong knowledge of Linux/Unix operating systems.Experience with containerization technologies such as Docker and Kubernetes.Familiarity with Infrastructure as Code (IaC) tools such as Terraform, Ansible, or similar.Experience with monitoring and logging tools such as Grafana, Prometheus, ELK Stack, Datadog, or similar.Strong knowledge of networking, security best practices, and system performance optimization.Experience working in Agile/Scrum environments.Good troubleshooting, analytical, and problem-solving skills.Strong communication and stakeholder management skills.Responsibilities:
Design, implement, and maintain CI/CD pipelines to support application deployment and release processes.Manage and optimize cloud infrastructure to ensure scalability, reliability, and performance.Automate infrastructure provisioning, configuration management, and operational processes.Monitor system health, application performance, and infrastructure availability.Troubleshoot and resolve production issues, incidents, and performance bottlenecks.Collaborate with Development, QA, Security, and Infrastructure teams to improve deployment efficiency and system reliability.Implement and maintain containerized environments using Docker and Kubernetes.Ensure infrastructure and applications comply with security standards and best practices.Develop and maintain monitoring, alerting, and logging solutions.Support disaster recovery planning, backup strategies, and business continuity initiatives.Participate in architecture discussions and provide recommendations for system scalability and operational excellence.