Description
Role Overview
We are looking for an experienced and highly skilled Senior DevOps / SRE Engineer to design, implement, automate, and manage modern DevOps and Site Reliability Engineering practices across enterprise technology environments.
The ideal candidate will have strong hands-on experience in CI/CD, GitHub Actions, Jenkins, container technologies, Infrastructure as Code (IaC), DevSecOps, monitoring, logging, automation, and cloud-native technologies.
The role requires someone who can work across development, infrastructure, security, and operations teams to improve deployment efficiency, system reliability, security, scalability, and operational excellence.
The successful candidate will be expected to take ownership of projects from inception through delivery and ongoing operational support, while continuously identifying opportunities to automate and improve existing processes.
Key Responsibilities
DevOps & CI/CD
Design, implement, maintain, and optimize CI/CD pipelines for enterprise applications and servicesDevelop and manage deployment pipelines using Jenkins and GitHub ActionsBuild and maintain Jenkins pipelines using:Jenkins Shared LibrariesDeclarative and Scripted PipelinesGroovy scriptingPython and Bash automationImplement automated build, test, security scanning, packaging, and deployment processesDevelop reusable CI/CD components and standards to improve consistency across development teamsSupport automated deployments across containerized and cloud-native environmentsIdentify and eliminate manual deployment activities through automationEstablish best practices around source control, branching, release management, artifact management, and deployment automationDevOps Tooling
Install, configure, administer, upgrade, integrate, and troubleshoot enterprise DevOps tooling, including:
JenkinsGitHub / GitHub ActionsNexus / Nexus IQSonarQubeCheckmarxSysdigCosignArgo CDJIRAConfluenceManage integrations between DevOps tools to establish an efficient and secure software delivery lifecycleMonitor tool availability, performance, capacity, and securityTroubleshoot issues related to CI/CD tools and their integrationsEstablish standards for tool configuration, access management, security, and governance
Containerization & Deployment
Implement and manage CI/CD pipelines for applications deployed on container technologiesSupport containerized application build, packaging, deployment, and lifecycle managementWork closely with development and infrastructure teams to improve container-based application deliveryImplement automated deployment strategies using modern DevOps and GitOps practicesSupport Argo CD-based GitOps workflows and ensure reliable application deploymentsTroubleshoot container, deployment, configuration, and runtime-related issues
Infrastructure Automation & Configuration Management
Automate infrastructure provisioning, deployment, and configuration management activitiesDevelop reusable automation to create consistent, scalable, and repeatable environmentsApply Infrastructure as Code principles to infrastructure and application configurationReduce operational dependency on manual activities through automationImplement configuration management standards and ensure consistency across environmentsSupport infrastructure changes through controlled, automated, and auditable processes
DevSecOps & Security
Integrate security controls into CI/CD pipelines following DevSecOps principlesImplement and maintain automated security and quality gates using tools such as:SonarQubeCheckmarxNexus IQSysdigCosignSupport vulnerability scanning, code quality analysis, dependency analysis, container security, and artifact verificationImplement secure software supply-chain practices, including artifact signing and verificationWork with security teams to address vulnerabilities and improve application and infrastructure securityEnsure DevOps processes comply with organizational security and governance standardsMonitoring, Logging & Observability
Implement and maintain enterprise logging, monitoring, and observability solutionsWork with technologies such as:ELK StackPrometheusGrafanaDevelop dashboards, alerts, and monitoring mechanisms for infrastructure and application environmentsMonitor system health, application performance, availability, and reliabilityAnalyze logs and metrics to identify performance issues and operational risksEstablish proactive monitoring and alerting to reduce incidents and improve service reliabilitySupport root-cause analysis of production incidents using monitoring and observability dataIncident Management & Troubleshooting
Troubleshoot and resolve complex infrastructure, application, deployment, and CI/CD issuesWork closely with development, infrastructure, security, and operations teams to resolve production and non-production issuesParticipate in incident management, problem management, and root-cause analysisIdentify recurring issues and implement permanent solutions rather than relying on manual workaroundsSupport production deployments and provide operational assistance when requiredContribute to continuous improvement initiatives based on incident trends and operational feedback
SRE & Operational Excellence
Apply Site Reliability Engineering (SRE) principles to improve system availability, scalability, performance, and resilienceImplement automation to reduce operational toilDefine and improve operational processes, reliability practices, and service standardsSupport capacity planning, performance optimization, availability, and resilience initiativesContribute to reliability engineering practices such as monitoring, alerting, incident response, and continuous improvementBalance project delivery responsibilities with operational support requirements
DevOps & Application Lifecycle
Work on projects from inception, design, implementation, testing, deployment, and transition to operationsCollaborate with application development teams to integrate DevOps practices into the software development lifecycleProvide technical guidance on build, deployment, configuration, automation, and operational requirementsPromote automation-first approaches across development and operations teamsSupport continuous improvement of engineering processes and delivery methodologies
Required Technical Skills
Mandatory Skills
Strong hands-on experience with DevOps methodologies and practicesStrong practical experience with GitHub ActionsStrong hands-on experience with Jenkins and CI/CD pipelinesGood knowledge of Jenkins Shared Libraries, Groovy, Python, and/or Bash scriptingExperience managing and administering enterprise DevOps toolingStrong understanding of container technologies and container-based deploymentsExperience with GitOps and Argo CDStrong understanding of Infrastructure as Code (IaC) principlesStrong knowledge of DevSecOps and security integration within CI/CD pipelinesExperience with ELK, Prometheus, and Grafana or equivalent monitoring and logging technologiesStrong infrastructure and security fundamentalsStrong troubleshooting and problem-solving capabilities
DevOps Tools
Hands-on experience with several of the following:
Jenkins | GitHub | GitHub Actions | Nexus | Nexus IQ | SonarQube | Checkmarx | Sysdig | Cosign | Argo CD | JIRA | Confluence
DevOps / SRE Knowledge
The candidate should have a comprehensive understanding of:
DevOps principles and methodologiesSRE principles and practicesCI/CD and continuous deliveryInfrastructure as CodeGitOpsDevSecOpsContainerizationAutomation and configuration management12-Factor Application principlesMonitoring and observabilityLogging and incident managementApplication lifecycle managementInfrastructure and application securityReliability, scalability, and availability engineering
Soft Skills & Behavioral Competencies
Strong communication skills with the ability to communicate effectively with technical and non-technical stakeholdersAbility to establish transparent and professional relationships with internal teams, clients, vendors, and key stakeholdersStrong ownership and accountability for assigned projects and servicesAbility to work independently while collaborating effectively within cross-functional teamsStrong analytical and pragmatic approach to problem-solvingAbility to work effectively under pressure and manage multiple prioritiesStrong stakeholder and client management skillsAbility to manage client expectations while maintaining delivery commitmentsStrong focus on operational excellence and continuous improvementWillingness to learn and adopt emerging technologies and industry best practicesAbility to work across both project delivery and operational support functionsAbility to adapt to changing business and technology requirementsDemonstrate professional integrity, accountability, and a collaborative approachAbility to contribute positively to team culture and promote organizational values
Key Success Factors
The successful candidate will be expected to:
Increase deployment automation and reduce manual interventionImprove CI/CD pipeline reliability and efficiencyStrengthen security throughout the software delivery lifecycleImprove infrastructure and application reliabilityReduce recurring production incidents through automation and root-cause analysisImprove monitoring, logging, and observabilityEstablish scalable and reusable DevOps practicesSupport faster and more reliable software deliveryBuild strong relationships with development, infrastructure, security, operations, and business stakeholdersContinuously identify opportunities for process, tooling, and technology improvements