Site Reliability Engineer

Tech Mahindra — Ireland · Posted ~4 hours ago

Senior Full-time Hybrid Visa History ✓ 70000 EUR per year

Skills

Site Reliability Engineering Cloud infrastructure Operations Automation SRE

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A site reliability engineering role focused on maintaining enterprise-scale technology platforms, improving operational stability, and implementing modern cloud and automation practices.

Highlights

Senior-level reliability engineering role with hybrid work model, enterprise technology exposure, and opportunities across cloud and infrastructure services.

Description

About Us: Tech Mahindra offers technology consulting and digital solutions to global enterprises across industries, enabling transformative scale at unparalleled speed. With 149k+ professionals across 90+ countries helping 1100+ clients, TechM provides a full spectrum of services including consulting, information technology, enterprise applications, business process services, engineering services, network services, customer experience & design services, AI & analytics, and cloud & infrastructure services. It is the first Indian company in the world to have been awarded the Sustainable Markets Initiative’s Terra Carta Seal, in recognition of actively leading the charge to create a climate and nature-positive future. Tech Mahindra (NSE: TECHM) is part of the Mahindra Group, founded in 1945, one of the largest and most admired multinational federations of companies. Job Details: Experience: 7+ years Salary: 70000 EUR Per Annum Work model: 3 days per week work from client office Job Description: Ultimately, the role of SRE is to align Product and Customer Focused priorities with Operational needs. We regularly review our run state not only from an internal perspective, but also understanding and providing the feedback loop to our development partners on how we can improve the customer experience of our applications. Engage in and improve the whole lifecycle of services—from inception and design, through deployment, operations, and refinement.Analyze ITSM activities of the platform and provide feedback loop to development teams on operational gaps or resiliency concernsSupport services before they go live through activities such as system design consulting, capacity planning and launch reviews.Maintain services once they are live by measuring and monitoring availability, latency and overall system health.Scale systems sustainably through mechanisms like automation and evolve systems by pushing for changes that improve reliability and velocity.Support the application CI/CD pipeline for promoting software into higher environments through validation and operational gating, and lead DevOps automation and best practices.Practice sustainable incident response and blameless postmortems.Take a holistic approach to problem solving, by connecting the dots during a production event thru the various technology stack that makes up the platform, to optimize mean time to recoverCollaborate with a global team spread across tech hubs in multiple geographies and time zonesShare knowledge and mentor junior resources.Develop and maintain automation pipelines for certificate renewal, traffic routing, alerting, and compliance reporting using tools like Ansible, Venafi.Drive improvements in ITSM and DQ SLOs, ensuring timely CRQ status updates and incident closure.Lead initiatives for Safety & Soundness and Operational Excellence across quarterly EPICs, covering areas such as PCI compliance, threat/toil management, self-healing, and ITSM defect resolution All About You: Background in operational resiliency and self-healing systems.Understanding of two factor authentication.Strong documentation and communication skills.Strong in Unix/LinuxIntermediate understanding of Active Directory (Users / Groups), SAML, LTPA, SSO, Oauth.Understanding of DEVOPS technologies like Chef, Jenkins, Groovy, shell scripting, bitbucket, GIT.Experience in working with or implementing automation workflows and/or scripting development.Understanding of:Client-server relationshipsNetwork concepts (Layer 1 to Layer 3)Stack trace analysis (TCP dumps, heap dumps, CPU/memory analysis, thread dumps).Load balancers and application firewalls.Operating System navigation.Logging and monitoring methods, standards, and tools.High availability and business continuity planningCaching conceptsConfiguration managementAwareness of security implementations, certificate management lifecycle, mutual TLS, SSL handshake, SSH keys, symmetric and asymmetric encryptions.o Experience with AWS infrastructure and secure access practices.Familiarity with ITSM processes, compliance frameworks, and incident management.Excellent communication and collaboration skills across cross-functional teams How To Apply: It's easy to apply online; you just need a copy of your up-to-date CV and to follow the step-by step process. Don't worry if you need to make changes - you'll have the opportunity to review and edit your work on the final page, or you can also share resume directly to provided email address. We look forward to receiving your application! Tech Mahindra is an Equal Employment Opportunity employer. We promote and support a diverse workforce at all levels of the company. All qualified applicants will receive consideration for employment without regard to race, religion, color, sex, age, national origin or disability. All applicants will be evaluated solely on the basis of their ability, competence, and performance of the essential functions of their positions.