Summary
Join a collaborative engineering team to improve the reliability, scalability, and automation of cloud infrastructure. Work with modern cloud technologies, infrastructure as code, incident response, and monitoring while contributing to highly available production systems.
Highlights
Contract opportunity with hybrid flexibility, remote work within the country, professional development, certifications, English classes, and comprehensive employee benefits.
Description
We are seeking a skilled Site Reliability Engineer to join our team on a contract basis.
As an SRE, you will work to ensure our systems, services and applications running on Google Cloud Platform (GCP) are reliable, performant and scalable.
The ideal candidate will have strong technical skills, be passionate about automation and infrastructure-as-code, and work well in a collaborative team environment.
Please note that the working hours are standard for candidates from Poland, with the ability to adjust to evening calls, a few times per week, up to 6-7 pm.
Responsibilities
Participation in on-call rotations to cover 24/7 support for critical systemsResponse to alerts of running services and applications, conducting RCADeployment of microservices according to release cadenceDesign, implementation and maintenance of scalable and reliable systems and applications on Google Cloud Platform (GCP)Development and maintenance of infrastructure as code using TerraformCollaboration with engineering teams to identify and prioritize reliability, performance improvements and rightsizing of the dedicated cloud resourcesInvolvement in incident management and response using ServiceNowManagement and resolution of technical issues and tickets using JiraDevelopment of knowledge base for maintaining existing infrastructure and monitoring services
Requirements
3+ years of experience in an SRE, DevOps or system administration roleKnowledge of Google Cloud Platform (GCP)Expertise in Linux operating system internals coupled with the ability to diagnose and resolve complex system-level problemsUnderstanding of containerization concepts and toolsExperience with incident management and response using ServiceNow or similar toolsStrong problem-solving skills and experience with debugging complex technical issuesUnderstanding of monitoring, logging and alerting systems, preferably Cloud MonitoringFamiliarity with version control using GitHubExperience with infrastructure-as-codeExcellent communication and collaboration skillsEnglish proficiency at B2 level or higher
Nice to have
Experience with Kubernetes and containerization technologiesExperience with Terraform for infrastructure-as-codeStrong understanding of SDLC and CI/CD pipelines and experience with CI/CD tools
We offer
We gather like-minded people:Top tech minds driving innovation in AI, cloud and digital platform modernizationSupportive team and agile, startup-like cultureHybrid by design mode and opportunity to work remotely within PolandChance to work abroad for up to 60 days annuallyBusiness-driven relocation opportunitiesWe provide growth opportunities:Career development programsThought leadership, mentoring, soft skills and well-being programsCertification (Anthropic, Gemini, GCP, Azure, AWS)English classesWe cover it all:Stable payParticipation in the Employee Stock Purchase Plan with a 15% discountBenefits package (health insurance, multisport, shopping vouchers)Referral bonuses up to $2,000Offices featuring entertainment and relaxation zones, table tennis and football, free snacks, coffee and moreCorporate, social and well-being eventsPlease, note:Benefits listed above are available to employees onlyWe are open for working with Contractors.
Terms of B2B cooperation agreements are agreed individuallyWe will reach out to selected candidates exclusively
EPAM is global leader in AI transformation engineering and integrated consulting, serving Forbes Global 2000 companies and ambitious startups.
With over thirty years of expertise in custom software, product and platform engineering, we empower our clients to become AI-Native enterprises, driving measurable value from innovation and digital investments.