Site Reliability Engineer

Thundersoft — Sweden · Posted ~3 hours ago

Mid

Skills

Site Reliability Engineering Operations and maintenance CI/CD DevOps Cloud operations Application deployment System stability Technical support Cloud platforms SRE

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Join an SRE team responsible for keeping application systems stable and supporting deployments across multiple cloud environments and regions. You will build and maintain CI/CD and DevOps capabilities, assist with application upgrades, maintain operational documentation, troubleshoot technical issues, and support development and testing environments. Experience with enterprise-scale operations is valued.

Highlights

Work on reliable systems across multiple cloud environments, contribute to CI/CD and DevOps initiatives, support deployments and upgrades, and collaborate with development teams. The role also provides exposure to enterprise-scale systems and international environments.

Description

Main responsibilities: 1. Ensure the stability of the application system and deploy it online in new regions; 2. Maintain the stability of daily systems across multiple cloud vendors; 3. Carry out CICD and DevOps construction; 4. Assist in the deployment and upgrade of application systems to ensure a smooth transition. 5. Maintain operation and maintenance documents, record system changes and operation manuals. 6. Provide technical support to solve technical problems raised by customers and internal teams. 7. Collaborate with the development team to support the development and testing environment of application systems. Job requirements: Educational background: 1. College degree or above in computer science, information technology, software engineering or related majors (equivalent overseas education). hands-on background: At least 3 years of SRE operation and maintenance experience. 2. Experience in SRE operations and maintenance for large enterprises or multinational corporations is preferred. Certification requirements: 1. Familiar with ITIL (Information Technology Infrastructure Library) framework and best practices for application operation and maintenance. 2. Holders of relevant technical certifications (such as RHCE, MCSE, AWS, Azure, CKA, SRE Foundation/Professional, etc.) are preferred. Skill requirements: 1. Proficient in Kubernetes (k8s) container orchestration, containerization technology (such as Docker), and deployment, operation, monitoring, and troubleshooting of microservice architecture. 2. Proficient in relevant CI/CD toolchains (such as Jenkins, GitLab CI, Argo CD, etc.) and DevOps practices. 3. Proficient in Linux/Unix system management, familiar with common application operation and maintenance monitoring tools (such as Prometheus, Grafana, ELK/Loki, etc.) and automated configuration management tools (such as Ansible, Terraform, etc.) are preferred. 4. Have basic programming skills, at least one general-purpose programming or scripting language (such as Python, Go, Shell, etc.). 5. Familiar with the core services of cloud platforms such as AWS, GCP, Azure, etc., with experience in cloud environment operation and maintenance. 6. Experience in managing databases such as MySQL, PostgreSQL, Redis, etc. is preferred. 7. Good communication and coordination skills, as well as teamwork spirit, enable efficient collaboration with development teams, testing teams, and other technical departments to jointly ensure system stability and promote continuous improvement. Language requirements: 1. Fluent English communication skills, capable of conducting professional oral and written communication. 2. Chinese or other foreign language proficiency is preferred (such as Spanish, Malay, German, Arabic, etc.).