Summary
✨ AI‑Generated
A junior reliability engineering role focused on monitoring services, troubleshooting issues, and improving deployment and operational workflows. Ideal for building practical experience in production engineering.
Highlights
Entry-level opportunity to develop SRE expertise while working on reliable production systems and improving operational processes.
Description
About the Role
We are currently looking for a Junior Site Reliability Engineer to help keep our services reliable for the people who use them every day.
You’ll work alongside our engineering team to monitor production systems, investigate issues, and improve how we deploy and operate our services.
This is an opportunity to build your SRE skills while contributing to Flip’s mission of providing fair financial services to all Indonesians.
About Flip
Flip was founded in 2015 by three friends from Universitas Indonesia — Rafi, Luqman, and Anjar — with the goal of making money transfers fairer for Indonesians.
Today, millions of individuals and businesses trust Flip to send trillions of rupiah across banks every year.
We’re backed by leading investors such as Sequoia India, Insight Partners, and Insignia, and our mission is simple: to give Indonesians access to one of the most progressive and fairest financial services in the world.
At Flip, we always strive to provide the fairest place for you to work, learn, and grow with talented and fun people in various opportunities to advance your career and get fair rewards.
We believe that we have to treat employees, customers, and all stakeholders fairly and respectfully.
Fair treatment for employees means we establish clear goals, facilitate our employees to achieve them, and value their contribution to the company with equitable benefits.
What you'll do
Help monitor and maintain Flip’s production systems and cloud infrastructure to keep our services reliable.Support the team in investigating incidents, troubleshooting issues, and documenting findings and follow-up actions.Assist in improving monitoring and alerting to help the team detect issues earlier.Work with engineers to improve the reliability, performance, and scalability of our services.Help automate recurring operational tasks and improve infrastructure and deployment processes.Maintain technical documentation, runbooks, and operational procedures.Learn and apply SRE practices, including incident management, service metrics, and security best practices, with guidance from the team.
What you'll need
Up to 2 years of experience in SRE, DevOps, infrastructure, software engineering, or a related role.
Relevant internship or project experience is welcome.A degree in Computer Science or a related field, or equivalent practical experience.Basic understanding of Linux systems, networking, and how web applications run in production.Familiarity with at least one cloud platform, such as Alibaba Cloud, GCP, or AWS.Ability to write scripts or code in at least one language, such as Python, Go, JavaScript, or Bash.Familiarity with Git and the basics of CI/CD, with some exposure to container technologies such as Docker or Kubernetes.Interest in troubleshooting, automation, monitoring, and building reliable systems.Good problem-solving skills, willingness to learn, and ability to collaborate with other teams.