Lead Backend Distributed Systems Engineer

Epam Systems — Colombia · Posted ~5 hours ago

Lead Full-time

Skills

Backend engineering Distributed systems Rust Scalable services System design Cloud infrastructure

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A global technology organization is seeking a lead engineer to design and operate scalable distributed backend systems. The role includes architecture ownership, reliability improvements, and cross-team technical leadership.

Highlights

Lead engineering role focused on high-performance distributed systems, global collaboration, and advanced backend architecture.

Description

EPAM is a leading global provider of digital platform engineering and development services. We are committed to having a positive impact on our customers, our employees, and our communities. We embrace a dynamic and inclusive culture. Here you will collaborate with multi-national teams, contribute to a myriad of innovative projects that deliver the most creative and cutting-edge solutions, and have an opportunity to continuously learn and grow. No matter where you are located, you will join a dedicated, creative, and diverse community that will help you discover your fullest potential. We are seeking a Lead Backend/Distributed Systems Engineer to build and operate high-performance, scalable services that integrate with distributed storage systems. You will lead design and implementation work, improve reliability through troubleshooting and on-call support, and collaborate across global time zones—apply now. Responsibilities Design scalable, high-availability backend services that interact with distributed storage systemsBuild high-performance Rust components and core service functionalityLead system design for distributed storage features including consistency, sharding, replication, and fault toleranceImplement RPC-based service interfaces using gRPC or Apache ThriftDevelop concurrent and asynchronous execution models for reliability and throughputTroubleshoot production issues, debug complex failures, and drive root-cause resolutionProvide on-call support and improve operational readiness through runbooks and diagnosticsCoordinate with global teams and support occasional early morning or late evening overlap meetingsIntegrate cloud-native capabilities such as service discovery, service mesh, and IPv6 where required Requirements 5+ years of backend software engineering experience with RustStrong leadership skills to guide technical decisions and mentor engineersProven project experience designing or operating distributed storage systems or componentsAdvanced concurrent programming expertise with async, multithreading, and concurrency patternsStrong RPC systems experience using gRPC or Apache Thrift for service-to-service communicationCloud-native architecture experience with service discovery tools such as Consul or ZooKeeperEnglish proficiency level B2 (Upper-Intermediate) for daily collaboration and documentation Nice to have Go LanguagePythonLinuxSystem design and analysisgRPC We offer International projects with top brandsWork with global teams of highly skilled, diverse peersHealthcare benefitsEmployee financial programsPaid time off and sick leaveUpskilling, reskilling and certification coursesUnlimited access to the LinkedIn Learning library and 22,000+ coursesGlobal career opportunitiesVolunteer and community involvement opportunitiesEPAM Employee GroupsAward-winning culture recognized by Glassdoor, Newsweek and LinkedIn EPAM is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, age, sexual orientation, gender identity or expression, disability, protected veteran status, or any other characteristic protected by applicable law.