Summary
✨ AI‑Generated
Join an AI engineering team building production-grade intelligent systems with LLMs, embeddings, vector databases, and cloud-native infrastructure. You will manage Kubernetes clusters, deploy and tune vector databases, build Python AI applications and pipelines, containerize services, and integrate cloud storage and other backend technologies.
Highlights
Build and scale production-grade AI systems using modern LLM, vector database, Kubernetes, and cloud-native technologies, with substantial hands-on infrastructure and application engineering.
Description
We are seeking a highly skilled Senior AI Engineer with strong expertise in Kubernetes and vector database technologies to join our team.
In this role, you will design, build, and scale production-grade AI systems, working with cutting-edge LLM frameworks, embeddings, and cloud-native infrastructure to deliver robust and high-performance solutions.
Responsibilities
Deploy and manage Milvus vector databases, including schema design and index tuning (HNSW, IVF-FLAT)Build and maintain embedding and LLM pipelines using OpenAI API, Hugging Face, or CohereManage Kubernetes clusters, Helm charts, and containerized microservices in productionDevelop and maintain Docker containerization workflows, including multi-stage builds and registry managementDesign and deliver production-grade Python applications, integrating with Go, Java, or C++ where requiredIntegrate object storage systems such as AWS S3, MinIO, or Google Cloud StorageEvaluate and implement alternative vector database solutions, including Qdrant, Pinecone, and WeaviateCollaborate cross-functionally with team members to deliver reliable, scalable AI servicesEnsure operational excellence, observability, and performance of deployed AI workloads
Requirements
Bachelor's degree in Engineering with 5+ years of relevant experienceExpertise in Milvus deployment, schema design, and index tuning (HNSW, IVF-FLAT)Familiarity with vector database alternatives such as Qdrant, Pinecone, Weaviate, PGVector, or ChromaProficiency in building embedding and LLM pipelines using OpenAI API, Hugging Face, or CohereSkills in Kubernetes cluster management, Helm charts, and containerized microservicesBackground in Docker containerization, multi-stage builds, and registry managementProduction-level Python development along with Go, Java, or C++Knowledge of object storage integration, including AWS S3, MinIO, or Google Cloud StorageExcellent verbal and written communication skills with strong team collaboration abilitiesProficiency in English at an Upper-Intermediate level (B2) or higher
Nice to have
Experience supporting large-scale RAG applications and multi-agent platformsHands-on familiarity with LangChain, LlamaIndex, or custom pipelinesUnderstanding of GPU scheduling, resource optimization, and inference accelerationProduction experience with hybrid search, metadata filtering, and index tuningImplementation of LLM evaluation, governance, tracing, and monitoring toolsFamiliarity with CI/CD pipelines, Infrastructure-as-Code, and cloud-native deployment practicesPrior work experience in the Oil and Gas industryExperience with Dataiku DSSKnowledge of SRE practices
We offer
We connect like-minded people:Delivering innovative solutions to industry leaders, making a global impactEnjoyable working environment, whether it is the vibrant office or the comfort of your own homeOpportunity to work abroad for up to two months per yearRelocation opportunities within our offices in 55+ countriesCorporate and social eventsWe invest in your growth:Leadership development, career advising, soft skills and well-being programsCertifications, including GCP, Azure and AWSUnlimited access to EPAM's internal learning databaseFree English classes with certified teachersWe cover it all:Monetary bonuses for engaging in the referral programMedical & family care packageSix trust days per year (sick leave without a medical certificate)Coverage of psychology sessions of your choiceDiscounts for fitness clubs and sports programsBenefits package (sports activities, a variety of stores and services)
EPAM is global leader in AI transformation engineering and integrated consulting, serving Forbes Global 2000 companies and ambitious startups.
With over thirty years of expertise in custom software, product and platform engineering, we empower our clients to become AI-Native enterprises, driving measurable value from innovation and digital investments.
Experience the freedom of remote work from anywhere in Kyrgyzstan, whether it's the comfort of your home or our modern office in Bishkek.