Summary
✨ AI‑Generated
A technology organization is seeking a senior data engineer to design and maintain cloud-based data platforms, streaming and batch pipelines, governance solutions, and analytics infrastructure.
Highlights
Build scalable data platforms using modern cloud technologies and work on production-grade analytics solutions.
Description
About The Role
In this role, you will build scalable, efficient data solutions on the Databricks Lakehouse platform in Azure environments.
You will work on batch and streaming data pipelines, support data platform optimization initiatives, and collaborate with technical and business stakeholders to deliver reliable, production-ready data solutions across different stages of the project lifecycle.
Responsibilities
Design, develop, and maintain scalable batch and streaming ETL pipelines on the Databricks Lakehouse platformBuild and optimize data processing solutions using PySpark, Python, and SQLConfigure and maintain Databricks clusters and SQL WarehousesImplement and support data governance practices using Unity Catalog, including access control, lineage, and cross-workspace sharingDevelop and deploy data solutions using Databricks Asset Bundles and CI/CD processesDesign and maintain bronze–silver–gold Lakehouse architectures using Delta LakeExpose and visualize data using Databricks SQL and BI tools such as Power BICollaborate with technical and business stakeholders to understand requirements and deliver effective data solutionsParticipate in architecture discussions and contribute to continuous improvement initiatives within the teamSupport implementation and usage of modern Databricks AI/BI capabilities such as Dashboards, Metric Views, MLflow, Genie, and agents where applicable
Requirements
5+ years of experience in Data Engineering, including designing data models and scalable ETL pipelinesStrong expertise in PySpark, Python, and SQL, including performance tuning and optimizationHands-on experience with Databricks, including clusters and SQL WarehousesPractical knowledge of Unity Catalog governance, including fine-grained access control, lineage, and data sharingExperience delivering end-to-end data solutions using Databricks Asset Bundles and CI/CD practicesStrong understanding of Delta Lake and Lakehouse architecture conceptsExperience working with Databricks SQL and BI tools such as Power BIStrong communication and collaboration skills with technical and business stakeholdersUpper-intermediate or higher level of EnglishFamiliarity with Databricks AI/BI tools such as MLflow, Dashboards, Metric Views, Genie, or agents (is a plus)Experience with Azure data services such as ADF, Synapse, ADLS, Azure DevOps, or Microsoft Fabric (nice to have)
SoftServe is an equal opportunity employer.
Qualified applicants will receive consideration regardless of race, color, ancestry, ethnicity, national origin, religion, sex, sexual orientation, gender identity or expression, age, citizenship, disability, health condition, marital or family status, veteran status, or any other characteristic protected by applicable law.