Senior Machine Learning Engineer

Auxotalent — Poland · Posted ~3 hours ago

Senior Contract Remote

Skills

Python machine learning Apache Spark SQL data pipelines distributed systems Spark Airflow DBT

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A senior machine learning engineering contract role focused on scalable data processing, AI models, search capabilities, and high-volume data systems.

Highlights

Contract role working on large-scale AI and data platforms with modern machine learning technologies.

Description

Senior Machine Learning Engineer - Warsaw Location: Warsaw/remote Contract: 12 months Pay rate: Up to 450 euros per day My client is a leading global consulting and technology organisation seeking a Senior Machine Learning Engineer to build and scale a high-impact entity resolution platform that leverages AI, machine learning, search technologies, and distributed systems to process hundreds of millions of records. Key Responsibilities: Design and operate large-scale data processing systems using Python, Spark, and SQL.Develop entity matching, deduplication, clustering, and search capabilities.Build and optimise data pipelines using Spark, Airflow, and DBT.Improve matching accuracy, recall, performance, and scalability.Develop analytical data models and optimise large-scale SQL workloads.Implement validation frameworks and monitor data quality metrics. Requirements: 5+ years of software engineering experience.Strong Python development skills.Hands-on production experience with Apache Spark / Databricks.Strong SQL skills and experience with large-scale datasets.Solid understanding of distributed systems and data-intensive applications.Experience with similarity matching, search, entity resolution, or related domains.Ability to balance accuracy, performance, scalability, and cost. Desirable Skills: Snowflake or similar cloud data warehouse platforms.Airflow, DBT, or equivalent orchestration tools.Entity resolution, record linkage, or deduplication systems.Embeddings, vector search, semantic matching, or ML-assisted pipelines.Elasticsearch, OpenSearch, or vector databases.Azure, AWS, Kubernetes, serverless technologies, and GitHub Actions. If this opportunity is of interest to you, send in your application!