Data Ingestion Engineer

Pro Plus Pop Tv โ€” Slovenia ยท Posted ~2 hours ago

๐Ÿ”“ Log in to save this job, tailor your resume & track your apply process โ€” 7 days free, no card needed.

Log in to add to target list

Description

About the job Data Ingestion Engineer (m/f) We are looking for a Data Ingestion Engineer to build and operate the ingestion layer of a modern data platform. You will work on integrating source systems, streaming and batch pipelines, source-to-raw data flows, data contracts, and ingestion quality controls to ensure data arrives reliably, accurately, and on time. What you will do: * Design, build, and maintain integrations for structured, semi-structured, and event-based data * Develop and operate Kafka-based streaming pipelines for near-real-time data ingestion * Build and maintain batch ingestion workflows using Logstash, object storage, Parquet, Python, and orchestration tools * Implement and support MySQL replica synchronization and other source-aligned extraction patterns * Ensure reliable, stable, observable, and scalable source-to-raw data flows * Define and maintain source-to-raw and infrastructure-level data contracts, including schema, freshness, and service-level expectations * Implement ingestion-layer data quality controls, including schema, completeness, freshness, and anomaly checks * Monitor ingestion SLAs, throughput, latency, failures, and overall pipeline health * Support cloud-based ingestion tooling, storage services, and scheduler-orchestrated jobs * Document source integrations, ingestion logic, runbooks, and incident-handling procedures * Collaborate with platform, infrastructure, analytics, and downstream engineering teams What you bring: * Hands-on experience building and operating data ingestion pipelines in both streaming and batch environments * Strong knowledge of Kafka, including producer/consumer patterns, partitioning, offsets, and monitoring * Experience with Logstash, object storage, and formats such as Parquet, JSON, and CSV * Strong Python and SQL skills * Experience with orchestration tools such as Airflow, including scheduling, retries, dependencies, and backfills * Practical understanding of relational data sources such as MySQL and analytical or real-time stores such as ClickHouse * Experience with cloud-based ingestion tooling, storage services, and data contract-oriented delivery * Solid understanding of CDC, incremental loading, schema evolution, idempotent processing, and raw-layer architecture * Experience with monitoring, observability, alerting, SLA tracking, and production support * Familiarity with Git-based workflows and CI/CD practices Nice to have: * Experience with GCP-based data environments and services supporting ingestion, storage, and orchestration * Exposure to Cloud Composer, cloud storage, and managed cloud ingestion frameworks * Experience with ClickHouse or similar high-performance analytical stores * Knowledge of infrastructure-level data contracts and platform-oriented ingestion service design What we offer: * Flexible working (hybrid/remote) * Learning & development opportunities * Strong engineering ownership and impact * Modern cloud-native tech stack and high-impact projects If you enjoy building reliable data pipelines and solving complex data engineering challenges, weโ€™d love to hear from you.