Data Ingestion Engineer
Pro Plus Pop Tv โ Slovenia ยท Posted ~2 hours ago
๐ Log in to save this job, tailor your resume & track your apply process โ 7 days free, no card needed.
Log in to add to target listDescription
About the job
Data Ingestion Engineer (m/f)
We are looking for a Data Ingestion Engineer to build and operate the ingestion layer of a modern data platform.
You will work on integrating source systems, streaming and batch pipelines, source-to-raw data flows, data contracts, and ingestion quality controls to ensure data arrives reliably, accurately, and on time.
What you will do:
* Design, build, and maintain integrations for structured, semi-structured, and event-based data
* Develop and operate Kafka-based streaming pipelines for near-real-time data ingestion
* Build and maintain batch ingestion workflows using Logstash, object storage, Parquet, Python, and orchestration tools
* Implement and support MySQL replica synchronization and other source-aligned extraction patterns
* Ensure reliable, stable, observable, and scalable source-to-raw data flows
* Define and maintain source-to-raw and infrastructure-level data contracts, including schema, freshness, and service-level expectations
* Implement ingestion-layer data quality controls, including schema, completeness, freshness, and anomaly checks
* Monitor ingestion SLAs, throughput, latency, failures, and overall pipeline health
* Support cloud-based ingestion tooling, storage services, and scheduler-orchestrated jobs
* Document source integrations, ingestion logic, runbooks, and incident-handling procedures
* Collaborate with platform, infrastructure, analytics, and downstream engineering teams
What you bring:
* Hands-on experience building and operating data ingestion pipelines in both streaming and batch environments
* Strong knowledge of Kafka, including producer/consumer patterns, partitioning, offsets, and monitoring
* Experience with Logstash, object storage, and formats such as Parquet, JSON, and CSV
* Strong Python and SQL skills
* Experience with orchestration tools such as Airflow, including scheduling, retries, dependencies, and backfills
* Practical understanding of relational data sources such as MySQL and analytical or real-time stores such as ClickHouse
* Experience with cloud-based ingestion tooling, storage services, and data contract-oriented delivery
* Solid understanding of CDC, incremental loading, schema evolution, idempotent processing, and raw-layer architecture
* Experience with monitoring, observability, alerting, SLA tracking, and production support
* Familiarity with Git-based workflows and CI/CD practices
Nice to have:
* Experience with GCP-based data environments and services supporting ingestion, storage, and orchestration
* Exposure to Cloud Composer, cloud storage, and managed cloud ingestion frameworks
* Experience with ClickHouse or similar high-performance analytical stores
* Knowledge of infrastructure-level data contracts and platform-oriented ingestion service design
What we offer:
* Flexible working (hybrid/remote)
* Learning & development opportunities
* Strong engineering ownership and impact
* Modern cloud-native tech stack and high-impact projects
If you enjoy building reliable data pipelines and solving complex data engineering challenges, weโd love to hear from you.
We have 145,133 jobs that might be an even better fit for you
DontApply's real value goes far beyond a single job link or company name. Just upload your resume โ in under a minute we'll analyze all 145,133 jobs and tell you exactly which ones you should apply to right now.
Upload My Resume