Senior Data Engineer

Wearehaystack — Germany · Posted ~9 hours ago

Senior Full-time

Skills

Data engineering Python ETL/ELT Apache Spark Scala Data quality Data lineage Observability CI/CD Automated testing Schema management FastAPI

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A Senior Data Engineer is needed to design and operate robust data pipelines and services that transform curated data into reliable, performant datasets. The role combines Python, distributed data processing, ETL/ELT, data quality, lineage, observability, testing, CI/CD, and schema management. You will collaborate with product, AI, and architecture teams while using modern AI-assisted engineering workflows.

Highlights

Build and operate reliable data products and pipelines supporting AI-native features and integrations, with a strong focus on engineering quality, observability, automation, and scalability.

Description

We're hiring on behalf of a Haystack partner! The Role • Design and operate data pipelines that power SaaS products and partner integrations, transforming curated data into reliable, up-to-date, and performant datasets. • Build and maintain robust ETL/ELT pipelines using Python with FastAPI services, and Apache Spark/Scala workloads, focusing on data quality, lineage, observability, and cost optimization. • Collaborate closely with Product Management, AI Engineers, and Architects to design Data Products enabling new AI-native features and partner integrations. • Utilize AI tools and agents to accelerate pipeline development, schema migrations, and analyses, building reusable data engineering skills and workflows. • Apply strong engineering principles to data management, including automated testing, CI/CD, observability, schema management, and clear ownership of pipelines and services. What You'll Need • Several years of senior-level experience in Data Engineering with Python and a proven track record of operating production data pipelines in regulated or business-critical environments. • A degree in Computer Science, Data Engineering, Mathematics, Physics, or a comparable technical field with excellent academic performance. • Strong SQL skills and experience with modern data stacks (e.g., Airflow, dbt, Spark/PySpark, Kafka) and cloud data platforms, ideally AWS. • Proficiency with Infrastructure-as-Code using Terraform and GitOps/ArgoCD, along with version control, testing, and monitoring. • Demonstrated effective use of AI in software and data engineering, ideally with agentic tools, and experience designing AI pipelines, harnesses, and context engineering. • Fluency in both German and English, strong problem-solving skills, innovative mindset, high proactivity, and an entrepreneurial spirit aligned with healthcare's purpose and patient safety. What's On Offer • Competitive salary • Flexible work arrangements with hybrid options. • Opportunities for professional development through internal and external training programs. • Access to wellness facilities, including a company canteen with healthy options and a fully equipped fitness studio. Apply via Haystack today!