Senior Data Engineer

Wearehaystack — Germany · Posted ~10 hours ago

Senior

Skills

Data Engineering Python ETL/ELT Apache Spark Scala Data quality Data lineage Observability CI/CD Automated testing Schema management FastAPI

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A senior data engineering opportunity focused on designing and operating reliable, high-performance data pipelines that transform curated data into production-ready datasets. You will work with Python, FastAPI, Apache Spark, and Scala while emphasizing data quality, lineage, observability, automated testing, CI/CD, schema management, and cost efficiency. The role involves close collaboration with product, AI, and architecture teams to create data products and accelerate development through AI-assisted engineering workflows.

Highlights

Senior-level data engineering role focused on building reliable, scalable data pipelines and data products that support AI-driven capabilities and integrations. Offers opportunities to work cross-functionally with product, AI, and architecture teams while applying modern engineering practices and AI-assisted development workflows.

Description

We're hiring on behalf of a Haystack partner! The Role • Design and operate data pipelines that power SaaS products and partner integrations, transforming curated data into reliable, up-to-date, and performant datasets. • Build and maintain robust ETL/ELT pipelines using Python with FastAPI services, and Apache Spark/Scala workloads, focusing on data quality, lineage, observability, and cost optimization. • Collaborate closely with Product Management, AI Engineers, and Architects to design Data Products enabling new AI-native features and partner integrations. • Utilize AI tools and agents to accelerate pipeline development, schema migrations, and analyses, building reusable data engineering skills and workflows. • Apply strong engineering principles to data management, including automated testing, CI/CD, observability, schema management, and clear ownership of pipelines and services. What You'll Need • Several years of senior-level experience in Data Engineering with Python and a proven track record of operating production data pipelines in regulated or business-critical environments. • A degree in Computer Science, Data Engineering, Mathematics, Physics, or a comparable technical field with excellent academic performance. • Strong SQL skills and experience with modern data stacks (e.g., Airflow, dbt, Spark/PySpark, Kafka) and cloud data platforms, ideally AWS. • Proficiency with Infrastructure-as-Code using Terraform and GitOps/ArgoCD, along with version control, testing, and monitoring. • Demonstrated effective use of AI in software and data engineering, ideally with agentic tools, and experience designing AI pipelines, harnesses, and context engineering. • Fluency in both German and English, strong problem-solving skills, innovative mindset, high proactivity, and an entrepreneurial spirit aligned with healthcare's purpose and patient safety. What's On Offer • Competitive salary • Flexible work arrangements with hybrid options. • Opportunities for professional development through internal and external training programs. • Access to wellness facilities, including a company canteen with healthy options and a fully equipped fitness studio. Apply via Haystack today!