Senior Data Engineer

Wearehaystack — Germany · Posted ~3 hours ago

Senior

Skills

Data Engineering Python ETL/ELT FastAPI Apache Spark Scala Data Quality Data Lineage Observability CI/CD Automated Testing Schema Management

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A senior Data Engineer role responsible for designing and operating reliable data pipelines that support software products and integrations. The work combines Python, API services, distributed data processing, data quality, lineage, observability, automated testing, CI/CD, schema management, and cost optimization, while collaborating with product, AI, and architecture teams.

Highlights

Senior data engineering opportunity focused on building reliable, scalable data pipelines and data products, with strong engineering practices, modern AI-assisted workflows, and close collaboration across product, AI, and architecture teams.

Description

We're hiring on behalf of a Haystack partner! The Role • Design and operate data pipelines that power SaaS products and partner integrations, transforming curated data into reliable, up-to-date, and performant datasets. • Build and maintain robust ETL/ELT pipelines using Python with FastAPI services and Apache Spark/Scala workloads, focusing on data quality, lineage, observability, and cost optimization. • Collaborate closely with Product Management, AI Engineers, and Architects to design Data Products enabling new AI-native features and partner integrations. • Utilize AI tools and agents to accelerate pipeline development, schema migrations, and analyses, building reusable data engineering skills and workflows. • Apply strong engineering principles to data management, including automated testing, CI/CD, observability, schema management, and clear ownership of pipelines and services. What You'll Need • Several years of senior-level experience in Data Engineering with Python and a proven track record of operating production data pipelines in regulated or business-critical environments. • A degree in Computer Science, Data Engineering, Mathematics, Physics, or a comparable technical field with strong academic performance. • Excellent SQL skills and experience with modern data stacks (e.g., Airflow, dbt, Spark/PySpark, Kafka) and cloud data platforms, preferably AWS. • Proficiency with Infrastructure-as-Code (Terraform), GitOps (ArgoCD), version control, testing, and monitoring. • Demonstrated effective use of AI in software and data engineering, ideally with agentic tools, and experience designing AI pipelines, harnesses, and context engineering. • Fluency in both German and English, strong problem-solving skills, a proactive and innovative mindset, and an entrepreneurial spirit aligned with healthcare purposes. What's On Offer • Competitive salary • Flexible work arrangements with hybrid options. • Opportunities for professional development through internal and external training programs. • Access to health and wellness resources, including healthy meal options and fitness facilities. Apply via Haystack today!