Summary
✨ AI‑Generated
A senior data engineering opportunity focused on designing and operating reliable, high-performance data pipelines that transform curated data into production-ready datasets. You will work with Python, FastAPI, Apache Spark, and Scala while emphasizing data quality, lineage, observability, automated testing, CI/CD, schema management, and cost efficiency. The role involves close collaboration with product, AI, and architecture teams to create data products and accelerate development through AI-assisted engineering workflows.
Highlights
Senior-level data engineering role focused on building reliable, scalable data pipelines and data products that support AI-driven capabilities and integrations. Offers opportunities to work cross-functionally with product, AI, and architecture teams while applying modern engineering practices and AI-assisted development workflows.
Description
We're hiring on behalf of a Haystack partner!
The Role
• Design and operate data pipelines that power SaaS products and partner integrations, transforming curated data into reliable, up-to-date, and performant datasets.
• Build and maintain robust ETL/ELT pipelines using Python with FastAPI services, and Apache Spark/Scala workloads, focusing on data quality, lineage, observability, and cost optimization.
• Collaborate closely with Product Management, AI Engineers, and Architects to design Data Products enabling new AI-native features and partner integrations.
• Utilize AI tools and agents to accelerate pipeline development, schema migrations, and analyses, building reusable data engineering skills and workflows.
• Apply strong engineering principles to data management, including automated testing, CI/CD, observability, schema management, and clear ownership of pipelines and services.
What You'll Need
• Several years of senior-level experience in Data Engineering with Python and a proven track record of operating production data pipelines in regulated or business-critical environments.
• A degree in Computer Science, Data Engineering, Mathematics, Physics, or a comparable technical field with excellent academic performance.
• Strong SQL skills and experience with modern data stacks (e.g., Airflow, dbt, Spark/PySpark, Kafka) and cloud data platforms, ideally AWS.
• Proficiency with Infrastructure-as-Code using Terraform and GitOps/ArgoCD, along with version control, testing, and monitoring.
• Demonstrated effective use of AI in software and data engineering, ideally with agentic tools, and experience designing AI pipelines, harnesses, and context engineering.
• Fluency in both German and English, strong problem-solving skills, innovative mindset, high proactivity, and an entrepreneurial spirit aligned with healthcare's purpose and patient safety.
What's On Offer
• Competitive salary
• Flexible work arrangements with hybrid options.
• Opportunities for professional development through internal and external training programs.
• Access to wellness facilities, including a company canteen with healthy options and a fully equipped fitness studio.
Apply via Haystack today!