Senior Data Engineer

Hashlist — United States · Posted ~2 hours ago

Senior Contract Hybrid

Skills

Data pipelines Apache Spark Distributed systems Cloud infrastructure CI/CD Terraform GitHub Actions

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A senior data engineering position focused on designing scalable pipelines, distributed processing systems, and reliable cloud-based data platforms.

Highlights

High-impact role working with large-scale data systems, modern engineering practices, and advanced software-driven programs.

Description

We're looking for a Senior Data Engineer to design, build, and optimise robust data pipelines that process massive datasets in both batch and real-time. Working at the intersection of software engineering and data science, you'll ensure our data architecture is scalable, reliable, and follows industry best practices - building systems that power B2B data products at global scale. Engagement detail: Location: Auburn HillsContract type: Permanent/ContractStart date: ASAPWork model: HybridBenefits: Competitive package; high-impact, high-visibility role within a major software-defined vehicle programme Responsibilities: Design and implement complex data processing pipelines using Apache SparkBuild scalable, distributed systems for high-throughput data streams and large-scale batch processingDesign and implement real-time pipelines with strong data quality and validation controlsManage and provision cloud infrastructure using TerraformImplement CI/CD workflows using GitHub Actions for automated testing and deploymentUphold rigorous engineering standards — unit/integration testing, code reviews, and clean documentationTranslate business requirements into technical specifications alongside cross-functional stakeholders Qualifications: 5+ years in data engineering and the software development lifecycle4+ years hands-on with AWS cloud services4+ years building production data applications across relational and columnar data storesDeep expertise in Apache Spark — Core, SQL, and Structured StreamingStrong proficiency in Scala or Java for JVM-based production developmentAdvanced SQL for transformation, analysis, and performance tuningHands-on experience with Terraform and GitHub ActionsExperience with Apache Flink for low-latency stream processingProficiency in Python for automation and scriptingFamiliarity with Kafka, event-driven architectures, and streaming dataExperience with workflow engines — Airflow, Luigi, or Azure Data FactoryFamiliarity with Lakehouse architectures (Delta Lake, Iceberg) or NoSQL databasesDegree in Computer Science, Engineering, Mathematics, or related field preferred Next Steps Press "Apply"We will review your applicationIf qualified, you will be accepted into the network and can be considered for this and similar positions