Founding Software Engineer - Data Infrastructure

Stealth Startup Community โ€” United States ยท Posted ~1 day ago

Senior Full-time

Skills

Software architecture Data pipelines Backend development FastAPI PostgreSQL React TypeScript Large-scale systems

๐Ÿ”“ Log in to save this job, tailor your resume & track your apply process โ€” 7 days free, no card needed.

Log in to add to target list

Summary

An early-stage technology company is hiring a senior engineer to build a large-scale data delivery platform, own critical architecture decisions, and develop reliable backend services and user interfaces.

Highlights

High-impact founding engineering role with ownership over architecture, large-scale data systems, and end-to-end product delivery.

Description

Senior/Staff Software Engineer (E5/E6) โ€” Data Delivery Infrastructure About the role You will drive our end-to-end data delivery platform โ€” the system that turns raw robot data into customer-ready datasets and lands them in customer clouds. You'll shape the architecture, build the pipeline, operate real customer deliveries, and keep the system fast and cost-efficient at multi-terabyte scale. Delivery commitments to customers are hard deadlines, and you'll carry them from pipeline design through final handoff. What you'll do Build the delivery platform - Design and evolve the delivery pipeline end-to-end: dataset querying and selection, quality-based data preparation, conversion into customer-specified formats, and batched delivery with staged progress tracking. - Build and maintain the services behind it: the delivery orchestration service (FastAPI, Postgres/asyncpg), its integration with our internal data platform, and the operator-facing UI (React/TypeScript). - Drive compute orchestration on our Kubernetes-based job platform: job definition, container images, retries, reconciliation, and throughput. - Develop cross-cloud transfer infrastructure spanning multiple providers (AWS, Cloudflare, Alibaba Cloud), including automated customer-bucket synchronization and verified multi-terabyte, cross-region movement. - Improve efficiency and cost continuously: transfer throughput and parallelism, compute utilization, egress and storage spend, and elimination of redundant work โ€” with per-delivery cost measured and driven down. Run real deliveries - Onboard new customers: translate each customer's format, quota, and compliance requirements into a reusable delivery configuration, then land their first delivery end-to-end. - Execute and monitor production delivery runs: track staged progress, recover stuck or partial runs, reconcile failures, and verify completeness. - Ensure delivery correctness: conformance to customer specifications, deduplication and eligibility rules, and quota and budget enforcement. - Act as the engineering point of contact for delivery escalations, debugging from the customer bucket back through the pipeline to source data. What we're looking for - 6+ years (E5) / 8+ years (E6) building backend or data infrastructure, with experience taking a production system from design through operation. - Strong Python and SQL; deep familiarity with Postgres, Kubernetes, and object storage at multi-terabyte scale. - Operational rigor: you treat delivery deadlines as SLAs and failed runs as incidents. - A cost-and-efficiency mindset: you profile before you scale, understand what egress, storage, and compute actually cost, and have made a data-heavy system measurably cheaper or faster. - For E6: a track record of setting technical direction across team boundaries and simplifying systems rather than only extending them.