Data Product Engineer

David Joseph Inc — United States · Posted ~16 hours ago

Senior Full-time Onsite Visa History ✓ $230K-$280K + equity

Skills

data engineering backend engineering data infrastructure AI systems data pipelines backend

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A founding data engineering role responsible for building the complete data layer behind intelligent applications. Combines data infrastructure, backend development, and AI-focused engineering challenges.

Highlights

Foundational engineering opportunity with ownership of end-to-end data systems, applied AI challenges, and potential leadership growth.

Description

San Francisco, CA On-site (5 days/week) Full-time Compensation: $230K-$280K + equity About The Company Our client is a seed-stage AI company building an intelligence platform for teams in a large, highly regulated financial-services vertical. Their product turns dense operating knowledge (filings, regulations, and manuals) into trusted, governed systems that both people and AI agents can act on. Backed by top-tier venture investors, the team is small, technically deep, and tackling hard applied-AI problems including long-context reasoning and multi-agent coordination. Founded 2025 Small team (under 20 people) Industry: AI, B2B, Data, Enterprise, Insurance The Role This is a founding data hire who owns the full data layer end-to-end, from raw external sources through to the production signals the platform's agents reason over. It is a roughly even split between data infrastructure and product/backend engineering, with a clear path to grow into leading the data function. What You'll Be Doing Partner with customers to pinpoint high-value data use cases and decide which external sources to bring onto the platformDesign, build, and operate end-to-end pipelines that ingest, extract, and synthesize data from new external sourcesSurface data through product interfaces so it is cleanly consumable by AI agents, with a focus on accuracyBuild and maintain evaluation harnesses that keep data quality and agent reliability high at scale Tech stack: Python, data pipelines, multi-agent systems, LLMs, search infrastructure, Git, SQL Requirements Around 5 to 8 years in data engineering, with a track record of building and running production pipelines and systemsHas personally built and scaled a data system end-to-end, connecting new external sources and owning them from ingestion through productionRecent experience at a high-caliber environment: a fast-growing startup (seed to Series D), a strong big-tech team, an AI-native company, or a fast-moving fintechHands-on strength in production pipeline design, ingestion, and orchestrationWorking experience with AI/ML agent frameworks and evaluation harnessesA CS or STEM degree from a strong, top-tier programBased in San Francisco or open to relocating, and comfortable working on-site 5 days a weekAuthorized to work in the US (client is open to visa transfers such as OPT or H-1B transfers) Nice to Haves Experience building agent harnesses or LLM-powered extraction and validation workflowsExperience processing large volumes of unstructured documents (e.g., PDFs and filings) Why Join Founding role with genuine end-to-end ownership of the data layer and room to grow into a leadership seatZero-to-one building with no existing playbook at an early, well-backed companyMeaningful, current AI work: agents, RAG, and eval harnesses over real-world documentsCompetitive base plus equity Details Location: San Francisco, CAWork policy: On-site, 5 days/weekCompensation: $230K-$280K + equityVisa sponsorship: Open to visa transfers (e.g. OPT, H-1B transfers)Employment type: Full-time