Senior Data/strong Spark & Backend API Engineer-PHP/Symfony+AI+Java/Python+AWS+Iceberg with must-have AdTech experience || Remote role

Verito Solutions Inc โ€” United States ยท Posted ~4 hours ago

๐Ÿ”“ Log in to save this job, tailor your resume & track your apply process โ€” 7 days free, no card needed.

Log in to add to target list

Description

Senior Data/strong Spark & Backend API Engineer-PHP/Symfony+AI+Java/Python+AWS+Iceberg with must-have AdTech experience Remote role Long Term Contract Strong Spark + Java/Python/PySpark experience AWS (S3), Airflow, Snowflake, Iceberg & Advanced SQL PHP/Symfony + REST API development experience AdTech/advertising measurement experience โ€” attribution, conversions, impressions, identity matching, partner feeds Experience with external partner integrations and large-scale data pipelines Strong production troubleshooting, testing, backfills & data reconciliation Hands-on AI-assisted/agentic development experience Job Description Responsibilities: โ— Build, optimize, and operate production Apache Spark pipelines that generate attributed conversion and optimization feeds from large-scale impression and conversion data. โ— Design and evolve data processing and models across modern data lake and warehouse technologies, orchestrated with a production workflow scheduler on AWS. โ— Build and operate partner optimization feeds and integrations with external partners, including schemas, identity fields and unique IDs, file delivery, reconciliation, and SLAs. โ— Build, change, and operate production REST APIs, including PHP APIs, delivering secure, backward-compatible changes end to end (implementation, validation, authorization, data access, documentation, automated testing, and production diagnostics). โ— Trace and modify behavior across controllers, database queries, templates, JavaScript, and automated tests. โ— Diagnose production failures, reconcile partner-facing data, execute backfills, and manage staged releases and rollbacks. โ— Write and maintain unit and integration tests across pipelines, APIs, and feeds. Qualifications and Education Requirements: โ— 7+ years of professional software or data engineering experience. โ— 4+ years building, optimizing, and maintaining production Apache Spark pipelines, with strong Java and Python skills across JVM Spark and PySpark. โ— Advanced SQL, relational data modeling, and large-scale data processing experience, including production work with Snowflake, MySQL, and Iceberg. โ— 3+ years operating AWS data workloads using S3 and Airflow (or equivalent production workflow orchestration); container orchestration and data-catalog experience a plus. โ— 3+ years building, changing, and operating production REST APIs, including PHP APIs built with Symfony or a closely equivalent PHP MVC framework. โ— Demonstrated proficiency using AI development tools to deliver production-quality work efficiently โ€” including AI-assisted coding, code review, documentation, and investigation, working with agentic code harnesses, and spec-based (spec-driven) AI development โ€” with the judgment to know when to rely on AI output and when not to. โ— Ability to independently deliver secure, backward-compatible API changes, including implementation, validation, authorization, data access, documentation, automated testing, and production diagnostics. โ— Experience implementing external provider or client integrations involving schemas, APIs, file delivery, identity fields, reconciliation, privacy-sensitive data, and SLAs. โ— Production full-stack experience with PHP, Symfony (or a comparable framework), Doctrine, Twig (or another server-rendered template system), and JavaScript. โ— Ability to trace and modify behavior across controllers, database queries, templates, JavaScript, and automated tests. โ— Experience writing unit and integration tests for data pipelines, APIs, and web applications. โ— Ability to diagnose production failures, reconcile data, execute backfills, manage staged releases and rollbacks, and support delivery commitments. โ— Hands-on experience building and running big data systems, with a focus on performance, reliability, and data quality. โ— Experience in ad tech or advertising measurement, such as attribution, conversions, audience and impression data, identity matching, or partner optimization feeds. โ— Ability to ramp quickly and deliver with minimal onboarding in an existing, complex codebase. Preferred Skills: โ— Direct experience building partner optimization or advertising feeds with external platforms (DSPs, publishers, or measurement partners). โ— Experience with deterministic and probabilistic identity matching โ€” unique IDs, device IDs, IP- based matching, and identity graphs. โ— Familiarity with edge/log delivery infrastructure and pixel/impression tracking. โ— Experience with data lake table formats and query engines at large scale. โ— Experience with privacy and compliance-driven data workflows (deletion, opt-out, suppression, data retention). โ— Comfort operating in uncharted territory โ€” turning ambiguity into production systems without a detailed guide. Additional Notes: AI-forward engineering. is an AI-forward engineering organization, and we expect engineers โ€” including contractors โ€” to work with AI, not around it. We use AI-assisted development as a standard part of our workflow, and our engineers use it for tickets, design docs, runbooks, and scaffolding to move fast without sacrificing quality. You should be comfortable using AI for code generation, review, documentation, and investigation, and equally comfortable knowing when not to trust the output. We care about outcomes: how effectively you leverage AI to deliver reliable, secure, maintainable systems is part of the job. On-call. This role includes availability for occasional off-hours release or incident support. It does not include formal on-call duty unless separately established. Sayantan Das Senior Tech Recruiter | Talent Acquisition | People & Culture [ Verito Solutions - An E-verified company] [ sayantan@veritosolutions.com ]