Remote job
Senior Data Engineer
Job details
About this role
Role overview A Senior Data Engineer is needed to take ownership of the analytical data platform that powers advertising, campaign delivery, product, and financial reporting for a contextual commerce business. The work spans the full data stack, from streaming event ingestion through warehouse modeling and business intelligence, with high autonomy to reshape architecture and engineering practices. The position combines hands-on pipeline and model building with cross-functional partnership across engineering, campaign operations, finance, product, sales, and leadership.
Responsibilities - Design and evolve the cloud warehouse layout across raw, staging, core, and reporting tiers, including table grain, keys, naming, partitioning, clustering, incremental loads, and clear ownership boundaries - Build and operate batch and streaming pipelines using Python, SQL, Apache Beam, and Dataflow to move ad events, partner files, API payloads, and internal system data into the warehouse - Author and maintain orchestration workflows with Cloud Composer and Apache Airflow, emphasizing safe reruns, retries, alerting, and dependency hygiene - Stand up a version-controlled SQL transformation layer so headline calculations move out of scheduled queries and dashboards into tested, reviewable models - Investigate discrepancies between source systems, partner reports, warehouse tables, and BI outputs; repair the underlying logic and document the correction - Monitor warehouse spend and query health, govern access to sensitive fields, and enforce retention and deletion rules through infrastructure-as-code practices
Requirements - Proven track record running a production analytical platform that includes a durable raw layer, a cloud warehouse, governed models, and business-facing data marts - Deep hands-on expertise with a major cloud data warehouse (BigQuery-level fluency), including schema design, partitioning, clustering, performance tuning, permissions, retention, and cost control - Strong Python and advanced SQL skills, with experience writing tested production code across ingestion, transformation, orchestration, and serving layers - Hands-on production experience with Apache Beam/Dataflow and Apache Airflow/Cloud Composer, covering idempotency, schema evolution, late-arriving events, replay, backfills, and incident recovery - Sound data-modeling judgment across dimensional, event, and domain designs, including facts, dimensions, slowly changing dimensions, semantic layers, and reusable analytical datasets - Demonstrated ability to operate as an autonomous senior technical owner, making architectural calls, sequencing migrations, leading incidents, and explaining tradeoffs to non-technical partners
Nice to have - Familiarity with data governance, least-privilege access, privacy handling, and auditability for revenue-impacting datasets - Experience with software-engineering practices for data systems, including Git, automated testing, CI/CD, code review, Terraform, and clear technical documentation
Benefits and work setup - Base salary range of $160,000 to $185,000 annually