zencastr
zencastr17d ago
New

Senior Data Engineer (Remote)

San Francisco OfficeRemotefull-timesenior
Data EngineerData
0 views0 saves0 applied

Quick Summary

Overview

About the Role As our first dedicated Data Engineer, you will build and own the data foundation that powers analytics, reporting, and decision-making across the organization.

Technical Tools
Data EngineerData

About the Role

~2 min read

As our first dedicated Data Engineer, you will build and own the data foundation that powers analytics, reporting, and decision-making across the organization. This is a hands-on role where you'll design the dimensional model, own the pipelines that feed it, and establish the standards our data practice is built on.

You won't be starting from zero or working alone. You'll join a data team with direct analytics experience, and partner closely with engineering and other technical teams who have built and run what we have today. There's real institutional knowledge here to draw on. What's been missing is someone whose focus is turning it into a single, well-modeled foundation the whole company can rely on.

You will work cross-functionally to understand how data is generated and used, and translate those needs into scalable models and structured reporting layers. A central part of the work is identity resolution: building the spine that reliably connects the same entity as it appears across systems that each have their own identifiers and lifecycles.

This role is ideal for someone who enjoys owning data systems end to end — from ingestion and transformation through modeling, governance, and performance — and who wants the autonomy to design a warehouse properly, with colleagues who can help you understand the business behind the data.

Responsibilities

~2 min read
  • →

    Design and build a conformed, Kimball-style dimensional model across our operational, behavioral, and transactional data

  • →

    Own ingestion end to end, including capturing change over time from sources that don't preserve history natively

  • →

    Consolidate transformation logic that currently lives in more than one place into a single governed, tested layer

  • →

    Implement and manage our data warehouse and transformation layer, taking ownership of the pipelines that move data from our operational systems into it

  • →

    Establish foundational best practices for data modeling, documentation, testing, and governance

  • →

    Improve data reliability, quality, and accessibility across systems

  • →

    Collaborate with analysts and business stakeholders to support evolving data needs

  • →

    Encode business metric definitions once, so that reporting stops drifting across teams

  • →

    Monitor and optimize performance and cost efficiency across pipelines, storage, and warehouse queries

  • →

    Design, build, and maintain reliable data pipelines that transform operational data into structured, analytics-ready datasets

  • →

    Design and maintain the dimensional model — dimensions, facts, and bridge tables with clearly defined grain

  • →

    Develop and maintain scalable data models and data marts to support reporting and business analysis

  • →

    Manage and optimize data ingestion and event workflows to ensure consistent, high-quality upstream data flows

  • →

    Implement and manage transformation processes that structure raw data for analytics use

  • →

    Implement orchestration, testing, freshness monitoring, and alerting so that data issues are caught before stakeholders encounter them

  • →

    Build and maintain change capture or snapshotting to support historical reporting and slowly-changing dimensions

  • →

    Improve data freshness — moving our core operational data from batch refreshes toward near-real-time availability, and establishing freshness SLAs stakeholders can rely on

  • →

    Ensure strong standards for data quality, validation, and consistency across systems

  • →

    Document models and definitions so analysts and stakeholders can self-serve with confidence

  • →

    Monitor and optimize performance, reliability, and cost efficiency within the analytics environment

  • →

    Partner cross-functionally to translate business requirements into scalable data solutions

  • →

    Proactively improve our data systems so they remain structured, consistent, and scalable as the organization grows

  • Have 5+ years of experience specifically in data engineering, analytics engineering, or a closely related role, including having built and owned a dimensional model in production

  • Have strong proficiency in SQL, with experience across document-based operational databases (e.g., MongoDB) and analytical data warehouses (e.g., BigQuery, Snowflake, Redshift, or similar)

  • Have experience building fact and dimension tables using star schema principles to support reporting and data marts, with a clear point of view on grain, conformed dimensions, and slowly-changing dimensions

  • Have hands-on experience with modern transformation and modeling frameworks (e.g., dbt, Dataform, or similar), including managing transformation layers within a warehouse environment with version control, testing, and CI

  • Have built and maintained reliable ETL/ELT pipelines that transform raw application data into structured, analytics-ready datasets

  • Have worked with orchestration tooling (e.g., Airflow, Dagster, Prefect, or similar) and think in terms of dependencies, retries, and backfills

  • Have experience with data ingestion or event streaming platforms (e.g., RudderStack, Segment, Pub/Sub, or similar) and ensuring consistent, reliable upstream data flows, including identity stitching across web and mobile

  • Have a solid understanding of data modeling best practices, including schema design, dimensional modeling, and performance considerations

  • Have a track record of inheriting and operating systems you didn't build

  • Have a strong focus on data quality, validation, and governance, with the ability to identify and resolve inconsistencies

  • Have an understanding of performance optimization across pipelines, storage, and warehouse queries

  • Can explain technical tradeoffs clearly to non-engineers

  • Are comfortable operating in a growing environment where you both execute technically and help shape our data architecture standards

Nice to Have

~1 min read
  • Change data capture patterns from operational databases

  • Subscription billing data — proration, refunds, failed payments, trials

  • Experience with distributed processing frameworks (e.g., Spark, Beam, Dataflow)

  • Experience as a first or early data hire

Location & Eligibility

Where is the job
Worldwide
Fully remote, anywhere in the world
Who can apply
Same as job location

Listing Details

Posted
September 11, 2026
First seen
September 26, 2026
Last seen
September 28, 2026

Posting Health

Days active
1
Repost count
0
Trust Level
34%
Scored at
September 28, 2026

Signal breakdown

freshnesssource trustcontent trustemployer trust
Newsletter

Stay ahead of the market

Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.

A
B
C
D
Join 12,000+ marketers

No spam. Unsubscribe at any time.

zencastrSenior Data Engineer (Remote)