hirly

Zencastr

Senior Data Engineer (Remote)

San Francisco Office

See how you match this job — and similar ones. Free.

Upload your resume and hirly scores it against this role at Zencastr first, then against similar open jobs, and shows where you fit and why.

PDF or DOCX, up to 12MB. No sign-up to see your matches.

Get past the screening software and onto a recruiter's desk

hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.

  • Keywords matched to this posting
  • Fit score before you apply
  • Cover letter included

Matched against 2.5M live jobs from 200,000+ employers in 200+ countries.

Tailor my resume for this job →

Apply from your AI assistant

Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.

Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.

hirly's read of this role

Role family
Data & ML
Seniority
Senior
Country
US
Work mode
Remote-friendly
First seen by hirly
11 Sept 2026

Derived automatically from the posting. Upload your resume above to see how the role scores against it.

the posting

About the Role

As our first dedicated Data Engineer, you will build and own the data foundation that powers analytics, reporting, and decision-making across the organization. This is a hands-on role where you'll design the dimensional model, own the pipelines that feed it, and establish the standards our data practice is built on.

You won't be starting from zero or working alone. You'll join a data team with direct analytics experience, and partner closely with engineering and other technical teams who have built and run what we have today. There's real institutional knowledge here to draw on. What's been missing is someone whose focus is turning it into a single, well-modeled foundation the whole company can rely on.

You will work cross-functionally to understand how data is generated and used, and translate those needs into scalable models and structured reporting layers. A central part of the work is identity resolution: building the spine that reliably connects the same entity as it appears across systems that each have their own identifiers and lifecycles.

This role is ideal for someone who enjoys owning data systems end to end — from ingestion and transformation through modeling, governance, and performance — and who wants the autonomy to design a warehouse properly, with colleagues who can help you understand the business behind the data.

What You’ll Do

Design and build a conformed, Kimball-style dimensional model across our operational, behavioral, and transactional data

Own ingestion end to end, including capturing change over time from sources that don't preserve history natively

Consolidate transformation logic that currently lives in more than one place into a single governed, tested layer

Implement and manage our data warehouse and transformation layer, taking ownership of the pipelines that move data from our operational systems into it

Establish foundational best practices for data modeling, documentation, testing, and governance

Improve data reliability, quality, and accessibility across systems

Collaborate with analysts and business stakeholders to support evolving data needs

Encode business metric definitions once, so that reporting stops drifting across teams

Monitor and optimize performance and cost efficiency across pipelines, storage, and warehouse queries

You're a Good Fit If You

Have 5+ years of experience specifically in data engineering, analytics engineering, or a closely related role, including having built and owned a dimensional model in production

Have strong proficiency in SQL, with experience across document-based operational databases (e.g., MongoDB) and analytical data warehouses (e.g., BigQuery, Snowflake, Redshift, or similar)

Have experience building fact and dimension tables using star schema principles to support reporting and data marts, with a clear point of view on grain, conformed dimensions, and slowly-changing dimensions

Have hands-on experience with modern transformation and modeling frameworks (e.g., dbt, Dataform, or similar), including managing transformation layers within a warehouse environment with version control, testing, and CI

Have built and maintained reliable ETL/ELT pipelines that transform raw application data into structured, analytics-ready datasets

Have worked with orchestration tooling (e.g., Airflow, Dagster, Prefect, or similar) and think in terms of dependencies, retries, and backfills

Have experience with data ingestion or event streaming platforms (e.g., RudderStack, Segment, Pub/Sub, or similar) and ensuring consistent, reliable upstream data flows, including identity stitching across web and mobile

Have a solid understanding of data modeling best practices, including schema design, dimensional modeling, and performance considerations

Have a track record of inheriting and operating systems you didn't build

Have a strong focus on data quality, validation, and governance, with the ability to identify and resolve inconsistencies

Have an understanding of performance optimization across pipelines, storage, and warehouse queries

Can explain technical tradeoffs clearly to non-engineers

Are comfortable operating in a growing environment where you both execute technically and help shape our data architecture standards

Nice to have

Change data capture patterns from operational databases

Subscription billing data — proration, refunds, failed payments, trials

Experience with distributed processing frameworks (e.g., Spark, Beam, Dataflow)

Experience as a first or early data hire

Key Responsibilities

Design, build, and maintain reliable data pipelines that transform operational data into structured, analytics-ready datasets

Design and maintain the dimensional model — dimensions, facts, and bridge tables with clearly defined grain

Develop and maintain scalable data models and data marts to support reporting and business analysis

Manage and optimize data ingestion and event workflows to ensure consistent, high-quality upstream data flows

Implement and manage transformation processes that structure raw data for analytics use

Implement orchestration, testing, freshness monitoring, and alerting so that data issues are caught before stakeholders encounter them

Build and maintain change capture or snapshotting to support historical reporting and slowly-changing dimensions

Improve data freshness — moving our core operational data from batch refreshes toward near-real-time availability, and establishing freshness SLAs stakeholders can rely on

Ensure strong standards for data quality, validation, and consistency across systems

Document models and definitions so analysts and stakeholders can self-serve with confidence

Monitor and optimize performance, reliability, and cost efficiency within the analytics environment

Partner cross-functionally to translate business requirements into scalable data solutions

Proactively improve our data systems so they remain structured, consistent, and scalable as the organization grows

Original posting on Zencastr's site ↗

Browse similar roles

Want this one?

Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.

Tailor my resume for this job