What this is
A data warehouse is a centralised, integrated repository that stores large volumes of structured and unstructured data drawn from many systems. It is built to support business intelligence and analytics work by giving that data a scalable, high-performance environment for storage, processing and analysis. For customer data in particular, a warehouse such as Snowflake or BigQuery is often where the most complete view of a customer already lives — joined, cleaned and modelled. Bringing it into Zeotap CDP lets you act on it: make data-driven decisions, personalise customer experiences, sharpen marketing strategies and improve overall customer satisfaction. Data Warehouse is the source category that brings that data in. Instead of exporting files and pushing them to Zeotap, you give Zeotap CDP read access to a table — or, for Amazon S3, to a bucket and folder — and Zeotap pulls records from it on the schedule you choose. It is also the literal label you select in the source-creation form: every connector listed on this page is created by choosing Data Warehouse as the Category. Data Warehouse sources carry both Customer Data (CE) and Non-Customer Entity (NCE) data, so the same category covers customer profiles and the reference data around them — product catalogues, order data, campaign data and similar.Prerequisites
- You have permission to create sources in your Zeotap CDP tenant.
- You have an account with the warehouse or storage service, and the data you want to import already resides in it.
- You have read credentials Zeotap CDP can use, and the exact location of the object to read — warehouse, database, schema and table for Snowflake; project, dataset and table for BigQuery; bucket, bucket region and folder for Amazon S3; catalog, schema and table for Databricks; database and table for Microsoft Fabric. Each connector’s setup page lists the precise fields it asks for.
- You know the region of upload for the source, and your Catalogue has the target attributes defined, or you are ready to add them during mapping.
Key concepts
Every Data Warehouse source, whichever connector you pick, is configured around the same few settings.- Category and Data Source — in CREATE SOURCE you first choose Data Warehouse as the Category, then the specific connector as the Data Source. The rest of the form is connector-specific.
- Sync Frequency — how often Zeotap CDP re-reads the source. The Microsoft Fabric source form labels this Refresh Frequency. The first sync runs as soon as the source is created; every sync after that follows the frequency you set. Some connectors add further fields for these cadences — for BigQuery and Databricks you also set the sync time and, for a weekly or monthly cadence, the day. The cadences on offer differ by connector, so check the setup page for the one you are creating.
- Delta Queries Selection — offered by the connectors that support incremental sync. Set to true and Zeotap CDP gathers only the additions and changes made since the last sync; the connector’s setup page names the fields that identify them. Set to false and the full dataset is retrieved and transferred on every sync, whether or not anything changed. For a Snowflake source, prepare the table as described on its setup page before you switch delta queries on.
- Data Entity — mark the source as Customer Data or Non Customer Data, depending on whether the records describe customers or the entities around them.
- Region of upload — the storage region for the ingested records, chosen when you create the source. See Region of Storage.
Choose a connector
Zeotap CDP offers five Data Warehouse source integrations.Snowflake
BigQuery
Amazon S3
Databricks
Microsoft Fabric
Create a Data Warehouse source
The first four steps are identical for all five connectors. The connection fields in step 5 are the part that differs, and each connector’s setup page walks through them field by field.Open the Sources application
Start a new source
Choose Data Warehouse as the Category
Choose the Data Source
Enter the source details
Review and create
Verify the source was created
The source is configured correctly when:- It appears on the Source Listing page under the name you gave it.
- Its Implementation Details tab shows the connection parameters you entered.
- The first data sync begins once the source is created, and later syncs follow the Sync Frequency you selected.
FAQ
Why is Amazon S3 grouped with the data warehouses?
Why is Amazon S3 grouped with the data warehouses?
How often does Zeotap CDP re-read my table?
How often does Zeotap CDP re-read my table?
Do I have to re-ingest the whole table on every sync?
Do I have to re-ingest the whole table on every sync?
Can a Data Warehouse source carry data that is not about customers?
Can a Data Warehouse source carry data that is not about customers?