Blog — Data Vault & Data Warehouse Automation
Articles on Data Vault 2.0, data warehouse automation, Snowflake, Databricks, and practical data engineering.
-
Datavault Builder and the DAMA-DMBOK Framework — How We Support the Eleven Knowledge Areas
A walk through the eleven DAMA-DMBOK domains and how Datavault Builder supports each one — built specifically to make data warehousing, integration and modeling easier.
Read more → -
Data Sovereignty & AI: Enterprise DWH
How do you keep your data under control when AI models need access to everything? This webinar shows how to combine data governance with modern AI capabilities.
Read more → -
Data Vault & Databricks Medallion Stack
Data Vault and Lakehouse are often seen as conflicting approaches. This webinar shows why the opposite is true.
Read more → -
How I Think About Business Value in Data
Most data projects don’t fail because of technology. They fail because value arrives too late, in the wrong form, or at too high a cost.
Read more → -
ISO 27001:2022 Certification & SOC 2 Type 2
In today’s rapidly evolving digital landscape, data security is not just a priority—it's a necessity.
Read more → -
Import Flow.BI AI generated Data Models
Watch how to migrate Flow.BI AI-generated metadata into Datavault Builder using the Migration Vault approach — mapping hubs, links, and automated ETL pipelines.
Read more → -
Integrating and Unioning Data
In Data Vault Modeling, we use hubs to integrate data. This is one of the main reasons we choose Data Vault modeling.
Read more → -
Near Real-Time DWH Analytics
In today's data-driven world, credit cards, networks, IoT sensors, and numerous data sources provide real-time data. How can this be processed effectively?
Read more → -
SSO on Snowflake Snowpark Container Services
Using Datavault Builder with Snowflake's Snowpark container services lets data teams set up and manage data solutions without relying on infrastructure teams.
Read more → -
Bi-Temporal Data Processing
Whether you're dealing with customer accounts, financial transactions, or insurance claims, this webinar equips you to handle bi-temporal data.
Read more → -
DVB on Snowflake Snowpark Container Services
In this video we do demonstrate how simple it is to run Datavault Builder on Snowflake Container services.
Read more → -
Direct Model GIT Check-In and Check-Out
Welcome to our short presentation on Datavault Builder 7.1 and its new GIT check-in and check-out feature for agile development.
Read more → -
The Migration Vault Concept
A Migration Vault is a specialized metadata store designed as a Data Vault. It facilitates migrating data from an old format to a new one, ensuring compatibility.
Read more → -
Talend Open Studio Alternative Compared
Datavault Builder offers a large variety of features and modules for the same cost as previously free ETL tools.
Read more → -
Unified Star Schema Automation
In this video, viewers delve into the world of automated data warehousing with the Unified Star Schema, a concept by Francesco Puppini and Bill Inmon.
Read more → -
Yale University: Revolutionizing Data Management for 75% Savings and Faster Insights
Data management automation drives major productivity gains: how Yale University used Datavault Builder to cut 75% of billed hours, expand its data team, and accelerate time to insight.
Read more → -
BI-SPEKTRUM Case Study: C&A on Snowflake
BI-SPEKTRUM published an article in issue 2023/3 how one of our clients is using the Datavault Builder to integrate its SAP Data.
Read more → -
DDVUG Willibald Use Case
Watch the DDVUG Willibald use case: building a full Data Warehouse with 2 data sources in under 3 hours using Datavault Builder.
Read more → -
Is Data Modeling dead
Do we still need data models? What is the value of models? Why did data modeling fail in the past? How can we create value by modeling?
Read more → -
Data Vault Bi-temporal: Inscription Time
How to load bi-temporal data into the Data Vault using Inscription Time — patterns, pitfalls, and practical examples.
Read more → -
Data Vault vs. Data Mesh?
Should I still do Data Vault if there is Data Mesh? In the past few weeks and months, I got these very interesting questions which brought up a several times.
Read more → -
CI/CD with Datavault Builder on Snowflake
How to use Snowflake's Zero Copy Cloning with Datavault Builder to build a powerful CI/CD pipeline for your data warehouse.
Read more → -
Do Equi-Joins always matter?
It happens that from time to time I comeacross some statements about how databases work and how they shall be queried. And I like to read those recommendations.
Read more → -
3NF and Data Vault: Nothing to Fear
From time to time we receive an interesting question: does Datavault Builder support 3NF? The answer is yes — and here is how it works.
Read more → -
DWH Temporality Pt.4: SCD Type 2 Dimensions
Kimball Style dimensions - SCD Type 2 Output If you haven't read them I recommend reading the first 3 parts first.
Read more → -
DWH Temporality Part 3: Outputting Timelines
Although in many cases it is not necessary to output the timelines in the reports, there are some cases where the output of timelines is important.
Read more → -
DWH Temporality Part 2: Reducing Complexity
How to reduce temporal complexity in the Data Vault. Datavault Builder users frequently ask how to map changes over time correctly.
Read more → -
DWH Temporality Part 1: The Challenge
In the past years, I was confronted with the demand to create a reporting with SCD type 2 dimensions.
Read more → -
On Multi-Active Satellites in Data Vault
Petr Beles on implementing Multi-Active Satellites as Document Satellites in Data Vault: patterns, trade-offs, and practical guidance.
Read more → -
On Links
Petr Beles on Data Vault links representing transactions: patterns, pitfalls, and design decisions when modeling transaction links.
Read more → -
Qlik
Do Your Qlik Apps and Power BI Reports Show Different Numbers?
Neither Qlik nor Power BI is wrong. Each has its own load logic, its own definitions and its own lineage, so the same metric is calculated twice and nobody can reconcile them. The rules belong in one governed warehouse model both tools read.
Read more → -
Tableau
Does Your Tableau Server Have Too Many Versions of the Same Number?
Every published .tdsx was reasonable on the day it was made. Together they are hundreds of private definitions of the same metric, and no way to tell which one is right.
Read more → -
Power BI
Do Your Power BI Reports Show Different Numbers for the Same Thing?
Finance, Sales and Operations each built their own semantic model, and each is internally consistent. The definitions were never wrong in one place. They were never agreed in any place.
Read more → -
Qlik
Do You Clean the Same Data Again in Every Qlik Load Script?
The same mapping loads, string fixes and deduplication are written again in every app and QVD layer, and the copies drift apart. The cleansing is repeated per script because no integrated warehouse layer does it once.
Read more → -
Qlik
Are Your Qlik Set Analysis Expressions Too Long and Too Slow?
Set analysis is precise for genuine comparisons. Most of the long expressions in your app are there because the model never delivered history, flags or a single grain, so the chart rebuilds them on every selection.
Read more → -
Tableau
Are Your Tableau LOD Expressions Too Complex to Touch?
FIXED, INCLUDE and EXCLUDE are precise tools for genuine multi grain questions. Most of the ones in your workbook are there because the warehouse never resolved the grain or kept the history.
Read more → -
Power BI
Is Your Power BI DirectQuery Report Slow on Every Click?
DirectQuery and Direct Lake promise live data. What you get is a thirty second visual and a compute bill nobody wants to explain. The mode is not the problem. The schema underneath it is.
Read more → -
Qlik
Do Your Qlik Apps Keep Creating Synthetic Keys?
Synthetic keys and circular references are Qlik associating exactly what it was given. They appear because the data arrives without conformed dimensions or real keys, so every app has to invent them.
Read more → -
Qlik
Do Your Qlik Reloads Keep Failing as the Data Grows?
The in-memory engine is fast because everything sits in RAM. Reloads fail and apps slow down when that RAM is filled with row level detail and transformations that no warehouse did beforehand.
Read more → -
Tableau
Do Your Tableau Extract Refreshes Keep Failing or Running Late?
The backgrounder times out, the .hyper file keeps growing, and the dashboard shows yesterday. The extract is large because it is carrying raw rows that were never aggregated upstream.
Read more → -
Power BI
Is Your Power BI DAX Getting Too Long to Maintain?
Two hundred lines of CALCULATE and FILTER is not a sign of advanced DAX. It is usually a sign that the warehouse never gave you the keys, the history or the grain you needed.
Read more → -
Tableau
Is Your Tableau Dashboard Slow Every Time You Change a Filter?
Twenty seconds of "Executing Query" on every filter click. Tableau is not rendering slowly. It is waiting for a database that was handed a question it cannot answer quickly.
Read more → -
Power BI
Does Your Power BI Refresh Keep Failing Overnight?
Scheduled refresh times out, Power Query runs out of memory, and you find out when someone opens the dashboard. The cause is almost never Power BI. It is what Power BI is being asked to do.
Read more →
Nothing here yet.