As part of flawless data governance and data quality implementation, Data Drift metrics play a very important role. I have compiled a few Data Drift metrics like Jensen-Shannon Divergence Score, Correlation Breakdown Detection, Data Lineage Health and Blast Radius Score etc.,
If there are multiple features available in your dataset, it's important to measure the Jensen-Shannon Divergence score feature by feature and combine the individual JSD scores into an overall Population Drift Score.
What other innovative & novel KPIs did you use in your projects that measure the quality of the data before being consumed in to your workloads?
These are natural discussions that go hand-in-hand if you are using DQX framework in Databricks.
suryaprayaga