High data quality creates sustainable added value by:
Duplicates are double entries in databases caused by different spellings (e.g. “Müller” vs. “Mueller”). They are often the result of human error and can significantly impair data quality.
The main causes can be divided into four groups:
Data is of high quality if it:
Data quality is crucial for generating added value from data and making value-adding decisions. In the age of big data, it is essential that data is suitable and “fit for use” for the respective purpose.