Have a personal or library account? Click to login
The Challenges of Data Quality and Data Quality Assessment in the Big Data Era Cover

The Challenges of Data Quality and Data Quality Assessment in the Big Data Era

By: Li Cai and  Yangyong Zhu  
Open Access
|May 2015

Abstract

High-quality data are the precondition for analyzing and using big data and for guaranteeing the value of the data. Currently, comprehensive analysis and research of quality standards and quality assessment methods for big data are lacking. First, this paper summarizes reviews of data quality research. Second, this paper analyzes the data characteristics of the big data environment, presents quality challenges faced by big data, and formulates a hierarchical data quality framework from the perspective of data users. This framework consists of big data quality dimensions, quality characteristics, and quality indexes. Finally, on the basis of this framework, this paper constructs a dynamic assessment process for data quality. This process has good expansibility and adaptability and can meet the needs of big data quality assessment. The research results enrich the theoretical scope of big data and lay a solid foundation for the future by establishing an assessment model and studying evaluation algorithms.

Language: English
Published on: May 22, 2015
Published by: Ubiquity Press
In partnership with: Paradigm Publishing Services
Publication frequency: 1 issue per year

© 2015 Li Cai, Yangyong Zhu, published by Ubiquity Press
This work is licensed under the Creative Commons Attribution 4.0 License.