Big Data Integration

Big Data Integration
Authors
Dong, Xin Luna & Srivastava, Divesh
Publisher
Morgan & Claypool
Tags
compositor: windfall software , publisher: morgan & claypool publishers
ISBN
9781627052238
Date
2014-01-01T00:00:00+00:00
Size
0.91 MB
Lang
en
Downloaded: 30 times

The Big Data era is upon us: data is being generated, collected, and analyzed at an unprecedented scale, and data-driven decision making is sweeping through all aspects of society. Since the value of data explodes when it can be linked and fused with other data, addressing the big data integration (BDI) challenge is critical to realizing the promise of Big Data. BDI differs from traditional data integration in many dimensions: (i) the number of data sources, even for a single domain, has grown to be in the tens of thousands, (ii) many of the data sources are very dynamic, as a huge amount of newly collected data are continuously made available, (iii) the data sources are extremely heterogeneous in their structure, with considerable variety even for substantially similar entities, and (iv) the data sources are of widely differing qualities, with significant differences in the coverage, accuracy and timeliness of data provided. This book explores the progress that has been made by the data integration community on the topics of schema alignment, record linkage, and data fusion in addressing these novel challenges faced by big data integration, and identifies a range of open problems for the community.