The Four Generations of Entity Resolution (Synthesis Lectures on Data Management)

★★★★★ 4.6 68 reviews

$51.72
Price when purchased online
Free shipping Free 30-day returns

Sold and shipped by polyproblem.org
We aim to show you accurate product information. Manufacturers, suppliers and others provide what you see here.
$51.72
Price when purchased online
Free shipping Free 30-day returns

How do you want your item?
You get 30 days free! Choose a plan at checkout.
Shipping
Arrives Jul 1
Free
Pickup
Check nearby
Delivery
Not available

Sold and shipped by polyproblem.org
Free 30-day returns Details

Product details

Management number 232052293 Release Date 2026/06/18 List Price $20.69 Model Number 232052293
Category

Entity Resolution (ER) lies at the core of data integration and cleaning and, thus, a bulk of the research examines ways for improving its effectiveness and time efficiency. The initial ER methods primarily target Veracity in the context of structured (relational) data that are described by a schema of well-known quality and meaning. To achieve high effectiveness, they leverage schema, expert, and/or external knowledge. Part of these methods are extended to address Volume, processing large datasets through multi-core or massive parallelization approaches, such as the MapReduce paradigm. However, these early schema-based approaches are inapplicable to Web Data, which abound in voluminous, noisy, semi-structured, and highly heterogeneous information. To address the additional challenge of Variety, recent works on ER adopt a novel, loosely schema-aware functionality that emphasizes scalability and robustness to noise. Another line of present research focuses on the additional challenge ofVelocity, aiming to process data collections of a continuously increasing volume. The latest works, though, take advantage of the significant breakthroughs in Deep Learning and Crowdsourcing, incorporating external knowledge to enhance the existing words to a significant extent. This synthesis lecture organizes ER methods into four generations based on the challenges posed by these four Vs. For each generation, we outline the corresponding ER workflow, discuss the state-of-the-art methods per workflow step, and present current research directions. The discussion of these methods takes into account a historical perspective, explaining the evolution of the methods over time along with their similarities and differences. The lecture also discusses the available ER tools and benchmark datasets that allow expert as well as novice users to make use of the available solutions. Read more

ISBN10 3031007506
ISBN13 978-3031007507
Edition 1st
Language English
Publisher Springer
Dimensions 7.52 x 0.39 x 9.25 inches
Item Weight 11.9 ounces
Print length 172 pages
Publication date March 16, 2021

Correction of product information

If you notice any omissions or errors in the product information on this page, please use the correction request form below.

Correction Request Form

Customer ratings & reviews

4.6 out of 5
★★★★★
68 ratings | 28 reviews
How item rating is calculated
View all reviews
5 stars
84% (57)
4 stars
3% (2)
3 stars
2% (1)
2 stars
1% (1)
1 star
10% (7)
Sort by

There are currently no written reviews for this product.