Free shipping on orders over $99
The Four Generations of Entity Resolution

The Four Generations of Entity Resolution

by George PapadakisEkaterini Ioannou and Emanouil Thanos
Hardback
Publication Date: 30/03/2021

Share This Book:

RRP  $135.30

RRP means 'Recommended Retail Price' and is the price our supplier recommends to retailers that the product be offered for sale. It does not necessarily mean the product has been offered or sold at the RRP by us or anyone else.

$122.95
or 4 easy payments of $30.74 with
afterpay
This item qualifies your order for FREE DELIVERY
This book organizes entity resolution (ER) into four generations based on the challenges posed by "the four Vs," Veracity, Volume, Variety, and Velocity. Entity resolution lies at the core of data integration and cleaning and, thus, a bulk of the research examines ways for improving its effectiveness and time efficiency.

For each generation, we outline the corresponding ER workflow, discuss the state-of-the-art methods per workflow step, and present current research directions. The discussion of these methods takes into account a historical perspective, explaining the evolution of the methods over time along with their similarities and differences. The lecture also discusses the available ER tools and benchmark datasets that allow expert as well as novice users to make use of the available solutions.

The initial ER methods primarily target Veracity in the context of structured (relational) data that are described by a schema of well-known quality and meaning. To achieve high effectiveness, they leverage schema, expert, and/or external knowledge. Part of these methods are extended to address Volume, processing large datasets through multi-core or massive parallelization approaches, such as the MapReduce paradigm. However, these early schema-based approaches are inapplicable to Web Data, which abound in voluminous, noisy, semi-structured, and highly heterogeneous information. To address the additional challenge of Variety, recent works on ER adopt a novel, loosely schema-aware functionality that emphasizes scalability and robustness to noise. Another line of present research focuses on the additional challenge of Velocity, aiming to process data collections of a continuously increasing volume. The latest works, though, take advantage of the significant breakthroughs in Deep Learning and Crowdsourcing, incorporating external knowledge to enhance the existing words to a significant extent.
ISBN:
9781636390581
9781636390581
Category:
Human-computer interaction
Format:
Hardback
Publication Date:
30-03-2021
Publisher:
Morgan & Claypool Publishers
Country of origin:
United States
Pages:
109
Dimensions (mm):
235x191mm
Weight:
0.33kg

This title is in stock with our Australian supplier and should arrive at our Sydney warehouse within 1-2 weeks of you placing an order.

Once received into our warehouse we will despatch it to you with a Shipping Notification which includes online tracking.

Please check the estimated delivery times below for your region, for after your order is despatched from our warehouse:

ACT Metro 2 working days

NSW Metro 2 working days 

NSW Rural 2-3 working days

NSW Remote 2-5 working days

NT Metro 3-6 working days

NT Remote 4-10 working days

QLD Metro 2-4 working days

QLD Rural 2-5 working days

QLD Remote 2-7 working days

SA Metro 2-5 working days

SA Rural 3-6 working days

SA Remote 3-7 working days

TAS Metro 3-6 working days

TAS Rural 3-6 working days

VIC Metro 2-3 working days

VIC Rural 2-4 working days

VIC Remote 2-5 working days

WA Metro 3-6 working days

WA Rural 4-8 working days

WA Remote 4-12 working days

Reviews

Be the first to review The Four Generations of Entity Resolution.