DSpace at KOASAS: A node resistance-based probability model for resolving duplicate named entities

DSpace at KOASAS

RIMS Collection RIMS Journal Papers

A node resistance-based probability model for resolving duplicate named entities

Cited 2 time in

Cited 1 time in

Hit : 191
Download : 0

Export

Kang, Namyong / Kim, Jeong-Jae / On, Byung-Won / Lee, Ingyu

Duplicate entities tend to degrade the quality of data seriously. Despite recent remarkable achievement, existing methods still produce a large number of false positives (i.e., an entity determined to be a duplicate one when it is not) that are likely to impair the accuracy. Toward this challenge, we propose a novel node resistance-based probability model in which we view a given data set as a graph of entities that are linked each other via relationships, and then compute the probability value between two entities to see how similar the two entities are. Especially, in the graph, each node has its own resistance value equivalent to1-confidence(normalized in 0-1) and resistance probabilityvalue is filtered out per node during computing the probability value. To evaluate the proposed model, we performed intensive experiments with different data sets including ACM https://dl.acm.org), DBLP (https://dblp.uni-trier.de), and IMDB (https://imdb.com). Our experimental results show that the proposed probability model outperforms the existing probability model, improving average F1 scores up to 14%, but never worsens them.

Publisher: SPRINGER

Issue Date: 2020-09

Language: English

Article Type: Article

Citation: SCIENTOMETRICS, v.124, no.3, pp.1721 - 1743

ISSN: 0138-9130

DOI: 10.1007/s11192-020-03585-4

URI: http://hdl.handle.net/10203/279532

Appears in Collection: RIMS Journal Papers

Files in This Item: There are no files associated with this item.

This item is cited by other documents in WoS

⊙ Detail Information in WoSⓡ	Click to see
⊙ Cited 2 items in WoS	Click to see citing articles in

Display Full Item Record

qr_code

트윗하기

KOASAS

Knowledge Service Development Team, KAIST 291 Daehak-ro, Yuseong-gu, Daejeon 34141, Republic of Korea. T. 82-42-350-4493 Email. koasas@kaist.ac.kr
Copyright © 2016. Korea Advanced Institute of Science and Technology. All Rights Reserved.

KOASAS

KOASAS

Browse

A node resistance-based probability model for resolving duplicate named entities

This item is cited by other documents in WoS

KOASAS

Communities & Collections