A preliminary assessment of the article deduplication algorithm used for the OpenAIRE Research Graph
Contributo in Atti di convegno
Data di Pubblicazione:
2022
Abstract:
In recent years, a large number of Scholarly Knowledge Graphs (SKGs) have been introduced in the literature. The communities behind these graphs strive to gather, clean, and integrate scholarly metadata from various sources to produce clean and easy-to-process knowledge graphs. In this context, a very important task of the respective cleaning and integration workflows is deduplication. In this paper, we briefly describe and evaluate the accuracy of the deduplication algorithm used for the OpenAIRE Research Graph. Our experiments show that the algorithm has an adequate performance producing a small number of false positives and an even smaller number of false negatives.
Tipologia CRIS:
04.01 Contributo in Atti di convegno
Keywords:
Deduplication; Open Science; Scholarly data; Knowledge graphs
Elenco autori:
DE BONIS, Michele; Manghi, Paolo; Atzori, Claudio
Link alla scheda completa:
Link al Full Text:
Titolo del libro:
IRCDL 2022 Italian Research Conference on Digital Libraries 2022
Pubblicato in: