Evaluating the Use of Clustering for Automatically Organising Digital Library Collections

Hall, Mark Michael; Clough, Paul and Stevenson, Mark (2012). Evaluating the Use of Clustering for Automatically Organising Digital Library Collections. In: Theory and Practice of Digital Libraries, Springer Berlin / Heidelberg, 7489 pp. 323–334.

DOI: https://doi.org/10.1007/978-3-642-33290-6_35

URL: http://dx.doi.org/10.1007/978-3-642-33290-6_35

Abstract

Large digital libraries have become available over the past years through digitisation and aggregation projects. These large collections present a challenge to the new user who wishes to discover what is available in the collections. Subject classification can help in this task, however in large collections it is frequently incomplete or inconsistent. Automatic clustering algorithms provide a solution to this, however the question remains whether they produce clusters that are sufficiently cohesive and distinct for them to be used in supporting discovery and exploration in digital libraries. In this paper we present a novel approach to investigating cluster cohesion that is based on identifying instruders in a cluster. The results from a human-subject experiment show that clustering algorithms produce clusters that are sufficiently cohesive to be used where no (consistent) manual classification exists.

Viewing alternatives

Metrics

Public Attention

Altmetrics from Altmetric

Number of Citations

Citations from Dimensions

Item Actions

Export

About

Recommendations