Evaluating the Use of Clustering for Automatically Organising Digital Library Collections

Mark Hall, Paul Clough, Mark Stevenson

Research output: Contribution to conferencePaperpeer-review

12 Citations (Scopus)
267 Downloads (Pure)

Abstract

Large digital libraries have become available over the past years through digitisation and aggregation projects. These large collections present a challenge to the new user who wishes to discover what is available in the collections. Subject classification can help in this task, however in large collections it is frequently incomplete or inconsistent. Automatic clustering algorithms provide a solution to this, however the question remains whether they produce clusters that are sufficiently cohesive and distinct for them to be used in supporting discovery and exploration in digital libraries. In this paper we present a novel approach to investigating cluster cohesion that is based on identifying instruders in a cluster. The results from a human-subject experiment show that clustering algorithms produce clusters that are sufficiently cohesive to be used where no (consistent) manual classification exists.
Original languageEnglish
Pages323-334
DOIs
Publication statusPublished - 2012
EventTheory and Practice of Digital Libraries - Paphos, Cyprus
Duration: 23 Sept 201227 Sept 2012

Conference

ConferenceTheory and Practice of Digital Libraries
Country/TerritoryCyprus
CityPaphos
Period23/09/1227/09/12

Fingerprint

Dive into the research topics of 'Evaluating the Use of Clustering for Automatically Organising Digital Library Collections'. Together they form a unique fingerprint.

Cite this