Visual Analysis of Clustering Algorithms A Methodology and a Case Study - LIRMM - Laboratoire d’Informatique, de Robotique et de Microélectronique de Montpellier Access content directly
Reports Year : 2011

Visual Analysis of Clustering Algorithms A Methodology and a Case Study

Guillaume Artignan
  • Function : Author
  • PersonId : 862749
Mountaz Hascoët
  • Function : Author
  • PersonId : 837916

Abstract

Clustering is probably one of the most frequently used ap- proaches when facing a scaling problem in large collections of documents. In many situations, however, the choice of the most appropriate algo- rithm for clustering can turn into a real dilemma. Numerical criteria have been proposed to evaluate the quality of the results of clustering algorithms. However, so many different criteria have been proposed that the dilemma even worsens. Most criteria reveal different aspects of the quality of the results and hide others. The aim of this paper is to help with the understanding of clustering and to facilitate the comparison and the choice of clustering algorithm for a given purpose. Our proposal consists in studying both quality evaluation criteria and clustering algo- rithms. We start by discussing a selected set of representative criteria, and further conduct a case study on a large set of real data, measuring not only the quality of different representative clustering algorithms but also the impact of each criterion on the ranking of the algorithms. By providing empirical results on large scale corpus of either documents or lexical networks useful to digital library, we hope to clarify the field and facilitate designers' choices.
Fichier principal
Vignette du fichier
icadl2011_submission_101.pdf (2.54 Mo) Télécharger le fichier
Origin : Files produced by the author(s)
Loading...

Dates and versions

lirmm-00585390 , version 1 (12-04-2011)
lirmm-00585390 , version 2 (20-07-2011)

Identifiers

  • HAL Id : lirmm-00585390 , version 2

Cite

Guillaume Artignan, Mountaz Hascoët. Visual Analysis of Clustering Algorithms A Methodology and a Case Study. RR-11015, 2011. ⟨lirmm-00585390v2⟩
144 View
165 Download

Share

Gmail Facebook X LinkedIn More