Empirical Evaluation of Clustering Algorithms for Large Networks

Guillaume Artignan; Mountaz Hascoët

Rapport Année : 2011

Empirical Evaluation of Clustering Algorithms for Large Networks

(1) , (1)

Guillaume Artignan

Fonction : Auteur correspondant
PersonId : 862749

Connectez-vous pour contacter l'auteur

Hors Équipe

Mountaz Hascoët

Fonction : Auteur
PersonId : 837916

Hors Équipe

Résumé

Clustering is probably one of the most frequently used approaches when facing a scaling problem in large networks. In many situations, however, the choice of the most appropriate algorithm for clustering can turn into a real dilemma. Numerical criteria have been proposed to evaluate the quality of the results of clustering algorithms. However, so many different criteria have been proposed that the dilemma gets even worse. Most criteria reveal different aspects of the quality of the results and hide others. The aim of this paper is to help with the understanding of clustering and to facilitate the comparison and the choice of clustering algorithm for a given purpose. Our proposal consists of studying both quality evaluation criteria and clustering algorithms. We start by discussing a selected set of representative criteria, and further conduct a case study on a large set of real data, measuring not only the quality of different representative clustering algorithms but also the impact of each criterion on the ranking of the algorithms. By providing empirical results on several large-scale corpus of either inter-related documents or lexical networks, we hope to clarify the field and facilitate designers' choices.

Mots clés

component clustering networks quality visual analysis

Domaines

Mathématique discrète [cs.DM] Algorithme et structure de données [cs.DS]

Fichier principal

2011_rr_ag_mh.pdf (7.01 Mo)

Origine	Fichiers produits par l'(les) auteur(s)

Guillaume Artignan : Connectez-vous pour contacter le contributeur

https://hal-lirmm.ccsd.cnrs.fr/lirmm-00648389

Soumis le : lundi 5 décembre 2011-15:42:12

Dernière modification le : vendredi 24 mars 2023-14:52:55

Archivage à long terme le : mardi 6 mars 2012-02:35:52

Dates et versions

lirmm-00648389 , version 1 (05-12-2011)

Identifiants

HAL Id : lirmm-00648389 , version 1

Citer

Guillaume Artignan, Mountaz Hascoët. Empirical Evaluation of Clustering Algorithms for Large Networks. 2011. ⟨lirmm-00648389⟩

Exporter

BibTeX XML-TEI Dublin Core DC Terms EndNote DataCite

Collections

CNRS LIRMM HORSEQUIPE TDS-MACS LARA MIPS UNIV-MONTPELLIER

154 Consultations

227 Téléchargements

Empirical Evaluation of Clustering Algorithms for Large Networks

Résumé

Mots clés

Domaines

Dates et versions

Identifiants

Citer

Exporter

Collections

Partager