Unsupervised Clustering on Multi-Components Datasets: Applications on Images and Astrophysics Data - C2S Accéder directement au contenu
Communication Dans Un Congrès Année : 2008

Unsupervised Clustering on Multi-Components Datasets: Applications on Images and Astrophysics Data

Résumé

This paper proposes an original approach to cluster multi-component data sets with an estimation of the number of clusters. From the construction of a minimal spanning tree with Prim's algorithm and the assumption that the vertices are approximately distributed according to a Poisson distribution, the number of clusters is estimated by thresholding the Prim's trajectory. The corresponding cluster centroids are then computed in order to initialize the Generalized Lloyd's algorithm, also known as K-means, which allows to circumvent initialization problems. Metrics used for measuring similarity between multi-dimensional data points are based on symmetrical divergences. The use of these informational divergences together with the proposed method lead to better results than some other clustering methods in the framework of astrophysical data processing. An application of this method in the multi-spectral imagery domain with a satellite view of Paris is also presented.
Fichier principal
Vignette du fichier
GallMC08lausanne.pdf (611.43 Ko) Télécharger le fichier
Origine : Fichiers produits par l'(les) auteur(s)

Dates et versions

hal-00339542 , version 1 (18-11-2008)

Identifiants

  • HAL Id : hal-00339542 , version 1

Citer

Laurent Galluccio, Olivier J.J. Michel, Pierre Comon. Unsupervised Clustering on Multi-Components Datasets: Applications on Images and Astrophysics Data. EUSIPCO 2008 - 16th European Signal Processing Conference, Aug 2008, Lausanne, Switzerland. pp.P4-1. ⟨hal-00339542⟩
358 Consultations
111 Téléchargements

Partager

Gmail Facebook X LinkedIn More