Exploiting parallelism to support scalable hierarchical clustering.

A distributed memory parallel version of the group average hierarchical agglomerative clustering algorithm is proposed to enable scaling the document clustering problem to large collections. Using standard message passing operations reduces interprocess communication while maintaining efficient load...

Descripción completa

Detalles Bibliográficos
Publicado en:Journal of the American Society for Information Science & Technology Vol. 58; no. 8; pp. 1207 - 1222
Autores principales: Cathey RJ, Jensen EC, Beitzel SM, Frieder O, Grossman D
Formato: algorithm equations & formulas research tables/charts Journal Article
Publicado: Wiley-Blackwell Jun2007
Acceso en línea:Ver este registro en EBSCOhost