Warning: Undefined array key "mm" in /www/wwwroot/www.ai-bt.com/si.php on line 10 Deprecated: trim(): Passing null to parameter #1 ($string) of type string is deprecated in /www/wwwroot/www.ai-bt.com/si.php on line 10 Local optima in K-means clustering: what you don't know may hurt you.

Literature DB >> 14596492

Local optima in K-means clustering: what you don't know may hurt you.

Abstract

The popular K-means clustering method, as implemented in 3 commercial software packages (SPSS, SYSTAT, and SAS), generally provides solutions that are only locally optimal for a given set of data. Because none of these commercial implementations offer a reasonable mechanism to begin the K-means method at alternative starting points, separate routines were written within the MATLAB (Math-Works, 1999) environment that can be initialized randomly (these routines are provided at the end of the online version of this article in the PsycARTICLES database). Through the analysis of 2 empirical data sets and 810 simulated data sets, it is shown that the results provided by commercial packages are most likely locally optimal. These results suggest the need for some strategy to study the local optima problem for a specific data set or to identify methods for finding "good" starting values that might lead to the best solutions possible.

Entities: Gene

Mesh：

Year: 2003 PMID： 14596492 DOI： 10.1037/1082-989X.8.3.294

Source DB: PubMed Journal: Psychol Methods ISSN： 1082-989X

Keyword Cloud
Cited

31 in total

1. A model-based cluster analysis approach to adolescent problem behaviors and young adult outcomes.

Authors: Eun Young Mun; Michael Windle; Lisa M Schainker
Journal: Dev Psychopathol Date: 2008

2. Modeling differences in the dimensionality of multiblock data by means of clusterwise simultaneous component analysis.

Authors: Kim De Roover; Eva Ceulemans; Marieke E Timmerman; John B Nezlek; Patrick Onghena
Journal: Psychometrika Date: 2013-01-25 Impact factor: 2.500

Local optima in K-means clustering: what you don't know may hurt you.

1. A model-based cluster analysis approach to adolescent problem behaviors and young adult outcomes.

2. Modeling differences in the dimensionality of multiblock data by means of clusterwise simultaneous component analysis.

3. Taxicab Correspondence Analysis.

4. A Note on Using the Adjusted Rand Index for Link Prediction in Networks.

5. Local Optima in Mixture Modeling.

6. A comparison of latent class, K-means, and K-median methods for clustering dichotomous data.

7. Psychosocial costs of racism to Whites: Understanding patterns among university students.

8. KSC-N: Clustering of Hierarchical Time Profile Data.

9. Identifying subtypes of criminal psychopaths: A replication and extension.

10. Psychopathy Subtypes among African American County Jail Inmates.