K-means cluster analysis, Advanced Statistics

K-means cluster analysis is the method of cluster analysis in which from an initial partition of observations into K clusters, each observation in turn is analysed and reassigned, if suitable, to a different cluster in an attempt to optimize some predefined numerical criterion that measures in some sense the 'quality' of cluster solution. Several such clustering criteria have been suggested, but the most usually used arise from considering the features of the within groups, between groups and whole matrices of sums of squares and the cross products (W, B, T) which can be described for every partition of the observations into the particular number of groups. The two most ordinary of the clustering criteria developing from these matrices are given as follows

minimization of trace W

minimization of determinant W

The first of these has tendency to produce the 'spherical' clusters, the second to produce clusters that all have same shape, though this will not necessarily be spherical in shape. 


Posted Date: 7/30/2012 1:31:04 AM | Location : United States

Related Discussions:- K-means cluster analysis, Assignment Help, Ask Question on K-means cluster analysis, Get Answer, Expert's Help, K-means cluster analysis Discussions

Write discussion on K-means cluster analysis
Your posts are moderated
Related Questions
Johnson-Neyman technique:  The technique which can be used in the situations where analysis of the covariance is not valid because of the heterogeneity of slopes. With this method

The measure of the degree to which the particular model differs from the saturated model for the data set. Explicitly in terms of the likelihoods of the two models can be defined a

A statewide survey of 1,706 California adults’ residents include the following question: would you favor or oppose providing a path to citizenship for illegal immigrants in the U.S

Kappa coefficient : The chance corrected index of the agreement between, for instance, judgements and diagnoses made by the two raters. Calculated as the ratio of the noticed exces

The method of summarizing the large amounts of data by forming the frequency distributions, scatter diagrams, histograms, etc., and calculating statistics like means variances and

need answers to questions in book advanced and multivariate statistical methods

Matching distribution is  a probability distribution which arises in the following manner. Suppose that the set of n subjects, numbered 1; . . . ; n respectively, are arranged in

Multi-hit model is the model for a toxic response which results from the random occurrence of one or the more fundamental biological events. A response is supposed to be induced o

A value related with the square matrix which represents sums and products of its elements. For instance, if the matrix is   then the determinant of A (conventionally written as