K-means cluster analysis, Advanced Statistics

Assignment Help:

K-means cluster analysis is the method of cluster analysis in which from an initial partition of observations into K clusters, each observation in turn is analysed and reassigned, if suitable, to a different cluster in an attempt to optimize some predefined numerical criterion that measures in some sense the 'quality' of cluster solution. Several such clustering criteria have been suggested, but the most usually used arise from considering the features of the within groups, between groups and whole matrices of sums of squares and the cross products (W, B, T) which can be described for every partition of the observations into the particular number of groups. The two most ordinary of the clustering criteria developing from these matrices are given as follows

minimization of trace W

minimization of determinant W

The first of these has tendency to produce the 'spherical' clusters, the second to produce clusters that all have same shape, though this will not necessarily be spherical in shape. 

 


Related Discussions:- K-means cluster analysis

Treatment allocation ratio, Treatment allocation ratio is the ratio of the...

Treatment allocation ratio is the ratio of the number of subjects allocated to the two treatments in a clinical trial. The equal allocation is most usual in practice, but it might

Adjusted r-squared, R-squared is regarded as the coefficient of determinati...

R-squared is regarded as the coefficient of determination and is used to give the proportion of the fluctuation of the variance of one variable to another variable. R-squared also

Codominance, Codominance : The relationship between genotype at the locus a...

Codominance : The relationship between genotype at the locus and a phenotype to which it in?uences. If an individuals with heterozygote (such as, AB) genotype is phenotypically dif

Analysis of variance, Thomas Economic Forecasting, Inc. and Harmon Economet...

Thomas Economic Forecasting, Inc. and Harmon Econometrics have the same mean error in forecasting the stock market over the last ten years. However, the standard deviation for Thom

Line-intersect sampling, Line-intersect sampling is a technique of unequal...

Line-intersect sampling is a technique of unequal probability sampling for selecting the sampling units in the geographical area. A sample of lines is drawn in a study area and, w

Probability, show all the ways in which 3 games of football can be conclude...

show all the ways in which 3 games of football can be concluded(it can be a win W,a loss L,or a draw X)

Homework help, Q1: The growth in bad debt expense for Aptara Pvt. Ltd. Comp...

Q1: The growth in bad debt expense for Aptara Pvt. Ltd. Company over the last 20 years is as follows. 1997 0.11 1998 0.09 1999 0.08 2000 0.08 2001 0.1 2002 0.11 2003 0.12 2004 0.1

Machine learning, Machine learning  is a term which literally means the ab...

Machine learning  is a term which literally means the ability of a machine to recognize patterns which have occurred repetitively and to improve its performance based on the past

Describe indirect least squares, Indirect least squares: An estimation tech...

Indirect least squares: An estimation technique used in the fitting of structural equation models. Commonly least squares are first used to estimate reduced form parameters. Usi

Incidental parameter problem, Incidental parameter problem is a problem wh...

Incidental parameter problem is a problem which sometimes occurs when the number of parameters increases in the tandem with the number of observations. For instance, models for pa

Write Your Message!

Captcha
Free Assignment Quote

Assured A++ Grade

Get guaranteed satisfaction & time on delivery in every assignment order you paid with us! We ensure premium quality solution document along with free turntin report!

All rights reserved! Copyrights ©2019-2020 ExpertsMind IT Educational Pvt Ltd