K-means cluster analysis, Advanced Statistics

Assignment Help:

K-means cluster analysis is the method of cluster analysis in which from an initial partition of observations into K clusters, each observation in turn is analysed and reassigned, if suitable, to a different cluster in an attempt to optimize some predefined numerical criterion that measures in some sense the 'quality' of cluster solution. Several such clustering criteria have been suggested, but the most usually used arise from considering the features of the within groups, between groups and whole matrices of sums of squares and the cross products (W, B, T) which can be described for every partition of the observations into the particular number of groups. The two most ordinary of the clustering criteria developing from these matrices are given as follows

minimization of trace W

minimization of determinant W

The first of these has tendency to produce the 'spherical' clusters, the second to produce clusters that all have same shape, though this will not necessarily be spherical in shape. 

 


Related Discussions:- K-means cluster analysis

Explain Genstat, Genstat: The basic purpose piece of statistical software ...

Genstat: The basic purpose piece of statistical software for the management and the analysis of data. The package incorporates the wide variety of data handling events and a wi

Explain information theory., Information theory: This is the branch of app...

Information theory: This is the branch of applied probability theory applicable to various communication and signal processing problems in the field of engineering and biology. In

Generalized additive models, Models which make use of the smoothing techniq...

Models which make use of the smoothing techniques such as locally weighted regression to identify and represent the possible non-linear relationships between the explanatory and th

Explain prevalence, Prevalence : The measure of the number of people in a p...

Prevalence : The measure of the number of people in a population who have a certain disease at a given point in time. It c an be measured by two methods, as point prevalence and p

Bimodal distribution, Bimodal distribution : The probability distribution, ...

Bimodal distribution : The probability distribution, or we can simply say the frequency distribution, with two modes. Figure 15 shows the example of each of them

Matlab help, Need help with Matlab assignments.

Need help with Matlab assignments.

Dot plot, The more effective display than a number of other methods or tech...

The more effective display than a number of other methods or techniques, for instance, pie charts and bar charts, for displaying the quantitative data which are labeled. An instanc

Double sampling, The procedure in which initially the sample of subjects is...

The procedure in which initially the sample of subjects is selected for generating the auxillary information only, and then the second sample is selected in which the variable of i

Battery reduction, Battery reduction : A common term for reducing the numbe...

Battery reduction : A common term for reducing the number of variables of the interest in a study for the purposes of study and perhaps later data collection. For instance, an over

Explain labour force survey, Labour force survey : This survey carried out ...

Labour force survey : This survey carried out in the UK on the quarterly basis since the spring of year 1992. It covers 60 000 households and gives labour force and other detail

Write Your Message!

Captcha
Free Assignment Quote

Assured A++ Grade

Get guaranteed satisfaction & time on delivery in every assignment order you paid with us! We ensure premium quality solution document along with free turntin report!

All rights reserved! Copyrights ©2019-2020 ExpertsMind IT Educational Pvt Ltd