K-means cluster analysis, Advanced Statistics

Assignment Help:

K-means cluster analysis is the method of cluster analysis in which from an initial partition of observations into K clusters, each observation in turn is analysed and reassigned, if suitable, to a different cluster in an attempt to optimize some predefined numerical criterion that measures in some sense the 'quality' of cluster solution. Several such clustering criteria have been suggested, but the most usually used arise from considering the features of the within groups, between groups and whole matrices of sums of squares and the cross products (W, B, T) which can be described for every partition of the observations into the particular number of groups. The two most ordinary of the clustering criteria developing from these matrices are given as follows

minimization of trace W

minimization of determinant W

The first of these has tendency to produce the 'spherical' clusters, the second to produce clusters that all have same shape, though this will not necessarily be spherical in shape. 

 


Related Discussions:- K-means cluster analysis

Log-linear models, Log-linear models is the models for count data in which...

Log-linear models is the models for count data in which the logarithm of expected value of a count variable is modelled as the linear function of parameters; the latter represent

Dot plot, The more effective display than a number of other methods or tech...

The more effective display than a number of other methods or techniques, for instance, pie charts and bar charts, for displaying the quantitative data which are labeled. An instanc

Cure models, Models for the analysis of the survival times, or the time to ...

Models for the analysis of the survival times, or the time to event, data in which it is expected that a fraction of the subjects will not experience the event of interest. In a cl

Explain maternal mortality, Maternal mortality : The maternal death is the ...

Maternal mortality : The maternal death is the death of a woman while pregnant, delivering a baby or within 42 days of the termination of pregnancy, from any reason related to or a

Partial least squares, Partial least squares is an alternative to the mult...

Partial least squares is an alternative to the multiple regressions which, in spite of using the original q explanatory variables directly, constructs the new set of k regressor v

Describe respondent-driven sampling (rds), Respondent-driven sampling (RDS ...

Respondent-driven sampling (RDS ): The form of snowball sampling which starts with the recruitment of the small number of people in the target population to serve as the seeds. Aft

General household survey, It is the survey which is carried out in Great Br...

It is the survey which is carried out in Great Britain on a continuous basis since 1971. About 100 000 households are included in this sample every year. The main goal of the surve

Mardia''s multivariate normality test, Mardia's multivariate normality test...

Mardia's multivariate normality test is a test that a set of the multivariate data arise from the multivariate normal distribution against departures due to the kurtosis. The test

Decision Models., An oil company thinks that there is a 60% chance that the...

An oil company thinks that there is a 60% chance that there is oil in the land they own. Before drilling they run a soil test. When there is oil in the ground, the soil test comes

Chapter 7&8, Chapter 7 2. Describe the distribution of sample means (shape...

Chapter 7 2. Describe the distribution of sample means (shape, expected value, and standard error) for samples of n =36 selected from a population with a mean of µ = 100 and a sta

Write Your Message!

Captcha
Free Assignment Quote

Assured A++ Grade

Get guaranteed satisfaction & time on delivery in every assignment order you paid with us! We ensure premium quality solution document along with free turntin report!

All rights reserved! Copyrights ©2019-2020 ExpertsMind IT Educational Pvt Ltd