Implement a simple k-means method, Applied Statistics

Assignment Help:

There exists an unclassified data set with hidden data structures in it. The task in this assignment is to perform comprehensive Cluster Analysis in order to reveal the structures and similar data groups.

1. Implement a simple K-means method, which is able to handle real values data in attributes. Also you need to add functionality in your program that allows utilization of Euclidean, City Block, Euclidean Squared and Chebyshev distances. You are free to use any kind of weights (for feature or data instance) in the program if necessary.

2. Find unlabeled data set test.txt and initial centroids data set centroids.txt in the archive, both files have the following format: [attribute1_value attribute2_value ... attribute90_value]. The unlabeled data set includes 350 samples and the initial centroids set consists of 15 samples. Data instances in both files have 90 attributes.


Related Discussions:- Implement a simple k-means method

Factor loadings matrix, As we stated above, we start factor analysis with p...

As we stated above, we start factor analysis with principal component analysis, but we quickly diverge as we apply the a priori knowledge we brought to the problem. This knowled

Bernoulli trial, Statistician is searching the \home ground" effect and is ...

Statistician is searching the \home ground" effect and is studying 20 football games, of which 14 were won by the home team and 6 by the visitors. Therefore the game is a Bernoulli

Standard deviation and variance, Explanation of standard deviation and vari...

Explanation of standard deviation and variance Describe the importance of standard deviation and variance, what they calculate and why they are required. Importance of char

Correlation, Correlation The correlation is commonly used and a useful...

Correlation The correlation is commonly used and a useful statistic used to describe the degree of the relationships between two or more variables. Pearson's correlation refle

Simple random sampling, Simple Random Sampling In Simple Random Sampli...

Simple Random Sampling In Simple Random Sampling each possible sample has an equal chance of being selected. Further, each item in the entire population also has an equal chan

Half of market share, Your company has developed a new product .Your compan...

Your company has developed a new product .Your company is a reputed company with 50% market share of same range of products. Your competitors also come with their new products equa

Multivariate analysis of variance, Multivariate analysis of variance (MANOV...

Multivariate analysis of variance (MANOVA) is a technique to assess group differences across multiple metric dependent variables simultaneously, based on a set of categorical (non-

Decision making ., it is said that management is equivalent to decision mak...

it is said that management is equivalent to decision making? do you agree? explain

Simulation, Simulation When decisions are to be taken under conditions ...

Simulation When decisions are to be taken under conditions of uncertainty, simulation can be used. Simulation as a quantitative method requires the setting up of a mathematical

Harmonic mean, The Harmonic Mean is based on the reciprocals of numbers ave...

The Harmonic Mean is based on the reciprocals of numbers averaged. It is defined as the reciprocal of the arithmetic mean of the reciprocal of the given individual observations. Th

Write Your Message!

Captcha
Free Assignment Quote

Assured A++ Grade

Get guaranteed satisfaction & time on delivery in every assignment order you paid with us! We ensure premium quality solution document along with free turntin report!

All rights reserved! Copyrights ©2019-2020 ExpertsMind IT Educational Pvt Ltd