Implement a simple k-means method, Applied Statistics

Assignment Help:

There exists an unclassified data set with hidden data structures in it. The task in this assignment is to perform comprehensive Cluster Analysis in order to reveal the structures and similar data groups.

1. Implement a simple K-means method, which is able to handle real values data in attributes. Also you need to add functionality in your program that allows utilization of Euclidean, City Block, Euclidean Squared and Chebyshev distances. You are free to use any kind of weights (for feature or data instance) in the program if necessary.

2. Find unlabeled data set test.txt and initial centroids data set centroids.txt in the archive, both files have the following format: [attribute1_value attribute2_value ... attribute90_value]. The unlabeled data set includes 350 samples and the initial centroids set consists of 15 samples. Data instances in both files have 90 attributes.


Related Discussions:- Implement a simple k-means method

Evaluate central tendency and variability, Why are graphs and tables useful...

Why are graphs and tables useful when examining data? A researcher is comparing two middle school 7th grade classes. One class at one school has participated in an arts program

Rank correlation, Rank Correlation Sometimes the characteristics whose ...

Rank Correlation Sometimes the characteristics whose possible correlation is being investigated, cannot be measured but individuals can only be ranked on the basis of the chara

Probability, There are 15 types of ice cream: A,B,C,D,E,F,G,H,I,J,K,L,M,N, ...

There are 15 types of ice cream: A,B,C,D,E,F,G,H,I,J,K,L,M,N, and O. How many combinations are there to sample 5 flavors if you sample 1 flavor 4 times? How many combinations are t

Calculate the normal loss and abnormal loss, Chemical processors manufactur...

Chemical processors manufacture wondercool using two processes- mixing and distillation. The following details relate to the distillation process for a period. No opening work i

What are the charateristics of a population for which, what are characteris...

what are characteristics of a population for which it would be appropiate to use mean/median/mode

B) Distinguish between:, X 110 120 130 120 140 135 155 160 165 155 ...

X 110 120 130 120 140 135 155 160 165 155 Y 12 18 20 15 25 30 35 20 25 10

Association of attributes, In an examination 600 candidates appeared, boys ...

In an examination 600 candidates appeared, boys outnumbered girls by 16% of all candidates. number of passed candidates exceeded the number of failed candidates by 310. Boys failin

Ogive graphs, how many types of ogive are there

how many types of ogive are there

PERCENTAGES, CALCULATE THE PERCENTAGE OF REFUNDS EXPECTED TO EXCEED $1000 U...

CALCULATE THE PERCENTAGE OF REFUNDS EXPECTED TO EXCEED $1000 UNDER THE CURRENT WITHHOLDING GUIDELINES

Write Your Message!

Captcha
Free Assignment Quote

Assured A++ Grade

Get guaranteed satisfaction & time on delivery in every assignment order you paid with us! We ensure premium quality solution document along with free turntin report!

All rights reserved! Copyrights ©2019-2020 ExpertsMind IT Educational Pvt Ltd