Transformation of data, Applied Statistics

Assignment Help:

PCA is a linear transformation that transforms the data to a new coordinate system such that the greatest variance by any projection of the data comes to lie on the first coordinate (called the first principal component), the second greatest variance on the second coordinate, and so on. The PCA can be used for dimensionality reduction in a dataset while retaining those characteristics of the dataset that contribute most to its variance, by keeping lower-order principal components and ignoring higher-order ones. Such low-order components often contain the "most important" aspects of the data. But this is not necessarily the case, depending on the application. Let p and tn denote respectively the original and reduced number of variables. The original variables are denoted X. In the simplest case our measure of accuracy of reconstruction is the sum ofp squared multiple correlations between X-variables and the predictions of X made froin the factors. In the more general case we can weight each squared multiple correlation by the variance of the corresponding X-variable.

Since we can set those variances ourselves by multiplying scores on each variable,by any constant we choose, this amounts to the ability to assign any weights we choose to the different variables.


Related Discussions:- Transformation of data

Statistics to support learning, Scenario: Many of the years 5 and year 6 l...

Scenario: Many of the years 5 and year 6 learners' at Woodlands Park School were excited about being chosen for the cross-country team.  Every day, they were able to run laps of t

Probability theory, Origin and Development of probability Theory: The c...

Origin and Development of probability Theory: The credit for origin and development of probability goes to the European gamblers of 17 th century. They  used to gamble  on gam

Simple random sampling, Simple Random Sampling In Simple Random Sampli...

Simple Random Sampling In Simple Random Sampling each possible sample has an equal chance of being selected. Further, each item in the entire population also has an equal chan

Find relative maxima and minima, Q. Find relative maxima and minima? Wh...

Q. Find relative maxima and minima? When finding relative maxima and minima in the Chapters absolute extrema problem, don't forget to use the first or second derivative test to

Chi square test, who invented the chi square test and why? what is central ...

who invented the chi square test and why? what is central chi square and non central chi square test? what is distribution free statistics? what are the conditions when the chi squ

Statistical procedures - estimation of a mean, Old Faithful Geyser in Yello...

Old Faithful Geyser in Yellowstone National Park derives its names and fame from the regularity (and beauty) of its eruptions. Rangers usually post the predicted times of eruptions

Simple linear regression model, The data le for this assignment is brain-b...

The data le for this assignment is brain-body-wts.txt, which lists the averages brain weights (gm) and body weights (kg) for a number of animal species. Your task is to t an appr

Correlation, prove that coefficient of correlation lies between -1 and+1

prove that coefficient of correlation lies between -1 and+1

Write Your Message!

Captcha
Free Assignment Quote

Assured A++ Grade

Get guaranteed satisfaction & time on delivery in every assignment order you paid with us! We ensure premium quality solution document along with free turntin report!

All rights reserved! Copyrights ©2019-2020 ExpertsMind IT Educational Pvt Ltd