Explain ridge regression, Applied Statistics

Assignment Help:

Using log(x1), log(x2) and log(x3) as the predictors, do pair wise scatterplots of all pairs of variables (including the response) and comment (use the pairs function). Do you think that multi collinearity might be a problem with these data?

Plot the ridge trace for a grid of 50 values for the shrinkage parameter  over the range [0; 1]. Based on this plot suggest a reasonable value for . Find the estimates of the coecients for a ridge re gression with your chosen value of  (using centred and scaled predictors).

(The following question is based on Exercise 8.5 of Myers (1990), Classical and Modern Regression with Applications (Second Edition)," Duxbury).

With centred and scaled predictor variables, the ridge regression estimator for the coecients of the predictors is where y is the vector of responses, X is the design matrix for the centred and scaled predictors, is

1709_basic linear models.png

the shirnkage parameter and I denotes the identity matrix. We write n for the number of observations and k for the number of predictors. Writing biR for the ith component of bR, we will prove in this question that where 2 is the variance of the responses, and vi, i = 1,.......k are the eigenvalues of XTX. The di erent parts of the question below lead you through the proof.

735_basic linear models1.png

(a) Write XTX = QDQT for the eigenvalue decomposition of XTX, where D = diag(v1,........vk) is the diagonal matrix of eigenvalues and Q is an orthogonal matrix (QTQ = I) where the columns are the eigenvectors of XTX. Show that XTX +I = Q(D+I)QT .

2344_basic linear models2.png

where V ar(bR) denotes the covariance matrix of bR. (Hint: recall the result from basic linear models that if Y is a k  1 random vector with V ar(Y ) = V and if A is a k  k matrix and Z = AY then V ar(Z) = AV AT ).


Related Discussions:- Explain ridge regression

Uses of arthematic mean, give me question on mean is the aimplest average t...

give me question on mean is the aimplest average to understand and easy to compute

Regression analysis , The data used is from a statistical software Minitab;...

The data used is from a statistical software Minitab; London.MPJ is the file that consists of 1519 households drawn from 1980 - 1982 British Family Expenditure Surveys. Data that i

Normal curve applications, Replacement times for TV sets are normally distr...

Replacement times for TV sets are normally distributed with a mean of 8.2 years and a standard deviation of 1.1 years. Find the replacement time that separates the top 20% from the

Correlation, Correlation The board of directors of Bata Company is face...

Correlation The board of directors of Bata Company is faced with the problem of estimating what the annual sales might be in a shop to be opened in Bagpur where Bata has not op

Chi square test as a distributional goodness of fit, Chi Square Test as a D...

Chi Square Test as a Distributional Goodness of Fit In day-to-day decision making managers often come across situations wherein they are in a state of dilemma about the applica

Index Number of formulae, discuss the mathematical test of adequacy of inde...

discuss the mathematical test of adequacy of index number of formulae. prove algebraically that the laspeyre, paasche and fisher price index formulae satisfies this test. What is

Physics, fixed capacitor and variable capacitor

fixed capacitor and variable capacitor

Age at first marrage, get a questionnaire that captured age at first marria...

get a questionnaire that captured age at first marriage

Prediction, Differentiate between prediction, projection and forecasting.

Differentiate between prediction, projection and forecasting.

Write Your Message!

Captcha
Free Assignment Quote

Assured A++ Grade

Get guaranteed satisfaction & time on delivery in every assignment order you paid with us! We ensure premium quality solution document along with free turntin report!

All rights reserved! Copyrights ©2019-2020 ExpertsMind IT Educational Pvt Ltd