Explain ridge regression, Applied Statistics

Assignment Help:

Using log(x1), log(x2) and log(x3) as the predictors, do pair wise scatterplots of all pairs of variables (including the response) and comment (use the pairs function). Do you think that multi collinearity might be a problem with these data?

Plot the ridge trace for a grid of 50 values for the shrinkage parameter  over the range [0; 1]. Based on this plot suggest a reasonable value for . Find the estimates of the coecients for a ridge re gression with your chosen value of  (using centred and scaled predictors).

(The following question is based on Exercise 8.5 of Myers (1990), Classical and Modern Regression with Applications (Second Edition)," Duxbury).

With centred and scaled predictor variables, the ridge regression estimator for the coecients of the predictors is where y is the vector of responses, X is the design matrix for the centred and scaled predictors, is

1709_basic linear models.png

the shirnkage parameter and I denotes the identity matrix. We write n for the number of observations and k for the number of predictors. Writing biR for the ith component of bR, we will prove in this question that where 2 is the variance of the responses, and vi, i = 1,.......k are the eigenvalues of XTX. The di erent parts of the question below lead you through the proof.

735_basic linear models1.png

(a) Write XTX = QDQT for the eigenvalue decomposition of XTX, where D = diag(v1,........vk) is the diagonal matrix of eigenvalues and Q is an orthogonal matrix (QTQ = I) where the columns are the eigenvectors of XTX. Show that XTX +I = Q(D+I)QT .

2344_basic linear models2.png

where V ar(bR) denotes the covariance matrix of bR. (Hint: recall the result from basic linear models that if Y is a k  1 random vector with V ar(Y ) = V and if A is a k  k matrix and Z = AY then V ar(Z) = AV AT ).


Related Discussions:- Explain ridge regression

What is the p-value, Use the information given below to find the P-value. ...

Use the information given below to find the P-value. Also, use a 0.05 significance level and state the conclusion about the null hypothesis (reject the null hypothesis or fail to

Utility index , If the economy does well, the investor's wealth is 2 and if...

If the economy does well, the investor's wealth is 2 and if the economy does poorly the investor's wealth is 1. Both outcomes are equally likely. The investor is offered to invest

Statisttics., Explain any two applications of statistics

Explain any two applications of statistics

Regression analysis, Meaning and Definitions of Regression The dictiona...

Meaning and Definitions of Regression The dictionary meaning of regression is just opposite the meaning of progression. Progression means to move forward while regression means

ANOVA, Your company operates a machine shop, and, having heard you had expe...

Your company operates a machine shop, and, having heard you had experience in statistics and design of experiments, consulted you for your opinion on an experiment they want to run

Two methods of isolating trend values in a time series, a) What is meant by...

a) What is meant by secular trend? Discuss any two methods of isolating trend values in a time series.

Non-sampling errors, Statistics Can Lead to Errors The use of st...

Statistics Can Lead to Errors The use of statistics can often lead to wrong conclusions or wrong estimates. For example, we may want to find out the average savings by i

Chi square test as a distributional goodness of fit, Chi Square Test as a D...

Chi Square Test as a Distributional Goodness of Fit In day-to-day decision making managers often come across situations wherein they are in a state of dilemma about the applica

Applications of standard error, Applications of Standard Error   ...

Applications of Standard Error   Standard Error is used to test whether the difference between the sample statistic and the population parameter is significant or is d

Compare the t interval with the bootstrap interval, Jocko's Garage has been...

Jocko's Garage has been accused of insurance fraud. Data on estimates made by Jocko and another garage were obtained for 10 damaged vehicles (available in 'jockogarage.txt'). Here

Write Your Message!

Captcha
Free Assignment Quote

Assured A++ Grade

Get guaranteed satisfaction & time on delivery in every assignment order you paid with us! We ensure premium quality solution document along with free turntin report!

All rights reserved! Copyrights ©2019-2020 ExpertsMind IT Educational Pvt Ltd