Best subsets regression, Advanced Statistics

Assignment Help:

In the time series plot and scatter graphs there were many outliers that were clearly visible. These have been removed to identify if they were influential or had high leverage and in order to see if the multiple regression model assumptions have been met.

Below are the rows of the outliers that I removed out of the 1519 observations:

77, 674, 448, 757, 317, 549, 1187, 1198, 26, 456, 405, 307, 1205, 1348, 611, 368, 309

Best Subsets Regression: wfood versus totexp, income, age, nk

Response is wfood

                                                                   t i

                                                                   o n

                                                                    t c

                                                                    e o a

                               Mallows                         x m g n

Vars  R-Sq  R-Sq(adj)       Cp         S             p e e k

   1  22.9       22.9     67.4            0.092326  X

   1   5.5        5.4      424.9           0.10222    X

   2  24.8       24.7     31.3            0.091236  X     X

   2  24.2       24.1     42.7           0.091572  X   X

   3  26.1       26.0      6.1            0.090461  X   X X

   3  24.8       24.7     32.3           0.091239  X X   X

   4  26.3       26.1      5.0            0.090397  X X X X

The best subset is a way of identifying which independent variable such as the totexp, income, age and nk are best suited to the regression model.  According to the results above income is the variable that has the highest Cp and the lowest R-squared value therefore it will be the variable that will be dropped to see if the data fits the model.


Related Discussions:- Best subsets regression

Pascal''s triangle, Pascal's triangle  is an arrangement of numbers describ...

Pascal's triangle  is an arrangement of numbers described by Pascal in his Traité du Triangle Arithmétique published in the year 1665 as 'The number in each cell is equal to in the

Daycare, facts and statistics about daycare

facts and statistics about daycare

Generalized additive model, The linear component ηi, de?ned just in the tra...

The linear component ηi, de?ned just in the traditional way: η i = x' 1 A monotone differentiable link function g that describes how E(Yi) = µi is related to the linear compon

Advanced managerial statistics, The objective of this assignment is to test...

The objective of this assignment is to test your understanding in the learning outcome (LO2) and learning outcome (LO3) and learning outcome (LO4). 1) This is a grouped assignme

Compound symmetry, Compound symmetry : The property possessed by the varian...

Compound symmetry : The property possessed by the variance-covariance matrix of the set of multivariate data when its chief diagonal elements are equal to each other, and in additi

Finite mixture distribution, The probability distribution which is a linear...

The probability distribution which is a linear function of the number of component probability distributions. This type of distributions is used to model the populations thought to

Relative poverty statistics, Relative poverty statistics is the statistics...

Relative poverty statistics is the statistics on the properties of populations falling below given fractions of average income which play a central role in debate of poverty. The

Explain negative hyper geometric distribution, Negative hyper geometric dis...

Negative hyper geometric distribution : In sampling without replacement from the population comprising of r elements of one kind and N - r of another, if two elements corresponding

Paired samples, Paired samples are the two samples of the observations wit...

Paired samples are the two samples of the observations with the characteristic feature with each of the observation in one sample have only one matching observation in the other s

To create a relative frequency histogram, The total amount of protein produ...

The total amount of protein produced by a dairy cow can be estimated from periodic testing of her milk.  The following are the total annual protein production values (lb) for 28 tw

Write Your Message!

Captcha
Free Assignment Quote

Assured A++ Grade

Get guaranteed satisfaction & time on delivery in every assignment order you paid with us! We ensure premium quality solution document along with free turntin report!

All rights reserved! Copyrights ©2019-2020 ExpertsMind IT Educational Pvt Ltd