Linear regression assignment help, Advanced Statistics

Assignment Help:

Using World Bank (2004) World Development Indicators; Washington: International Bank for Reconstruction & Development/ The World Bank, located in the reference section of the Learning Centre (Stats 330.9 WOR), collect a sample of data comprising cereal yield in kg/ha and fertilizer consumption in hundreds of grams /ha of arable land for the period 2000-2002 from 25 countries around the world. These data can be found in Table 3.3 pp123-126 and Table 3.2 pp 119-122 respectively.

For the regression analysis, use cereal yield as the dependent variable (Y) and fertilizer consumption as the independent variable (X). Enter the two variables into an SPSS file and carry out the following exercise: 

1.         Regress cereal yield (Y) on fertilizer consumption (X).

2.         Produce a plot of the 30 observations, the calculated regression line and the 95% confidence limits.                                                               

3.         What is the correlation between cereal yield and fertilizer consumption?             

4.          State whether the modeled regression relationship is significant.

 5.         Examine the plotted residuals and attempt to explain two of the extreme positive and negative values (max 300 words).                       

6.         Calculate the runs test and the Durbin-Watson statistic on the residuals and indicate whether auto-correlation is present at the 0.05 significance level.

 

1)

We run the regression of cereal yield on fertilizer consumption. The fitted regression line is given by:

 

cereal yield= 2254.069+0.253 * fertilizer consumption

 

 

Regression

 

Variables Entered/Removed

Model

Variables Entered

Variables Removed

Method

1

fertilizer consumptiona

.

Enter

a. All requested variables entered.

 

b. Dependent Variable: cereal yields

 

 

Residuals Statisticsa

 

Minimum

Maximum

Mean

Std. Deviation

N

Predicted Value

2254.07

3054.49

2400.90

195.404

25

Residual

-2.251E3

4481.879

.000

1822.796

25

Std. Predicted Value

-.751

3.345

.000

1.000

25

Std. Residual

-1.209

2.407

.000

.979

25

a. Dependent Variable: cereal yields

 

 

 

 

3) The Correlation coefficient is 0.01

 

4) From the ANOVA table we see that the regression is not significant at 5% level of significance.

 

ANOVAb

Model

Sum of Squares

df

Mean Square

F

Sig.

1

Regression

916388.422

1

916388.422

.264

.612a

Residual

7.974E7

23

3467045.255

 

 

Total

8.066E7

24

 

 

 

a. Predictors: (Constant), fertilizer consumption

 

 

b. Dependent Variable: cereal yields

 

 

 

 

6)

The Durbin-Watson statistic is 1.588 which is close to 2 indicating there may be no or little positive autocorrelation

Model Summary

Model

R

R Square

Adjusted R Square

Std. Error of the Estimate

Durbin-Watson

1

.107a

.011

-.032

1862.000

1.588

a. Predictors: (Constant), fertilizer consumption

 

b. Dependent Variable: cereal yields

 

 

However we next perform the run test which clearly implies that the residuals are independent at 5% level.

 

NPar Tests 

 

Runs Test

 

Standardized Residual

Test Valuea

-.15308

Cases < Test Value

12

Cases >= Test Value

13

Total Cases

25

Number of Runs

15

Z

.417

Asymp. Sig. (2-tailed)

.676

a. Median

 

 

 

7)

For less developed countries the intercept term will be very low as compared to high developed countries.

Moreover the slope of the fertilizer consumption will also be low in less developed countries indicating slow growth rate of cereal yield.

 

7.         What differences would you expect to find between less developed and more developed countries in terms of the relationship between cereal yield and fertilizer consumption?


Related Discussions:- Linear regression assignment help

Goodmanand kruskal measures of association, Goodmanand kruskal measures of ...

Goodmanand kruskal measures of association is the measures of associations which are useful in the situation where two categorical variables cannot be supposed to be derived from

Describe multiple imputation, Multiple imputation : The Monte Carlo techniq...

Multiple imputation : The Monte Carlo technique in which missing values in the data set are replaced by m> 1 simulated versions, where m is usually small (say 3-10). Each of simula

Logistic regression - computing log odds without probabiliti, Please help w...

Please help with following problem: : Let’s consider the logistic regression model, which we will refer to as Model 1, given by log(pi / [1-pi]) = 0.25 + 0.32*X1 + 0.70*X2 + 0.

Expectaton, sales per day for a product are as follows: x= 10, 11, 12, 13 (...

sales per day for a product are as follows: x= 10, 11, 12, 13 (p)= 0.2, 0.4, 0.3, 0.1 obtain mean and variance of daily sale. if the profit is described by the following equation p

MC0074 –Statistical and Numerical methods using C++, Write a c++ program to...

Write a c++ program to find the sum of 0.123 ? 103 and 0.456 ? 102 and write the result in three significant digits

Histogram, Histogram is the graphical representation of the set of observat...

Histogram is the graphical representation of the set of observations in which class frequencies are represented by the regions of rectangles centred on the class interval. If the f

Resentful demoralization, Resentful demoralization is the possible phenome...

Resentful demoralization is the possible phenomenon in the clinical trials and intervention studies in which comparison groups not attaining a perceived desirable treatment become

Component bar chart, Component bar chart : A bar chart which shows the comp...

Component bar chart : A bar chart which shows the component parts of the aggregate represented by the whole length of the bar. The component parts are shown as the sectors of bar w

Computer-assisted interviews, Computer-assisted interviews : A method or te...

Computer-assisted interviews : A method or technique of interviewing subjects in which the interviewer reads the question from the computer screen instead of the printed page, and

Fisher''s exact test, The alternative process to make use of the chi-square...

The alternative process to make use of the chi-squared statistic for assessing the independence of the two variables forming a two-by-two contingency table particularly when expect

Write Your Message!

Captcha
Free Assignment Quote

Assured A++ Grade

Get guaranteed satisfaction & time on delivery in every assignment order you paid with us! We ensure premium quality solution document along with free turntin report!

All rights reserved! Copyrights ©2019-2020 ExpertsMind IT Educational Pvt Ltd