Assumptions in regression, Applied Statistics

Assignment Help:

Assumptions in Regression

To understand the properties underlying the regression line, let us go back to the example of model exam and main exam. Now we can find an estimate of a student's main exam points, if we also know his or her points on the model exam. As we have stated, a student with score of 85 in the model exam should receive points for the main exam in the vicinity of 75 to 95.

If we knew the model exam scores of all students along with their main exam scores, we would then have the population of values. The mean and the variance of the population of the model exam would be μx and σx2 and respectively. The measurements for the main exam points are  μy  and  σy2 .

The assumptions in regression are:

  1. The relationship between the distributions X and Y is linear, which implies the formula E(Y|X=x) = A + Bx at any given value of X = x.

  2. At each X, the distribution of Yx is normal, and the variances  σx2  are equal. This implies that E's have the same variance,  σ2.

  3. The Y-values are independent of each other.

  4. No assumption is made regarding the distribution of X.

    Since we do not have all of the students' course points and main exam points we must estimate the regression line E(Y|X = x) = A + BX.

    The figure shows a line that has been constructed on the scatter diagram. Note that the line seems to be drawn through the collective mid-point of the plotted points. The term  2148_simple linear regression.png  is the estimate of the true mean of Y's at any particular X = x.

    Figure 8

    682_assumptions in regression.png

Related Discussions:- Assumptions in regression

PERCENTAGES, CALCULATE THE PERCENTAGE OF REFUNDS EXPECTED TO EXCEED $1000 U...

CALCULATE THE PERCENTAGE OF REFUNDS EXPECTED TO EXCEED $1000 UNDER THE CURRENT WITHHOLDING GUIDELINES

Genmod procedure, The following dataset is from a study of the effects of s...

The following dataset is from a study of the effects of second hand smoking in Baltimore, MD, and Washington, DC. For the 25 children involved in this study the outcome variable is

Time series, Measurement of trend , least square method

Measurement of trend , least square method

Central tendency, Definition of Central Tendency The central tendency o...

Definition of Central Tendency The central tendency of a variable means a typical value around which other values tend to concentrate which can be measured. Such concentration

Probability and expectation, Ten balls are put in 6 slots at random.Then ex...

Ten balls are put in 6 slots at random.Then expected total number of balls in the two extreme slots

Convenience sampling, Convenience Sampling It means a convenient sample...

Convenience Sampling It means a convenient sample is obtained by selecting convents units from the universe. Convenient sample is also known as chunk. It   means a fraction of

Coefficient of variation, Coefficient of Variation or C.V. To compare t...

Coefficient of Variation or C.V. To compare the variability between or more series, coeffiecnt of variation is used, it is relative measure of dispersion, it innovated and used

Standard deviation for grouped data, Grouped data  For ...

Grouped data  For grouped data, the formula applied is  σ = Where f = frequency of the variable, μ= population mea

Logistic regression model, A marketing research firm was engaged by an auto...

A marketing research firm was engaged by an automobile manufacturer to conduct a pilot study to examine the feasibility of using logistic regression for ascertaining the likelihood

Write Your Message!

Captcha
Free Assignment Quote

Assured A++ Grade

Get guaranteed satisfaction & time on delivery in every assignment order you paid with us! We ensure premium quality solution document along with free turntin report!

All rights reserved! Copyrights ©2019-2020 ExpertsMind IT Educational Pvt Ltd