Assumptions in regression, Applied Statistics

Assignment Help:

Assumptions in Regression

To understand the properties underlying the regression line, let us go back to the example of model exam and main exam. Now we can find an estimate of a student's main exam points, if we also know his or her points on the model exam. As we have stated, a student with score of 85 in the model exam should receive points for the main exam in the vicinity of 75 to 95.

If we knew the model exam scores of all students along with their main exam scores, we would then have the population of values. The mean and the variance of the population of the model exam would be μx and σx2 and respectively. The measurements for the main exam points are  μy  and  σy2 .

The assumptions in regression are:

  1. The relationship between the distributions X and Y is linear, which implies the formula E(Y|X=x) = A + Bx at any given value of X = x.

  2. At each X, the distribution of Yx is normal, and the variances  σx2  are equal. This implies that E's have the same variance,  σ2.

  3. The Y-values are independent of each other.

  4. No assumption is made regarding the distribution of X.

    Since we do not have all of the students' course points and main exam points we must estimate the regression line E(Y|X = x) = A + BX.

    The figure shows a line that has been constructed on the scatter diagram. Note that the line seems to be drawn through the collective mid-point of the plotted points. The term  2148_simple linear regression.png  is the estimate of the true mean of Y's at any particular X = x.

    Figure 8

    682_assumptions in regression.png

Related Discussions:- Assumptions in regression

Regression coefficient, Regression Coefficient While analysing regressi...

Regression Coefficient While analysing regression in two related series, we calculate their regression coefficients also. There are two regression coefficients like two regress

Residual, regression line drawn as Y=C+1075x, when x was 2, and y was 239, ...

regression line drawn as Y=C+1075x, when x was 2, and y was 239, given that y intercept was 11. calculate the residual

Coefficient of determination, Coefficient of Determination The c...

Coefficient of Determination The coefficient of determination is given by r 2 i.e., the square of the correlation coefficient. It explains to what extent the variation

Multiple regressions, A sample of 43 houses that were purchased in the Sout...

A sample of 43 houses that were purchased in the Southern California town Monrovia within a month was collected. We are interested in the study of the relationships between Price a

Initial centroids data set, Find unlabeled data set test.txt and initial ...

Find unlabeled data set test.txt and initial centroids data set centroids.txt in the archive, both files have the following format: [attribute1_value attribute2_value ...

Bernoulli''s theorem, Bernoulli's Theorem If a trial of an experiment c...

Bernoulli's Theorem If a trial of an experiment can result in success  with probability p and failure with probability q (i.e.1-p) the probability of exactly r success in n tri

Median for ungrouped data, If the data set contains an odd number of items,...

If the data set contains an odd number of items, the middle item of the array is the median. If there is an even number of items, the median is the average of the two items. If the

Quantitative Business Analysis, Motion Picture Industry (95 Points) The m...

Motion Picture Industry (95 Points) The motion picture industry is a competitive business. More than 50 studios produce a total of 300 to 400 new motion pictures each year, and t

Write Your Message!

Captcha
Free Assignment Quote

Assured A++ Grade

Get guaranteed satisfaction & time on delivery in every assignment order you paid with us! We ensure premium quality solution document along with free turntin report!

All rights reserved! Copyrights ©2019-2020 ExpertsMind IT Educational Pvt Ltd