Scatter diagram - correlation analysis, Applied Statistics

Assignment Help:

Scatter Diagram

The first step in correlation analysis is to visualize the relationship. For each unit of observation in correlation analysis there is a pair of numerical values. One is considered the independent variable; the other is considered dependent upon it and is called the dependent variable. One of the easiest ways of studying the correlation between the two variables is with the help of a scatter diagram.

A scatter diagram can give us two types of information. Visually, we can look for patterns that indicate whether the variables are related. Then, if the variables are related, we can see what kind of line, or estimating equation, describes this relationship.

The scatter diagram gives an indication of the nature of the potential relationship between the variables.

Example 

A sample of 10 employees of the Universal Computer Corporation was examined to relate the employees' score on an aptitude test taken at the beginning of their employment and their monthly sales volume. The Universal Computer Corporation wishes to estimate the nature of the relationship between these two variables

Aptitude Test Score

Monthly Sales (Thousands of Rupees)

Aptitude Test Score

Monthly Sales (Thousands of Rupees)

X

Y

X

Y

50

30

70

60

50

35

70

45

60

40

80

55

60

50

80

50

70

55

90

65

To determine the nature of the relationship for example, we initially draw a graph to observe the data points.

Figure 1

2406_scatter diagram.png

On the vertical axis, we plot the dependent variable monthly sales. On the horizontal axis we plot the independent variable aptitude test score. This visual display is called a scatter diagram.

In the figure given above, we see that larger monthly sales are associated with larger test scores. If we wish, we can draw a straight line through the points plotted in the figure. This hypothetical line enables us to further describe the relationship. A line that slopes upward to the right indicates that a direct, or a positive relation is present between the two variables. In the figure given above we see that this upward-sloping line appears to approximate the relationship being studied.

The figures below show additional relations that may exist between two variables. In figure 2(a), the nature of the relationship is linear. In this case, the line slopes downward. Thus, smaller values of Y are associated with larger values of X. This relation is called an inverse (linear) relation.

Figure 2

705_scatter diagram1.png

 

Figure 2(b) represents a relationship that is not linear. The nature of the relationship is better represented by a curve than by a straight line - that is, it is a curvilinear relation. The relationship is inverse since smaller values of Y are associated with larger values of X.

Figure 2(c) is another curvilinear relation. In this case, however, larger values of Y are associated with larger values of X. Hence, the relation is direct and curvilinear.

In figure 2(d), there is no relation between X and Y. We can draw neither a straight line nor a curve that adequately describes the data. The two variables are not associated.


Related Discussions:- Scatter diagram - correlation analysis

Example of discrete random variable, Example of discrete random variable: ...

Example of discrete random variable: 1. What is a discrete random variable? Give three examples from the field of business. 2. Of 1000 items produced in a day at XYZ Manufa

Explain ridge regression, Using log(x1), log(x2) and log(x3) as the predict...

Using log(x1), log(x2) and log(x3) as the predictors, do pair wise scatterplots of all pairs of variables (including the response) and comment (use the pairs function). Do you thin

Effect in frequency domain, A) The three images shown below were blurred us...

A) The three images shown below were blurred using square masks of sizes n=23, 25, and 45, respectively. The vertical bars on the le_ lower part of (a) and (c) are blurred, but a c

Multi stage or cluster random sampling, Multi stage or Cluster Random sampl...

Multi stage or Cluster Random sampling  Under this method, the random selection is made of primary, intermediate and final units from a given population. The area of investigat

Draw a network diagram for this problem, The project of building a backyard...

The project of building a backyard swimming pool consists of eight major activities and has to be completed within 19 weeks. The activities and related data are given in the follow

Stratified sampling, Stratified Sampling Stratified Sampling is ...

Stratified Sampling Stratified Sampling is generally used when the population is heterogeneous. In this case, the population is first subdivided into several parts (or s

Systematic sampling, Systematic Sampling In Systematic Sampling ...

Systematic Sampling In Systematic Sampling each element has an equal chance of being selected, but each sample does not have the same chance of being selected. Here,

Regression and anova, The first step in this case is to ensure that you ar...

The first step in this case is to ensure that you are adequately clear on the General Linear Model and its relationship to both ANOVA and regression. The distinction is approxim

Statistics, Theories of Business forecasting

Theories of Business forecasting

Measurement error models., how can we use measurement error method with eig...

how can we use measurement error method with eight responses variables (we do not have explanatory variable in the data )?.the data analyse 521 leaves ..

Write Your Message!

Captcha
Free Assignment Quote

Assured A++ Grade

Get guaranteed satisfaction & time on delivery in every assignment order you paid with us! We ensure premium quality solution document along with free turntin report!

All rights reserved! Copyrights ©2019-2020 ExpertsMind IT Educational Pvt Ltd