Calculate cutoff values and analyzing histograms, Advanced Statistics

Assignment Help:

1. You are interested in investigating if being above or below the median income (medloinc) impacts ACT means (act94) for schools. Complete the necessary steps to examine univariate grouped data in order to respond to the questions below. Although deletions and/or transformations may be implied from your examination, all steps will examine original variables.

a. How many subjects have missing values for medlonic and act94?

b. Is there a severe split in frequencies between groups?

According to the descriptive analysis, no severe split is detected. This is also reflected in the skewness number which is lower than .5.

c. What are the cutoff values for outliers in each group?

d. Which outlying cases should be deleted for each group?

Average ACT score 1994 Stem-and-Leaf Plot for

medloinc= below the median for low inc % 1993

 Frequency Stem & Leaf

 7.00 14 . 1223789

 9.00 15 . 234478888

 5.00 16 . 12788

 4.00 17 . 1378

 2.00 18 . 09

 1.00 19 . 6

 3.00 20 . 069

 1.00 Extremes (>=22.5)

 Stem width: 1.0

 Each leaf: 1 case(s)

e. Analyzing histograms, normal Q-Q plots, and tests of normality, what is your conclusion regarding normality? If a transformation is necessary, which one would you use?

Tests of Normality

 

above or below median loinc

Kolmogorov-Smirnova

Shapiro-Wilk

 

Statistic

df

Sig.

Statistic

df

Sig.

average ACT score 1994

below the median for low inc % 1993

.162

32

.032

.903

32

.007

above the median for low inc % 1993

.166

32

.025

.921

32

.023

 

According to the information and the test of normality, it appears that this is a normal distribution.  Therefore, for the transformation, we would select 'Square Root."

 

 

f. Do the results from Levene's Test of Equal Variances indicate homogeneity of variance? Explain.

In running the test, there were no significant differences between the categories. Therefore; we can assume that this indicates homogeneity of variance.

2. Examination of the variable of scienc93 indicates a substantial to serve positively skewed distribution. Transform this variable using the most two appropriate methods. After examining the distribution for these transformed variables, which produced the best alteration?


Related Discussions:- Calculate cutoff values and analyzing histograms

Describe ignorability., Ignorability : The missing data mechanism is said t...

Ignorability : The missing data mechanism is said to be ignorable for likelihood inference if (1) the joint likelihood for the responses of the interest and missing data indicators

Baddeley''smetric, Baddeley'smetric : A manner of measuring the 'error' in ...

Baddeley'smetric : A manner of measuring the 'error' in the image processing technique or method. The metric is derived using the fundamental theory from the stochastic geometry an

What is the expectation of the number of tosses required, Question 1 A box...

Question 1 A box contains 20 fuses of which 5 are defective If 2 fuses are chosen together at random what is the probability that both the fuses are defective? Question 2 A c

Component bar chart, Component bar chart : A bar chart which shows the comp...

Component bar chart : A bar chart which shows the component parts of the aggregate represented by the whole length of the bar. The component parts are shown as the sectors of bar w

Define misspecification, Misspecification  is the term is applied to descri...

Misspecification  is the term is applied to describe the assumed statistical models which are incorrect for one of the several of reasons, for instance, using the wrong probability

Probabilistic matching, Probabilistic matching is a method developed to ma...

Probabilistic matching is a method developed to maximize the accuracy of the linkage decisions based on the level of agreement and disagreement among the identifiers on different

Attack rate, Attack rate : This term frequently used for the incidence of t...

Attack rate : This term frequently used for the incidence of the disease or condition in the particular group, or during a limited interval of time, or under the special circumstan

Dirichlet process, The distribution over distributions in the sense that ea...

The distribution over distributions in the sense that each draw from the process is itself the distribution. The name Dirichlet process or procedure is due to the fact that the ?ni

Categorizing continuous variables, Categorizing continuous variables : A pr...

Categorizing continuous variables : A practice which involves the conversion of the continuous variables into the series of the categories, which is common in the field of medical

Bayesian confidence interval, Bayesian confidence interval : An interval of...

Bayesian confidence interval : An interval of the posterior distribution which is so that the density of it at any point inside the interval is greater than that of the density at

Write Your Message!

Captcha
Free Assignment Quote

Assured A++ Grade

Get guaranteed satisfaction & time on delivery in every assignment order you paid with us! We ensure premium quality solution document along with free turntin report!

All rights reserved! Copyrights ©2019-2020 ExpertsMind IT Educational Pvt Ltd