You are on page 1of 70

CHAPTER 3

DATA DESCRIPTION

WEBEX 2

We lead | Kami Memimpin

www.usm.my
CHAPTER 3: OVERVIEW

www.usm.my
We lead | Kami Memimpin
LEARNING
OUTCOMES

1. Summarize data using measures of central


tendency.
2. Describe data using measures of
variation.
3. Identify the position of a data value in a data set.
4. Use boxplots and five-number summaries to
discover various aspects of data.

www.usm.my
We lead | Kami Memimpin
Introduction
Traditional Statistics:
 Average
 Variation
 Position

www.usm.my
We lead | Kami Memimpin
 Mean
3–1  Median
Measures of  Mode

Central Tendency  Midrange


 Weighted Mean

www.usm.my
We lead | Kami Memimpin
Measures of Central Tendency
 A statistic is a characteristic or measure obtained by using the data
values from a sample.
 A parameter is a characteristic or measure obtained by using all the
data values for a specific population.

General Rounding Rule

• The basic rounding rule is that rounding should not be done


until the final answer is calculated.
• The use of parentheses on calculators or use of
spreadsheets help to avoid early rounding error.

www.usm.my
We lead | Kami Memimpin
What Do We Mean By Average?

 Mean

 Median

 Mode

 Midrange

 Weighted Mean

www.usm.my
We lead | Kami Memimpin
The Mean

www.usm.my
We lead | Kami Memimpin
Rounding Rule: The Mean
The mean should be rounded to one more decimal place
than occurs in the raw data.
The mean, in most cases, is not an actual data value.
www.usm.my
We lead | Kami Memimpin
The Mean for Grouped Data

The mean for grouped data is calculated


by multiplying the frequencies and
midpoints of the classes.

X  f X m

www.usm.my
We lead | Kami Memimpin
www.usm.my
We lead | Kami Memimpin
Below is a frequency distribution of miles run per week.
Find the mean.

www.usm.my
We lead | Kami Memimpin
www.usm.my
We lead | Kami Memimpin
The Median

www.usm.my
We lead | Kami Memimpin
www.usm.my
We lead | Kami Memimpin
www.usm.my
We lead | Kami Memimpin
The Mode

 It is sometimes said to be the most


typical case.
 There may be no mode, one mode
(unimodal), two modes (bimodal), or
many modes (multimodal).

www.usm.my
We lead | Kami Memimpin
www.usm.my
We lead | Kami Memimpin
www.usm.my
We lead | Kami Memimpin
www.usm.my
We lead | Kami Memimpin
www.usm.my
We lead | Kami Memimpin
The Midrange

www.usm.my
We lead | Kami Memimpin
www.usm.my
We lead | Kami Memimpin
The Weighted Mean

www.usm.my
We lead | Kami Memimpin
www.usm.my
We lead | Kami Memimpin
Properties of the Mean
 Uses all data values.
 Varies less than the median or mode.
 Used in computing other statistics, such as
the variance.
 Unique, usually not one of the data values.
 Cannot be used with open-ended classes.
 Affected by extremely high or low values,
called outliers.
www.usm.my
We lead | Kami Memimpin
Properties of the Median

 Gives the midpoint


 It is used when it is necessary to determine
whether the data values fall into the
distribution's upper or lower half.
 Can be used for an open-ended distribution.
 Affected less than the mean by extremely
high or extremely low values.

www.usm.my
We lead | Kami Memimpin
Properties of the Mode

 Used when the most typical


case is desired.
 Easiest average to compute.
 Can be used with nominal data.
 Not always unique or may not
exist.

www.usm.my
We lead | Kami Memimpin
Properties of the Midrange

 Easy to compute.
 Gives the midpoint.
 Affected by extremely high or
low values in a data set.

www.usm.my
We lead | Kami Memimpin
Types of Distributions

www.usm.my
We lead | Kami Memimpin
 Range
3–2  Variance

Measures of  Standard Deviation

Variation  Coefficient of Variation

 Chebyshev’s Theorem

 Empirical Rule (Normal)

www.usm.my
We lead | Kami Memimpin
How Can We Measure Variability?

 Range

 Variance

 Standard Deviation
 Coefficient of Variation
 Chebyshev’s Theorem
 Empirical Rule (Normal)
www.usm.my
We lead | Kami Memimpin
Range

www.usm.my
We lead | Kami Memimpin
www.usm.my
We lead | Kami Memimpin
Variance & Standard Deviation

www.usm.my
We lead | Kami Memimpin
Uses of Variance & Standard Deviation

www.usm.my
We lead | Kami Memimpin
Variance & Standard Deviation
(Population Theoretical Model)

www.usm.my
We lead | Kami Memimpin
www.usm.my
We lead | Kami Memimpin
Variance & Standard Deviation
(Sample Theoretical Model)

www.usm.my
We lead | Kami Memimpin
Variance & Standard Deviation
(Sample Computational Model)

www.usm.my
We lead | Kami Memimpin
Variance & Standard Deviation
(Sample Computational Model)

www.usm.my
We lead | Kami Memimpin
s2 1.28
s 1.13

www.usm.my
We lead | Kami Memimpin
Coefficient of Variation
A statistic that allows you to compare standard deviations
when the units are different is called the coefficient of
variation.

www.usm.my
We lead | Kami Memimpin
www.usm.my
We lead | Kami Memimpin
Range Rule of Thumb

The range can be used to approximate the standard


deviation. The approximation is called the range rule of
thumb.

www.usm.my
We lead | Kami Memimpin
www.usm.my
We lead | Kami Memimpin
Chebyshev’s Theorem

www.usm.my
We lead | Kami Memimpin
Chebyshev’s Theorem

www.usm.my
We lead | Kami Memimpin
www.usm.my
We lead | Kami Memimpin
www.usm.my
We lead | Kami Memimpin
Empirical Rule (Normal)

www.usm.my
We lead | Kami Memimpin
Empirical Rule (Normal)

www.usm.my
We lead | Kami Memimpin
 Z-score
3–3  Percentile
Measures of  Quartile
Position  Outlier

www.usm.my
We lead | Kami Memimpin
How Can We Measure the
Position of Data?

 Z-score

 Percentile

 Quartile

 Outlier

www.usm.my
We lead | Kami Memimpin
Z-score

www.usm.my
We lead | Kami Memimpin
She has a higher relative position in the Calculus class.

www.usm.my
We lead | Kami Memimpin
Percentiles

www.usm.my
We lead | Kami Memimpin
Percentiles

www.usm.my
We lead | Kami Memimpin
www.usm.my
We lead | Kami Memimpin
The value 5 corresponds to the 25th percentile.

www.usm.my
We lead | Kami Memimpin
Quartiles and Deciles

www.usm.my
We lead | Kami Memimpin
Quartiles and Deciles

www.usm.my
We lead | Kami Memimpin
Find Q1, Q2, and Q3 for the data set.
15, 13, 6, 5, 12, 50, 22, 18

www.usm.my
We lead | Kami Memimpin
Outliers

www.usm.my
We lead | Kami Memimpin
3–4
Exploratory
Data Analysis

www.usm.my
We lead | Kami Memimpin
The Five-Number Summary and Boxplots

 The Five-Number Summary is


composed of the following numbers:
Min, Q1, MD, Q3, Max
 The Five-Number Summary can be
graphically represented using a
Boxplot.

www.usm.my
We lead | Kami Memimpin
Procedure for Constructing Boxplots
1. Find the five-number summary.
2. Draw a horizontal axis with a scale that includes the
maximum and minimum data values.
3. Draw a box with vertical sides through Q1 and Q3, and
draw a vertical line through the median.
4. Draw a line from the minimum data value to the left side of
the box and a line from the maximum data value to the
right side of the box.

www.usm.my
We lead | Kami Memimpin
The number of meteorites found in 10 U.S. states is shown.
Construct a boxplot for the data.
89, 47, 164, 296, 30, 215, 138, 78, 48, 39

www.usm.my
We lead | Kami Memimpin
www.usm.my
We lead | Kami Memimpin
We lead | Kami Memimpin
End OF CHAPTER 3 www.usm.my

You might also like