You are on page 1of 14

Statistics for Business and Economics (13e)

BUSINESS STATISTICS

(MBA 803)

BY

MOHAMMED NUHU PhD

© 2017 Cengage Learning. May not be scanned, copied or duplicated, or posted to a publicly accessible website, in whole or in part, except for use as permitted in a license distributed with a certain product or service or 1
otherwise on a password-protected website or school-approved learning management system for classroom use.
Statistics for Business and Economics (13e)

Sampling

In statistics, quality assurance, and survey methodology, sampling is the


selection of a subset (a statistical sample) of individuals from within a statistical
population to estimate characteristics of the whole population.

Statisticians attempt for the samples to represent the population in question.

Two advantages of sampling are lower cost and faster data collection than
measuring the entire population.

© 2017 Cengage Learning. May not be scanned, copied or duplicated, or posted to a publicly accessible website, in whole or in part, except for use as permitted in a license distributed with a certain product or service or 2
otherwise on a password-protected website or school-approved learning management system for classroom use.
Statistics for Business and Economics (13e)

Types of Sampling

In applications:

Probability Sampling: Simple Random Sampling, Stratified Random Sampling, Multi-


Stage Sampling etc.
What is each and how is it done?
How do we decide which to use?
How do we analyze the results differently depending on the type of sampling?
Non-probability Sampling: voluntary response sampling, Judgement sampling,
Convenience sampling etc.
Why don't we use non-probability sampling schemes? Two reasons:
We can't use the mathematics of probability to analyze the results.
In general, we can't count on a non-probability sampling scheme to produce
representative samples.
© 2017 Cengage Learning. May not be scanned, copied or duplicated, or posted to a publicly accessible website, in whole or in part, except for use as permitted in a license distributed with a certain product or service or 3
otherwise on a password-protected website or school-approved learning management system for classroom use.
Statistics for Business and Economics (13e)

A probability sample is a sample in which every unit in the


population has a chance (greater than zero) of being selected
in the sample, and this probability can be accurately
determined.

The combination of these traits makes it possible to produce


unbiased estimates of population totals, by weighting sampled
units according to their probability of selection.

© 2017 Cengage Learning. May not be scanned, copied or duplicated, or posted to a publicly accessible website, in whole or in part, except for use as permitted in a license distributed with a certain product or service or 4
otherwise on a password-protected website or school-approved learning management system for classroom use.
Statistics for Business and Economics (13e)

Probability sampling includes: Simple Random


Sampling, Systematic Sampling, Stratified Sampling,
Probability Proportional to Size Sampling,
and Cluster or Multistage Sampling.

These various ways of probability sampling have two things in


common:
•Every element has a known nonzero probability of being
sampled and
•involves random selection at some point.

© 2017 Cengage Learning. May not be scanned, copied or duplicated, or posted to a publicly accessible website, in whole or in part, except for use as permitted in a license distributed with a certain product or service or 5
otherwise on a password-protected website or school-approved learning management system for classroom use.
Statistics for Business and Economics (13e)

Simple Random Sampling: A simple random sample (SRS) of size n is produced by a scheme which
ensures that each subgroup of the population of size n has an equal probability of being chosen as
the sample.

Stratified Random Sampling: Divide the population into "strata". There can be any number of these.
Then choose a simple random sample from each stratum. Combine those into the overall sample.
That is a stratified random sample. (Example: Church A has 600 women and 400 women as
members. One way to get a stratified random sample of size 30 is to take a SRS of 18 women from
the 600 women and another SRS of 12 men from the 400 men.).

Multi-Stage Sampling: Sometimes the population is too large and scattered for it to be practical to
make a list of the entire population from which to draw a SRS. For instance, when the a polling
organization samples US voters, they do not do a SRS. Since voter lists are compiled by counties,
they might first do a sample of the counties and then sample within the selected counties. This
illustrates two stages. In some instances, they might use even more stages. At each stage, they might
do a stratified random sample on sex, race, income level, or any other useful variable on which they
could get information before sampling.

© 2017 Cengage Learning. May not be scanned, copied or duplicated, or posted to a publicly accessible website, in whole or in part, except for use as permitted in a license distributed with a certain product or service or 6
otherwise on a password-protected website or school-approved learning management system for classroom use.
Statistics for Business and Economics (13e)

How does one decide which type of sampling to use?


The formulas in almost all statistics books assume simple random sampling. Unless you are willing to learn the
more complex techniques to analyze the data after it is collected, it is appropriate to use simple random sampling.
To learn the appropriate formulas for the more complex sampling schemes, look for a book or course on sampling.
Stratified random sampling gives more precise information than simple random sampling for a given sample size.
So, if information on all members of the population is available that divides them into strata that seem relevant,
stratified sampling will usually be used.
If the population is large and enough resources are available, usually one will use multi-stage sampling. In such
situations, usually stratified sampling will be done at some stages.

How do we analyze the results differently depending on the different type of sampling?
The main difference is in the computation of the estimates of the variance (or standard deviation). An excellent
book for self-study is A Sampler on Sampling, by Williams, Wiley. In this, you see a rather small population and
then a complete derivation and description of the sampling distribution of the sample mean for a particular small
sample size. I believe that is accessible for any student who has had an upper-division mathematical statistics
course and for some strong students who have had a freshman introductory statistics course. A very simple
statement of the conclusion is that the variance of the estimator is smaller if it came from a stratified random
sample than from simple random sample of the same size. Since small variance means more precise information
from the sample, we see that this is consistent with stratified random sampling giving better estimators for a given
sample size.
© 2017 Cengage Learning. May not be scanned, copied or duplicated, or posted to a publicly accessible website, in whole or in part, except for use as permitted in a license distributed with a certain product or service or 7
otherwise on a password-protected website or school-approved learning management system for classroom use.
Statistics for Business and Economics (13e)

Non-probability sampling is any sampling method where some elements of the population
have no chance of selection (these are sometimes referred to as 'out of
coverage'/'undercovered'), or where the probability of selection can't be accurately
determined.

It involves the selection of elements based on assumptions regarding the population of


interest, which forms the criteria for selection.

Hence, because the selection of elements is nonrandom, nonprobability sampling does not
allow the estimation of sampling errors.

These conditions give rise to exclusion bias, placing limits on how much information a sample
can provide about the population.

Information about the relationship between sample and population is limited, making it difficult
to extrapolate from the sample to the population.

© 2017 Cengage Learning. May not be scanned, copied or duplicated, or posted to a publicly accessible website, in whole or in part, except for use as permitted in a license distributed with a certain product or service or 8
otherwise on a password-protected website or school-approved learning management system for classroom use.
Statistics for Business and Economics (13e)

Non-probability sampling methods include convenience


sampling, quota sampling and purposive sampling.

In addition, non-response effects may turn any probability design


into a non-probability design if the characteristics of non-
response are not well understood, since non-response
effectively modifies each element's probability of being sampled.

© 2017 Cengage Learning. May not be scanned, copied or duplicated, or posted to a publicly accessible website, in whole or in part, except for use as permitted in a license distributed with a certain product or service or 9
otherwise on a password-protected website or school-approved learning management system for classroom use.
Statistics for Business and Economics (13e)

Non-probability sampling schemes


These include voluntary response sampling, judgement sampling, convenience sampling, and maybe
others.

In the early part of the 20th century, many important samples were done that weren't based on probability
sampling schemes. They led to some memorable mistakes. Look in an introductory statistics text at the
discussion of sampling for some interesting examples. The introductory statistics books I usually teach from
are Basic Practice of Statistics by David Moore, Freeman, and Introduction to the Practice of Statistics by
Moore and McCabe, also from Freeman. A particularly good book for a discussion of the problems of non-
probability sampling is Statistics by Freedman, Pisani, and Purves. The detail is fascinating. Or, ask a
statistics teacher to lunch and have them tell you the stories they tell in class. Most of us like to talk about
these! Someday when I have time, maybe I'll write some of them here.

Mathematically, the important thing to recognize is that the discipline of statistics is based on the
mathematics of probability. That's about random variables. All of our formulas in statistics are based on
probabilities in sampling distributions of estimators. To create a sampling distribution of an estimator for a
sample size of 30, we must be able to consider all possible samples of size 30 and base our analysis on how
likely each individual result is.

© 2017 Cengage Learning. May not be scanned, copied or duplicated, or posted to a publicly accessible website, in whole or in part, except for use as permitted in a license distributed with a certain product or service or 10
otherwise on a password-protected website or school-approved learning management system for classroom use.
Statistics for Business and Economics (13e)

Voluntary Sampling
The voluntary sampling method is a type of non-probability sampling. Volunteers choose to
complete a survey.

Volunteers may be invited through advertisements in social media. The target population for
advertisements can be selected by characteristics like location, age, sex, income, occupation,
education or interests using tools provided by the social medium. The advertisement may include a
message about the research and link to a survey. After following the link and completing the survey
the volunteer submits the data to be included in the sample population. This method can reach a
global population but is limited by the campaign budget. Volunteers outside the invited population
may also be included in the sample.

It is difficult to make generalizations from this sample because it may not represent the total
population. Often, volunteers have a strong interest in the main topic of the survey

© 2017 Cengage Learning. May not be scanned, copied or duplicated, or posted to a publicly accessible website, in whole or in part, except for use as permitted in a license distributed with a certain product or service or 11
otherwise on a password-protected website or school-approved learning management system for classroom use.
Statistics for Business and Economics (13e)

In quota sampling, the population is first segmented into mutually exclusive sub-
groups, just as in stratified sampling. Then judgement is used to select the subjects
or units from each segment based on a specified proportion. For example, an
interviewer may be told to sample 200 females and 300 males between the age of
45 and 60.

It is this second step which makes the technique one of non-probability sampling. In
quota sampling the selection of the sample is non-random. For example,
interviewers might be tempted to interview those who look most helpful. The
problem is that these samples may be biased because not everyone gets a chance
of selection. This random element is its greatest weakness and quota versus
probability has been a matter of controversy for several years.

© 2017 Cengage Learning. May not be scanned, copied or duplicated, or posted to a publicly accessible website, in whole or in part, except for use as permitted in a license distributed with a certain product or service or 12
otherwise on a password-protected website or school-approved learning management system for classroom use.
Statistics for Business and Economics (13e)

Sampling methods

Within any of the types of frames identified above, a variety of sampling


methods can be employed, individually or in combination.

Factors commonly influencing the choice between these designs include:


Nature and quality of the frame
Availability of auxiliary information about units on the frame
Accuracy requirements, and the need to measure accuracy
Whether detailed analysis of the sample is expected
Cost/operational concerns

© 2017 Cengage Learning. May not be scanned, copied or duplicated, or posted to a publicly accessible website, in whole or in part, except for use as permitted in a license distributed with a certain product or service or 13
otherwise on a password-protected website or school-approved learning management system for classroom use.
Statistics for Business and Economics (13e)

© 2017 Cengage Learning. May not be scanned, copied or duplicated, or posted to a publicly accessible website, in whole or in part, except for use as permitted in a license distributed with a certain product or service or 14
otherwise on a password-protected website or school-approved learning management system for classroom use.

You might also like