← back to statistics

Sampling Distributions and the Central Limit Theorem

Statistics through experiments · 7 of 16 · CC BY-SA 4.0

Individual observations can be lopsided. Their averages have a surprising habit.

You have seen a sample mean move when you draw another sample. What if we kept every one of those estimates? They would form a distribution of their own.

The upper chart below describes individual values in a population. Start with one observation per sample and take 500 samples. With just one observation, each “average” is the observation itself. The two shapes should look similar, allowing for random variation.

Choose Keep these means. Then increase observations per sample to 30 and take 500 new samples. Predict which chart will change.

One sample becomes one mean

The population stays this shape
The population stays this shape. 101 values; exact bin counts are available below.02040Frequency050100Individual values

Dashed line: population mean = 20.3.

Read the chart as a table
The population stays this shape
Value rangeCount
0.0–2.540
2.5–5.08
5.0–7.55
7.5–10.04
10.0–12.53
12.5–15.03
15.0–17.52
17.5–20.02
20.0–22.52
22.5–25.02
25.0–27.52
27.5–30.02
30.0–32.51
32.5–35.01
35.0–37.52
37.5–40.01
40.0–42.51
42.5–45.01
45.0–47.52
47.5–50.01
50.0–52.51
52.5–55.01
55.0–57.51
57.5–60.01
60.0–62.50
62.5–65.01
65.0–67.51
67.5–70.01
70.0–72.51
72.5–75.01
75.0–77.50
77.5–80.01
80.0–82.51
82.5–85.01
85.0–87.50
87.5–90.01
90.0–92.51
92.5–95.00
95.0–97.51
97.5–100.01

Each range includes its lower end and excludes its upper end, except the last range, which includes both.

Draw independently with replacement: each observation uses the same population again. Average the values, and place that one mean in the lower chart.

0 sample means recorded.

Means of samples of 1
Means of samples of 1. 0 values; exact bin counts are available below.012Frequency050100Sample meansRun an experiment to add values.

Dashed line: population mean = 20.3.

Read the chart as a table
Means of samples of 1
Value rangeCount
0.0–1.30
1.3–2.50
2.5–3.80
3.8–5.00
5.0–6.30
6.3–7.50
7.5–8.80
8.8–10.00
10.0–11.30
11.3–12.50
12.5–13.80
13.8–15.00
15.0–16.30
16.3–17.50
17.5–18.80
18.8–20.00
20.0–21.30
21.3–22.50
22.5–23.80
23.8–25.00
25.0–26.30
26.3–27.50
27.5–28.80
28.8–30.00
30.0–31.30
31.3–32.50
32.5–33.80
33.8–35.00
35.0–36.30
36.3–37.50
37.5–38.80
38.8–40.00
40.0–41.30
41.3–42.50
42.5–43.80
43.8–45.00
45.0–46.30
46.3–47.50
47.5–48.80
48.8–50.00
50.0–51.30
51.3–52.50
52.5–53.80
53.8–55.00
55.0–56.30
56.3–57.50
57.5–58.80
58.8–60.00
60.0–61.30
61.3–62.50
62.5–63.80
63.8–65.00
65.0–66.30
66.3–67.50
67.5–68.80
68.8–70.00
70.0–71.30
71.3–72.50
72.5–73.80
73.8–75.00
75.0–76.30
76.3–77.50
77.5–78.80
78.8–80.00
80.0–81.30
81.3–82.50
82.5–83.80
83.8–85.00
85.0–86.30
86.3–87.50
87.5–88.80
88.8–90.00
90.0–91.30
91.3–92.50
92.5–93.80
93.8–95.00
95.0–96.30
96.3–97.50
97.5–98.80
98.8–100.00

Each range includes its lower end and excludes its upper end, except the last range, which includes both.

All horizontal axes stay at 0–100. Each vertical axis counts its own observations. At most the latest 10,000 means are shown. Changing sample size clears the current means; changing population clears both runs.

Put a number on the narrowing

The population’s standard deviation is 27.17. For samples of 1, the standard error of the mean is 27.17 ÷ √1 = 27.17. The population’s spread stays the same; the estimates become steadier.

One dot stands for a whole sample

The lower chart is an observed picture of a sampling distribution: the distribution of a statistic across repeated samples drawn by the same rule. Each entry is a sample mean, not an individual measurement.

Try the other population shapes. An even spread, a long tail, and two separated peaks look quite different. Yet averages of many independent observations tend toward a common, rounded shape.

Increasing the number of samples makes that picture clearer. Increasing the number of observations per sample changes the distribution of the mean itself. These are the two different knobs you met with coins.

A moment to think

You keep 30 observations per sample and collect 500 more sample means. What changes?

A name for steadier estimates

The standard deviation of a sampling distribution is called its standard error. It describes how much an estimate varies from sample to sample. The population’s standard deviation describes how much individuals vary.

For the independent draws here, multiplying sample size by four halves the standard error of the mean. Taking four times as many samples of the old size would not do that.

You have met a theorem

The Central Limit Theorem

For independent observations drawn from the same distribution with finite, positive variance, the distribution of their standardized mean approaches a standard normal distribution as sample size grows.

In ordinary units: for sufficiently large samples, the means are approximately bell-shaped around the population mean, with standard error equal to the population standard deviation divided by the square root of sample size.

SE(x̄) = σ / √n

σ, “sigma,” is the population standard deviation. n is observations per sample. SE measures the spread of sample means. This standard-error relationship is exact under these assumptions; the bell shape is an approximation.

What does “standardized” mean?

Subtract the population mean μ, then divide by the standard error: (x̄ − μ) / (σ / √n). This puts the estimate’s distance from its target in units of sampling uncertainty. That distribution approaches a normal distribution centered at zero with standard deviation one.

The population does not have to become normal. There is no universal “30 is enough” rule: strong skew or rare extremes can demand much larger samples. Dependence or infinite variance can break this version of the theorem.

The experiments illustrate the theorem; they do not prove it. See Penn State’s Central Limit Theorem lesson for a mathematical treatment.

Return to the experiment ↑

A moment to think

Delivery times have a long right tail. You average 100 independent deliveries from the same process. What does the CLT suggest?

Words you met

TermMeaning
Sampling variabilityThe change in an estimate across different samples.
Sampling distributionThe distribution of a statistic over repeated samples of the same size and design.
Standard errorThe standard deviation of a sampling distribution.
Central Limit TheoremUnder its conditions, standardized sample means approach a standard normal distribution.
Neighbors

Written by June Kim. The conversational approach was inspired by Danielle Navarro’s Learning Statistics with R (CC BY-SA 4.0). The prose, experiments, and questions here were created for this book. This chapter’s text and illustrations are also shared under CC BY-SA 4.0.