Skip to content

Recent Posts

  • Sampling Distribution Simulator
  • Glossary of Common Academic Acronyms and Jargon (Canada & U.S.)
  • 16 Tips for Succeeding in Graduate School
  • Taking Feedback from your Supervisor
  • How to Pass your Comprehensive Exams

Most Used Categories

  • Publishing (4)
  • Writing (9)
  • Tools (3)
  • Methods (25)
    • Design (3)
    • Methodology (4)
    • Quantitative Data Analysis (18)
      • STATA (1)
      • SPSS (10)
  • Graduate School (3)
Skip to content

Practical Advice for Research, Writing, and Academia’s Hidden Curriculum

Subscribe
  • Graduate School
  • Writing
  • Methods
    • Methodology
    • Design
    • Quantitative Data Analysis
      • Statistics
      • SPSS
      • STATA
  • Tools
  • Publishing
  • Podcast
  • Home
  • Tools
  • Sampling Distribution Simulator

Sampling Distribution Simulator

Phyllis L. F. RippeySeptember 5, 2026September 5, 2026

Table of Contents

Toggle
  • What is a sample?
  • What is a sampling distribution?
  • Sampling Distribution Simulator
  • What’s the Central Limit Theorem?
  • What do we use this for?

What is a sample?

The goal of inferential statistics is to make inferences (or estimates) about what’s happening in a population. We can’t usually ask everyone in a population to answer our survey, so we use a random sample from that population.

A random sample is always going to look a little bit different from the population from which it was drawn. In other words, there’s going to be a bit of sampling error between the average we might calculate for a population and the average we calculate for a sample. Based on the principles of a sampling distribution, we can estimate just how much error there might be in our estimates.

What is a sampling distribution?

A sampling distribution shows what would happen if we repeatedly drew random samples of the same size from the same population and calculated the same statistic each time. We don’t actually ever do this in real life, but knowing what would happen helps us make estimates about the one sample we do select.

Sampling distributions help us understand how much a sample statistic can vary simply because of random sampling. Some samples will produce means below the true population mean (μ) and some above it.

If we draw more and more random samples, we can see the theoretical sampling distribution more clearly. The theoretical mean of the sampling distribution of sample means (μxˉ\mu_{\bar{x}}) is equal to the population mean (μ\mu). In our simulation, the mean of the sample means will get closer to μ\mu as we add more samples.

Sampling Distribution Simulator

Below is a sampling distribution simulator that lets you see these ideas in action. The simulator begins with a population of 100 people who differ by their age. Because we can see the entire population, we know the true average age of everyone in the population (μ).

Choose a sample size (n) and draw a random sample from the population. The simulator will calculate the average age of the people in your sample (x̄). Notice which people were randomly selected and how the sample mean compares with the population mean.

Next, add your sample mean (x̄) to the sampling distribution. The x̄ that appears on the graph represents the mean from that one random sample. Draw another sample and add its mean. As you repeat this process, you will see a sampling distribution of means begin to take shape. You can then use the +1, +10, +50, and +100 buttons to speed things up.

Try changing the sample size (n) and starting again. Compare the sampling distribution for a very small sample (n = 3) with those for larger samples (n = 10, 30, or 50). What happens to the shape and spread of the sampling distribution? How closely do the sample means cluster around the population mean (μ)?

Age Sampling Distribution Simulator
Age Sampling Distribution Simulator
Population → random sample → sample mean → sampling distribution of means.
Population (N = 100)
Each person is one case. The number above each icon is that person’s age.
Population mean (μ) = 44.7 years
12
14
20
61
55
30
76
5
99
69
60
39
82
41
40
43
46
25
50
72
34
75
35
19
49
74
91
56
3
15
86
24
6
88
58
36
4
51
1
6
95
77
97
42
38
37
10
28
92
21
94
68
67
79
63
59
2
17
90
7
10
9
65
16
48
22
27
0
12
84
4
96
18
33
7
47
98
57
44
26
32
29
45
71
62
23
8
2
54
11
64
70
13
31
53
73
80
66
52
78
Random sample
Exactly 30 people selected from the population of 100 above (n = 30).
Sample mean (x̄) = 49.0 years
12
61
76
69
82
43
50
75
49
56
86
88
4
6
97
37
92
68
63
17
10
16
27
84
18
47
44
29
62
2
Sampling Distribution of x̄
Each x̄ = one sample mean.
Means added: 0
Population mean μ = 44.7 Sample mean (x̄): age in years Frequency
Sample Size (n):
Draw a sample, add its x̄ to the sampling distribution, then keep adding random sample means to watch the distribution take shape. Change n and compare how the spread changes.

What’s the Central Limit Theorem?

The Central Limit Theorem tells us what happens to the sampling distribution of the mean as we increase the size of each random sample (n). As sample size increases, the sampling distribution will:

  • become increasingly close to the shape of a normal distribution, even when the original population is not normally distributed;
  • become less spread out and have a smaller standard deviation (the standard deviation of a sampling distribution is called the standard error or σx̄); and
  • remain centred around the population mean (μ).

This means that larger samples tend to produce more precise estimates of population parameters—that is, estimates with less sampling error.

(Note: σ is the Greek letter sigma and μ is the Greek letter mu. We conventionally use Greek letters to represent population parameters—the values we are usually trying to estimate but cannot observe directly. We use Roman letters for the equivalent sample statistics: s for the sample standard deviation and x̄ (“x-bar”) for the sample mean.)

What do we use this for?

This is all mostly theoretical at this point, but we will use these principles to estimate how much sampling error is associated with our statistics when we calculate things like confidence intervals, t-tests, and the statistical significance of regression coefficients.


See also: Samples vs. Populations

Share this:

  • Email a link to a friend (Opens in new window) Email
  • Share on LinkedIn (Opens in new window) LinkedIn
  • Share on Threads (Opens in new window) Threads
  • Share on Facebook (Opens in new window) Facebook
  • Share on Pinterest (Opens in new window) Pinterest
  • Share on Bluesky (Opens in new window) Bluesky
  • Share on WhatsApp (Opens in new window) WhatsApp
  • Share on Reddit (Opens in new window) Reddit

Like this:

Like Loading…
central limit theorem, sampling, central tendency, means, inferential statistics, Social Statistics, Statistics

Post navigation

Previous: Glossary of Common Academic Acronyms and Jargon (Canada & U.S.)
Next: 16 Tips for Succeeding in Graduate School

Related Posts

Glossary of Common Academic Acronyms and Jargon (Canada & U.S.)

October 3, 2025July 24, 2026 Phyllis L. F. Rippey

Samples vs. Populations

January 23, 2020July 24, 2026 Phyllis L. F. Rippey

Levels of Measurement

October 1, 2019July 24, 2026 Phyllis L. F. Rippey

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Copyright All Rights Reserved | Theme: BlockWP by Candid Themes.
%d
    We use cookies to ensure that we give you the best experience on our website. If you continue to use this site we will assume that you are happy with it.