AP Statistics · Unit 4 · 10-20% of the exam · ~18 class periods
Inference for Quantitative Data: Means
AP Statistics Unit 4 covers inference for means: sampling distributions of sample means, the t-distribution, and one-sample, matched-pairs, and two-sample t confidence intervals and tests. It is 10 to 20 percent of the multiple-choice section and about 18 class periods.
AP Statistics: Unit 4 (topics 4.1 Sampling Distributions for Sample Means, 4.2 Constructing a Confidence Interval for a Population Mean or Population Mean Difference, 4.3 Justifying a Claim Based on a Confidence Interval for a Population Mean or Population Mean Difference, 4.4 Setting Up a Test for a Population Mean or Population Mean Difference, 4.5 Carrying Out a Test for a Population Mean or Population Mean Difference, 4.6 Sampling Distributions for the Difference Between Two Sample Means, 4.7 Constructing a Confidence Interval for the Difference Between Two Population Means, 4.8 Justifying a Claim Based on a Confidence Interval for the Difference Between Two Population Means, 4.9 Setting Up a Test for the Difference Between Two Population Means, 4.10 Carrying Out a Test for the Difference Between Two Population Means). Unit 4 (Inference for Quantitative Data: Means) is 10 to 20 percent of the multiple-choice section and about 18 class periods in the Fall 2026 redesigned course, covering one-sample, matched-pairs, and two-sample t procedures.
What Unit 4 covers
Unit 4 is where you do inference for quantitative data, meaning numerical measurements whose center is a mean. It builds directly on the ideas you used for proportions in Unit 3, but the parameter is now a population mean, written (the Greek letter mu), instead of a proportion.
The College Board weights Unit 4 at 10 to 20 percent of the multiple-choice section and allots about 18 class periods. There are 10 topics, split into two halves: one-sample and matched-pairs procedures (topics 4.1 to 4.5), then two-sample procedures (topics 4.6 to 4.10).
The big change from Unit 3 is that you almost never know the population standard deviation (sigma) for a quantitative variable. You estimate it with the sample standard deviation , and that forces you to use the t-distribution instead of the normal z-distribution. See t-test vs z-test for the full contrast.
The 10 topics in Unit 4
- 4.1 Sampling distributions for sample means. Find the mean and standard deviation of the sampling distribution of the sample mean (x-bar), and check when it is approximately normal.
- 4.2 Constructing a confidence interval for a population mean or mean difference. Describe t-distributions and build a one-sample t-interval, including for matched pairs.
- 4.3 Justifying a claim based on that interval. Interpret the interval and the confidence level, and relate sample size, width, and margin of error.
- 4.4 Setting up a test for a population mean or mean difference. Choose the one-sample t-test, state hypotheses about or the population mean difference , and check conditions.
- 4.5 Carrying out that test. Compute the t-statistic and p-value, then decide by comparing the p-value to the significance level (alpha).
- 4.6 Sampling distributions for the difference between two sample means. Find the mean and standard deviation of .
- 4.7 Constructing a confidence interval for the difference between two population means. Build a two-sample t-interval.
- 4.8 Justifying a claim based on that interval. Interpret it, and use whether the interval contains 0.
- 4.9 Setting up a test for the difference between two population means. Choose the two-sample t-test and state the hypotheses.
- 4.10 Carrying out that test. Compute the two-sample t-statistic and p-value and conclude in context.
Sampling distributions for sample means
For a population with mean and standard deviation , when the sampled values are independent, the sampling distribution of the sample mean has mean and standard deviation , where is the sample size.
That standard deviation shrinks as grows, which is why larger samples give more precise estimates. The central limit theorem tells you the shape. If the population is normal, is normal for any . If the population is not normal, is approximately normal once , and a strongly skewed population may need a sample much larger than 30.
You can watch this happen in the sampling distribution interactive.
Why means use the t-distribution
Because you rarely know , you replace it with the sample standard deviation . The quantity is the standard error of the mean, an estimate of the true standard deviation . Those two ideas trip up many students, so read standard error vs standard deviation.
Using adds extra uncertainty, so the standardized statistic follows a t-distribution rather than the standard normal. t-distributions are symmetric and bell-shaped with heavier tails than the normal curve. Each one is set by its degrees of freedom (df); for one sample, . As df grows, the t-distribution approaches the standard normal, and you look up critical values on the t-table.
One-sample and matched-pairs t procedures
A one-sample t-interval estimates as , where is the critical value for the central C percent of a t-distribution with degrees of freedom. The one-sample t-test uses the statistic
where is the value of the mean claimed by the null hypothesis.
Matched pairs, for example a pre-test and post-test on the same students, are handled by taking the difference within each pair to get one sample of differences. You then run the same one-sample procedures on those differences, estimating the population mean difference . The one-sample t-test calculator runs both the interval and the test.
Two-sample t procedures
When you compare two independent groups, the parameter is the difference between population means . The sampling distribution of has standard deviation , and its standard error replaces each with the matching sample standard deviation .
The two-sample t-interval is , and the two-sample t-test uses
The degrees of freedom fall between and the smaller of and , and on the exam you find them with technology. If a two-sample interval contains 0, you do not have convincing evidence of a difference. The two-sample t-test calculator handles the arithmetic.
Conditions to check before any t procedure
Every t procedure in Unit 4 asks you to verify conditions before you trust the result, and you should state them explicitly on the free-response section.
- Randomization: the data come from a random sample or a randomized experiment, using two independent random samples for two-sample procedures.
- 10 percent: when sampling without replacement, each sample is at most 10 percent of its population, written , where is the population size.
- Sample data (normality): the population is stated to be approximately normal, or . If , the sample should be free of strong skewness and outliers. For matched pairs, apply this to the differences.
State your hypotheses in terms of population parameters, not sample statistics, and write the conclusion in context by comparing the p-value to .
How Unit 4 fits the exam
Unit 4 mirrors Unit 3 step for step, so the College Board expects you to keep the procedures straight and pick the right one from the wording of a problem. A question asking whether data give convincing evidence of something is asking for a significance test, not a description.
Free-response question 3 is always a full inference problem, a hypothesis test or a confidence interval, and Unit 4 procedures are fair game there. A graphing calculator with statistical capabilities is expected, and a formula sheet and tables are provided. For the wider picture see the AP Statistics hub and the official AP Statistics course page.
95% confidence interval for a mean (one sample)
A random sample of 6 phones of one model runs for 8, 10, 12, 14, 9, and 13 hours on a full charge. Assume battery life is approximately normal. Construct a 95% confidence interval for the mean battery life .
Find the sample mean: hours.
Find the sample standard deviation. The deviations from 11 are -3, -1, 1, 3, -2, 2; their squares are 9, 1, 1, 9, 4, 4, which sum to 28. So hours.
Find the standard error: hours.
Degrees of freedom are . For 95% confidence, the critical value is .
Margin of error: hours.
Confidence interval: .
You are 95% confident that the mean battery life for this phone model is between about 8.52 and 13.48 hours. Because , this relies on the stated assumption that battery life is approximately normal.
One-sample t-test for a population mean
A machine should fill bottles to a mean of 500 mL. A random sample of 5 bottles contains 498, 502, 495, 501, and 499 mL. Assume the fills are approximately normal. Test at whether the mean fill differs from 500 mL.
State hypotheses: and , where is the mean fill volume in mL. The null value is .
Sample mean: mL.
Sample standard deviation. The deviations from 499 are -1, 3, -4, 2, 0; their squares are 1, 9, 16, 4, 0, summing to 30. So mL.
Standard error: mL.
Test statistic: , with .
Two-sided p-value from technology: .
Because , you fail to reject . There is not convincing evidence that the machine's mean fill volume differs from 500 mL.
Frequently asked questions
When do you use a t-test instead of a z-test for means?
Use a t procedure whenever the population standard deviation is unknown and you estimate it with the sample standard deviation , which is almost always the case for quantitative data. A z procedure for a mean would require you to already know . See t-test vs z-test for the full comparison.
What are the degrees of freedom for a t procedure?
For a one-sample or matched-pairs t procedure, the degrees of freedom are . For a two-sample t procedure, the degrees of freedom fall between and the smaller of and , and the exam expects you to find them with technology.
How do matched pairs differ from two independent samples?
Matched pairs come from linked measurements, such as a pre-test and post-test on the same students, so you take the difference within each pair and analyze that one sample of differences with a one-sample t procedure. Two independent samples come from separate groups, so you use a two-sample t procedure on the difference .
Every topic in Unit 4
- 4.1Sampling Distributions for Sample Means
- 4.2Constructing a Confidence Interval for a Population Mean or Population Mean Difference
- 4.3Justifying a Claim Based on a Confidence Interval for a Population Mean or Population Mean Difference
- 4.4Setting Up a Test for a Population Mean or Population Mean Difference
- 4.5Carrying Out a Test for a Population Mean or Population Mean Difference
- 4.6Sampling Distributions for the Difference Between Two Sample Means
- 4.7Constructing a Confidence Interval for the Difference Between Two Population Means
- 4.8Justifying a Claim Based on a Confidence Interval for the Difference Between Two Population Means
- 4.9Setting Up a Test for the Difference Between Two Population Means
- 4.10Carrying Out a Test for the Difference Between Two Population Means