AP Statistics · Topic 4.10 · Unit 4
AP Stats 4.10: Carrying Out a Two-Sample t-Test
By Jude Wallis · Published
Topic 4.10 carries out the two-sample t-test: t equals the difference in sample means minus 0, divided by the square root of s-1 squared over n-1 plus s-2 squared over n-2. Technology gives the p-value and degrees of freedom. Reject the null if the p-value is at most alpha, then conclude in context.
AP Statistics: Unit 4 (topics 4.10). Topic 4.10 (Carrying Out a Test for the Difference Between Two Population Means) in the Fall 2026 AP Statistics course.
What topic 4.10 covers
Topic 4.10 finishes the two-sample test you set up in topic 4.9. You calculate the test statistic and p-value, interpret the p-value, and justify a conclusion about the difference between two population means, . It closes Unit 4 and is a common setting for free-response question 3.
The logic is identical to the one-sample test in topic 4.5; only the standard error and degrees of freedom change to handle two independent samples.
The test statistic and p-value
Under the null hypothesis the two means are equal, so the hypothesized difference is 0. The two-sample t-statistic is
where the denominator is the standard error of the difference; when the null hypothesis is true, this statistic follows a t-distribution. As in topic 4.7, the degrees of freedom fall between the smaller of and and the value , and you read the exact df and p-value from technology. Match the tail direction to the alternative. The two-sample t-test calculator and the p-value calculator confirm the numbers.
The standard error uses each sample's own standard deviation and size, so the two groups are not pooled. Because the exact degrees of freedom come from a formula best handled by technology, report the calculator value; by hand, using the smaller of and is a valid conservative choice that yields a slightly larger p-value.
Interpreting the p-value and concluding
The p-value is the probability of a test statistic as extreme or more extreme than the one observed, in the direction of the alternative, assuming the null hypothesis is true, that is assuming the two population means are equal. See what does a p-value mean for the careful reading.
Make the formal decision by comparing to the significance level (alpha): if the p-value , reject ; if the p-value , fail to reject . Write the conclusion in context, in terms of the alternative, using non-definitive language, and refer to the parameters and both populations. You never prove that the means are equal.
The test and the two-sample interval from topic 4.7 agree: a two-sided test that rejects at level matches a interval that excludes 0, and a test that fails to reject matches an interval that contains 0.
Two-sample t-test for a difference of means
Students are randomly assigned to a new lesson method (group 1) or the standard method (group 2), 36 in each. Group 1 has , ; group 2 has , . Test against at .
Point estimate: points.
Standard error: .
Test statistic: .
Using the conservative (the smaller of and ), the one-sided p-value is about 0.011 from technology.
Compare to : , so reject .
Because the p-value of about 0.011 is at most 0.05, reject . There is convincing evidence that the new method produces a higher mean test score than the standard method.
Frequently asked questions
Why is the hypothesized difference 0 in the test statistic?
Because the null hypothesis says the two population means are equal, so their difference mu-1 minus mu-2 is 0. That is the value you subtract in the numerator. On the rare occasion a problem tests a specific nonzero difference, you would subtract that value instead, but the AP course uses 0.
Do I need equal sample sizes or equal variances?
No. The AP two-sample t-test does not assume equal population variances and does not require equal sample sizes; it uses each sample's own standard deviation and size in the standard error. This is the unpooled procedure that technology reports by default, so you do not pool the two standard deviations.