Checking conditions for inference practice
By Jude Wallis · Updated
These 8 problems make you verify inference conditions with numbers on the page: Random, 10%, and Large Counts for proportions, the sample data condition for means, and expected counts for chi-square. Four of them end in 'not justified', which is the correct answer when a condition fails.
AP Statistics: Unit 3 (topics 3.3 Constructing a Confidence Interval for a Population Proportion, 3.5 Setting Up a Test for a Population Proportion, 3.12 Setting Up a Test for the Difference Between Two Population Proportions, 3.14 Setting Up a Chi-Square Test for Homogeneity or Independence, 4.2 Constructing a Confidence Interval for a Population Mean or Population Mean Difference, 4.4 Setting Up a Test for a Population Mean or Population Mean Difference). These problems cover the condition-verification objectives spread across Units 3 and 4 of the Fall 2026 AP Statistics course, where every procedure is justified by checking randomization, the 10% condition, and either a normality condition on counts, a sample data condition on shape, or an expected counts condition.
What these problems build
These 8 problems train the step most students rush: showing, with numbers, that an inference procedure is allowed before running it. Naming a condition earns nothing. You have to write the count, the product, or the plot description that makes it true, and you have to be willing to write 'not justified' when it is not.
Three ideas run through the set. A test checks its counts at the hypothesized value (read 'p-naught', the number stated in the null hypothesis), while a confidence interval has no such value and must check them at the observed (read 'p-hat', the sample proportion). For a mean, settles the shape question on its own, and below 30 the description of the sample plot decides it. For a (chi-square) test the rule is about expected counts, not the counts you observed.
Work each problem before opening the steps, and write every condition as a sentence a reader could check. The conditions for inference checklist collects the rules in one place, why 10 successes and 10 failures explains where the proportion rule comes from, and t-tests with a small sample size covers the case. The course wording is on topic 3.5 and topic 4.2, and more sets are on the practice page.
Problem 1
A regional transit agency states that 35% of its riders use a monthly pass. The agency serves about 24,000 riders. A researcher takes a random sample of 240 riders and finds 96 with a monthly pass. Check every condition for a one-sample z-test of the null hypothesis against the alternative hypothesis , showing the numbers.
Show the worked solution
Random. The 240 riders are stated to be a random sample of the agency's riders, so the randomization condition is met. Say what makes it random rather than writing the word 'random' by itself.
10%. Sampling is without replacement, so the sample size must be at most 10% of the population size : . Here , and , so the condition is met.
Large Counts. For a test, the counts are checked at the hypothesized value , not at the sample proportion. and .
Confirm both counts clear 10. and , so the condition is met. As an arithmetic check, , which is .
Note what plays no role here. The sample proportion is the statistic you will standardize next, but it does not enter the Large Counts check for a test.
Conclude. All three conditions are met, so a one-sample z-test for a population proportion is justified for these data.
Random: stated random sample of riders. 10%: . Large Counts: and , both at least 10. All three met, so the one-sample z-test is justified.
Problem 2
A seed library has about 900 members. Staff take a random sample of 60 members and find that 8 of them returned seeds at the end of the season. (a) Are the conditions met for a one-sample z-test of against ? (b) Are the conditions met for a one-sample z-interval for ? (c) Explain why the two answers differ.
Show the worked solution
(a) Random and 10%. The 60 members are a random sample, and with , so both conditions are met.
(a) Large Counts at . A test has a hypothesized value, so use : and . Both are at least 10, so the condition is met and the test is justified.
(b) Random and 10% again. These do not depend on the procedure, so they are still met.
(b) Normality at . An interval has no hypothesized value, so the condition uses the observed counts: successes and failures .
(b) Compare to 10. The 52 failures are fine, but the 8 successes fall below 10, so the condition fails and a one-sample z-interval is not justified for these data.
(c) Explain the gap. A test asks whether this sample would be surprising if were , so its sampling distribution is centered at and the counts are checked there. An interval has no such anchor and must use , and at that rate 60 members produce too few successes for the normal approximation.
(c) Take the general lesson. The same dataset can justify one procedure and rule out another, so check the counts for the procedure you are actually running.
(a) Yes: and , both at least 10, so the test is justified. (b) No: the interval uses observed counts and is below 10. (c) A test checks the counts at , while an interval has no and must check them at .
Problem 3
A hospital employs 3,200 people and wants a 95% confidence interval for the mean one-way commute time of its employees. A random sample of 46 employees gives minutes (read 'x-bar', the sample mean) and minutes (the sample standard deviation). Nothing is known about the shape of the population of commute times. Check the conditions for a one-sample t-interval, then find the standard error.
Show the worked solution
Random. The 46 employees are a random sample of the hospital's employees, so the condition is met.
10%. Sampling is without replacement, and with , so the condition is met.
Sample data condition. The population shape is unknown, but , which satisfies the condition on its own. You do not need a plot and you do not need to assume the commute times are normal.
Say why is enough. By the central limit theorem, the sampling distribution of is approximately normal for a sample this size even when the population of individual commute times is skewed.
Pick the right family of curves. The population standard deviation (sigma) is unknown, so you use and a t-distribution with degrees of freedom , not a z-interval.
Find the standard error. minutes.
All three met: random sample; ; and , so no plot is needed. The one-sample t-interval is justified with and minutes.
Problem 4
A bookshop owner wants a confidence interval for the mean time a customer spends in the shop. She randomly selects 14 of the roughly 300 customers from one week and records their times. (a) The dotplot of the 14 times is roughly symmetric with no outliers. Are the conditions for a one-sample t-interval met? (b) Suppose instead the dotplot is strongly right-skewed with one customer at 95 minutes, far above the rest. Are the conditions met? (c) What would change your answer in (b)?
Show the worked solution
(a) Random and 10%. The 14 customers are randomly selected, and with , so both are met.
(a) Sample data condition. Here , so the sample size alone does not settle it and you fall back on the shape of the data. Roughly symmetric with no outliers means free from strong skewness and outliers, so the condition is met.
(a) Conclude. All three conditions are met, so a one-sample t-interval with is justified.
(b) Recheck each condition separately. Random and 10% are unchanged and still met, because changing the shape of the data does not change how the sample was chosen.
(b) Apply the sample data condition. With and a plot showing strong right skew plus a clear outlier at 95 minutes, the condition fails. The one-sample t-interval is not justified, so describe the data instead of reporting an interval.
(c) Name what would fix it. Either a larger sample, since removes the need to inspect the plot, or credible information that the population of visit times is approximately normal.
(c) Name what would not fix it. Lowering the confidence level, using a different formula for the standard error, or dropping the outlier without a documented reason. The condition is about the shape of the data relative to .
(a) Met: random, , and with a roughly symmetric plot with no outliers satisfies the sample data condition. (b) Not met: strong right skew and an outlier at , so the t-interval is not justified. (c) A sample of at least 30, or evidence that the population is approximately normal.
Problem 5
A city tests three recycling-bin labels. Sixty volunteers are randomly assigned to each label, 180 in all, and each volunteer sorts one bag of mixed waste. The number who sort correctly is 39 for Label A, 45 for Label B, and 30 for Label C.
| Label | Sorted correctly | Not sorted correctly | Row total |
|---|---|---|---|
| A | 39 | 21 | 60 |
| B | 45 | 15 | 60 |
| C | 30 | 30 | 60 |
| Total | 114 | 66 | 180 |
Check the conditions for a test for homogeneity, showing every expected count.
Show the worked solution
Confirm the table. The 'not sorted correctly' counts are , , and . Column totals are and , and .
Random. Volunteers were randomly assigned to the three labels, which meets the randomized-experiment version of the randomization condition for a test for homogeneity.
10%. This condition is not needed here, because the data come from a randomized experiment rather than from sampling without replacement out of a population.
Set up the expected counts. , computed under the null hypothesis that the three labels give the same distribution of outcomes.
Compute the 'sorted correctly' column. Every row total is 60, so each expected count is .
Compute the 'not sorted correctly' column. Each expected count is .
Check the arithmetic. Each row gives , and the columns give and , matching the observed totals.
Apply the rule. All six expected counts, , are greater than 5, so the expected counts condition is met and the test for homogeneity is justified. Note that the small observed count of 15 is irrelevant, since the condition is about expected counts.
Random: randomized experiment. 10%: not needed for a randomized experiment. Expected counts are 38 in each 'sorted correctly' cell and 22 in each 'not sorted correctly' cell, all greater than 5. The chi-square test for homogeneity is justified.
Problem 6
A city runs a randomized experiment at its parking meters. Of 180 drivers randomly assigned a paper receipt prompt, 117 pay for the full time; of 180 drivers randomly assigned a text prompt, 135 pay for the full time. Let and be the true full-payment proportions for the paper and text prompts. (a) Check the conditions for a two-sample z-test of . (b) Check the conditions for a two-sample z-interval for . (c) State which counts each procedure uses and why.
Show the worked solution
Find the sample proportions. for the paper prompt and for the text prompt.
(a) Random and 10%. Drivers were randomly assigned to the two prompts, which meets the randomized-experiment version of the condition, and the 10% condition is unnecessary for a randomized experiment.
(a) Pooled proportion. A test assumes , so the counts are checked at the combined proportion .
(a) Expected counts for the test. For each group, and . All four values are at least 10, so the two-sample z-test is justified.
(b) Observed counts for the interval. An interval estimates a difference and never assumes the proportions are equal, so it uses each group's own counts: 117 and for the paper prompt, 135 and for the text prompt.
(b) Compare to 10. All four observed counts are at least 10, so the two-sample z-interval is also justified.
(c) Explain the difference. The test pools because its null hypothesis says the two proportions share one common value, and the best estimate of that value uses both samples. The interval has no null value to borrow, so it keeps the groups separate.
(c) Note the shared discipline. Both procedures happened to pass here, but each required four specific counts written down, not the phrase 'large counts is met'.
(a) Justified: randomized experiment, and the pooled gives expected counts and in each group, all at least 10. (b) Justified: the observed counts 117, 63, 135, and 45 are all at least 10. (c) The test uses pooled counts, the interval uses each group's own observed counts.
Problem 7
A podcast host reads a link on air and asks listeners to visit a page and say whether they skip the ads. 1,500 listeners respond and 1,080 say they skip. The host wants to test against for all of the show's listeners, of whom there are several million. (a) Check each condition. (b) Is the test justified? (c) The host offers to leave the page open until 15,000 people respond. Does that fix the problem?
Show the worked solution
Find the sample proportion. .
(a) 10%. The show has several million listeners, so 1,500 is far below 10% of the population. This condition is met with room to spare.
(a) Large Counts. At , and . Both are far above 10, so this condition is met.
(a) Random. This one fails. Respondents selected themselves by choosing to visit the page, which makes this a voluntary response sample, not a random sample of the show's listeners.
(a) Say why the failure matters. Listeners with strong feelings about ads are the ones most likely to answer, so estimates the skip rate among people motivated to respond, not among all listeners.
(b) Conclude. The test is not justified. Two of the three conditions pass easily but that does not rescue the procedure, because randomization is the condition that lets you treat as an unbiased estimate of and lets the standard error formula describe the real sampling variability.
(c) Answer the sample size question. No. Bias does not shrink as grows; only the standard error does, since has in the denominator.
(c) Describe the result. With 15,000 voluntary responses you would get a narrower interval centered on the same wrong value, which is a more confident wrong answer rather than a better one.
(a) 10% and Large Counts pass easily ( and ), but Random fails: this is a voluntary response sample. (b) Not justified. (c) No, a larger voluntary response sample shrinks the standard error without touching the bias.
Problem 8
A county fair drew about 18,000 attendees. Organizers survey a random sample of 150 of them. Of those surveyed, 27 arrived on the free shuttle and 123 did not. The age breakdown of the 150 is: under 18, 30 people; 18 to 59, 96 people; 60 and over, 24 people. (a) Are the conditions met for a one-sample z-interval for the proportion of attendees who arrived on the shuttle? (b) Organizers also want a test for independence between age group and whether the attendee took the shuttle. Are those conditions met? (c) Explain how the same 150 responses can justify one procedure and not the other, and say what would fix (b).
Show the worked solution
(a) Random and 10%. The 150 attendees are a random sample, and with , so both are met.
(a) Normality at . An interval uses observed counts: successes and failures . Both are at least 10, so the condition is met and the one-sample z-interval is justified, with .
(b) Random and 10%. A test for independence needs one random sample from a single population, which is what the organizers have, and again, so both are met.
(b) Set up the expected counts. , where the column totals are 27 for the shuttle and 123 for no shuttle, and the row totals are 30, 96, and 24.
(b) Compute the shuttle column. Under 18: . Ages 18 to 59: . Ages 60 and over: .
(b) Compute the no-shuttle column. Under 18: . Ages 18 to 59: . Ages 60 and over: .
(b) Check the arithmetic. The shuttle column gives and the other column gives , matching the observed totals.
(b) Apply the rule. The expected count for '60 and over, shuttle' is not greater than 5, so the expected counts condition fails and the test for independence is not justified. One failing cell is enough to stop the procedure.
(c) Explain the split. Conditions attach to procedures, not to datasets. The interval only needs enough shuttle riders in the sample overall, and 27 clears the bar, while the chi-square test needs enough expected in every one of the six cells, and the 60-and-over group has only 24 people crossed with a column that is just 18% of the sample.
(c) Say what would fix it. A larger overall sample, or a design that surveys more attendees aged 60 and over, would raise that expected count above 5. Choosing a larger (alpha, the significance level) would not, and neither would pointing at the observed counts, since the rule is about expected counts.
(a) Met: random sample, , and observed counts 27 and 123 are both at least 10, so the z-interval is justified. (b) Not met: the expected count for '60 and over, shuttle' is , which is not greater than 5, so the chi-square test is not justified. (c) Conditions are checked per procedure; a larger sample, or more attendees sampled from the 60 and over group, would fix it.