Sample proportion vs population proportion
By Jude Wallis · Published
The population proportion p is a fixed parameter describing the whole group, and you almost never know it. The sample proportion p-hat is a statistic you compute from one sample, so it changes from sample to sample. p-hat estimates p: its values center on p, with a spread set by the sample size.
AP Statistics: Unit 3 (topics 3.1 Estimators, 3.2 Sampling Distributions for Sample Proportions). In the Fall 2026 AP Statistics course, this pair opens Unit 3, Inference for Categorical Data: Proportions. Topic 3.1 is titled Estimators and topic 3.2 is titled Sampling Distributions for Sample Proportions, so the estimator and the distribution of its values around are the unit's first two topics, in that order.
p vs p-hat: the short answer
Both are proportions, both are written with the letter p, and only one of them is a number you will normally get to compute.
The population proportion is written , said "p". It is the fraction of the entire population with the characteristic you care about. It is a parameter: one fixed number that exists whether or not anyone measures it, and in practice you do not know its value.
The sample proportion is written , said "p-hat". It is the fraction of your sample with that characteristic. It is a statistic: you can compute it from the data in front of you, and a different sample gives a different value.
So is what you want to know and is what you have. The hat is the marker, and it always means sample. For the wider vocabulary of parameters and statistics, see parameter vs statistic.
What p is: fixed, and usually unknown
The population proportion is . Right now, for a fully defined population, that fraction has one exact value.
What makes hard is access, not existence. To know it you would have to check every member of the population, which is a census, and a census is the exception rather than the rule (see census vs sample survey). Everywhere else, stays unknown and you reason toward it.
One consequence shows up on every inference question you will write. Hypotheses are claims about , never about . Writing asserts something about the population that data can bear on. Writing asserts something about a number you already have in your hand, which is not worth testing.
What p-hat is: computable, and different every time
The sample proportion is , where counts the successes in your sample and is the sample size. Both are counts of individuals, so always lands between 0 and 1, and it can only take the values , , , and so on.
Because it depends on which individuals happened to land in your sample, is a moving target. Draw again and you get a different value. That is sampling variability, not a mistake, and it is the thing the rest of this page is about. The mechanics of getting out of a word problem, a two-way table, or a reported percentage are worked through in how to find p-hat.
The differences side by side
| Feature | Population proportion | Sample proportion |
|---|---|---|
| Kind of quantity | Parameter | Statistic |
| What it describes | Every member of the population | Only the individuals you sampled |
| Value | Fixed | Changes from sample to sample |
| Do you know it | Almost never | Yes, you computed it |
| How you would get it | A census of the population | from your data |
| Role in inference | The unknown a hypothesis is about | The evidence you compute |
Every row is the same split. One number is the target, the other is the estimate aimed at it.
One population, five samples, five p-hats
A district has 2,000 seniors, and district records show that 1,200 of them hold a driver's license. So , and for once you actually know it. Now take five separate random samples of seniors.
| Sample | Licensed, | ||
|---|---|---|---|
| 1 | 34 | 0.68 | +0.08 |
| 2 | 28 | 0.56 | -0.04 |
| 3 | 31 | 0.62 | +0.02 |
| 4 | 27 | 0.54 | -0.06 |
| 5 | 30 | 0.60 | 0.00 |
Five samples, five different answers, and the entire time. The population never moved. Nothing went wrong in samples 1 through 4; different seniors landed in different samples, so the fractions differ.
Two things are worth staring at. Sample 5 hit exactly, and from inside that sample there is no way to tell: all five look identical from the inside, 50 seniors and a fraction. And these five values happen to average to exactly 0.60, which is a nicety of the five I picked, not a demonstration. Being unbiased is a claim about every possible sample, not about a handful of them.
The sampling distribution is the whole payload
Collect from every possible sample of size and you get the sampling distribution of the sample proportion. When the observations are independent, it has
Read as "mu sub p-hat" and as "sigma sub p-hat". Those two lines are the entire relationship between the two proportions, and every proportion procedure you will run is built on them.
The first says is unbiased: across all possible samples its values center on , with no systematic pull above or below. The second says is variable, and says exactly how variable. For the district above, , then , and . All five sample proportions in the table sit within 0.08 of , which is about 1.15 of those standard deviations.
Unbiased and variable at once is the point. One is usually not , and on the samples where it is you have no way of knowing. What you get instead is a known amount of scatter around , and that known amount is what a margin of error and a p-value are both computed from. Raising shrinks it: at , , exactly half of 0.069282, so quadrupling the sample halves the standard deviation.
Before leaning on the normal shape, check two conditions. Large Counts wants and , here 30 and 20 (see why 10 successes and 10 failures). Independence is usually argued with the 10% condition, sampling no more than a tenth of the population, here 50 out of 2,000.
The mix-ups that cost points
- Reporting as if it were . "68% of district seniors are licensed" is a claim about 2,000 people built from 50 of them. Say "68% of the sampled seniors" and attach the sample size.
- Writing inside a hypothesis. Hypotheses are about , so , never .
- Substituting for in without saying so. In a test you have a hypothesized and you use it. In a confidence interval you do not, so you put in its place and the result is called the standard error, not the standard deviation. The distinction is spelled out in standard error vs standard deviation.
- Reading a new as evidence that changed. Samples 2 and 3 above reported 0.56 and 0.62 from a population that never moved, so check a gap against before reading it as a real shift: at a 0.06 gap is 0.61 standard deviations and unremarkable, at it is 2.73 and hard to explain away.
- Blurring the two groups behind the two symbols. If the population and the sample are not clear in your head, start at population vs sample.
Five samples, one parameter, and the spread to expect
A district has 2,000 seniors, 1,200 of whom hold a driver's license. Five separate random samples of seniors contain 34, 28, 31, 27, and 30 licensed seniors. Find the population proportion and each sample proportion, then find the mean and standard deviation of the sampling distribution of for and check the conditions.
Population proportion: . This is a parameter, and it is the same number for all five samples because the district did not change.
Sample proportions, one per sample: , , , , . Five statistics, five values.
Center of the sampling distribution: . That is what unbiased means, and it does not depend on the sample size.
Spread, digit by digit. Multiply: . Divide by : . Square root: , so .
Large Counts: and . Both are at least 10.
Independence via the 10% condition: , which is below the population of 2,000, so sampling 50 seniors without replacement is fine.
Compare the misses. The largest is sample 1: , and , about 1.15 standard deviations. Under the normal approximation roughly 75% of samples land within 0.08 of , so five out of five doing so is unremarkable.
for all five samples. The sample proportions are 0.68, 0.56, 0.62, 0.54, and 0.60. The sampling distribution of has and , both conditions hold, and the largest deviation is 1.15 standard deviations.
Two samples, two intervals, one p
Now suppose is unknown. Sample 1 above found 34 licensed seniors out of 50, and sample 4 found 27 out of 50. Build a 95% confidence interval for from each, using , and check whether each captures the true value 0.60.
Sample 1: . Because is unknown, put where sits in the standard deviation formula; the result is the standard error.
Multiply: . Divide by 50: . Square root: .
Margin of error: . Interval: , so .
Sample 4: . Multiply: . Divide by 50: . Square root: .
Margin of error: . Interval: , so .
Compare. The two intervals have different centers and different widths, because moved and the standard error moved with it. Both contain 0.60. The interval is what shifts around; sits still.
Sample 1 gives and the interval . Sample 4 gives and the interval . Both capture , from two samples that disagreed by 0.14.
Frequently asked questions
Which symbol is the sample one, p or p-hat?
(p-hat) is the sample one. The hat marks a value computed from a sample, so it is a statistic you know. Plain is the population proportion, a parameter that is fixed and almost always unknown.
Can p-hat ever equal p exactly?
Yes, and it happens fairly often with a small , since can only take values . In the district example above one of five samples landed exactly on 0.60. The catch is that you cannot tell from inside the sample, so it changes nothing about how you report the result.
Does a bigger sample make p-hat closer to p?
It shrinks the typical distance, but it does not guarantee any one sample is closer. The standard deviation is inversely proportional to , so quadrupling halves it: 0.069282 at becomes 0.034641 at when . Size does not fix a biased sampling method, which is a separate problem covered in does a bigger sample fix bias.
Why do hypotheses always use p and never p-hat?
A hypothesis is a claim about the population, and is the only one of the two symbols that describes the population. is the number you compute from your sample to test that claim, so it belongs in the test statistic, not in or .
What is the difference between sigma sub p-hat and the standard error?
They are the same formula with different inputs. uses the true or hypothesized , which you have in a significance test. The standard error substitutes your sample value, which is what a confidence interval has to do because is unknown there.