07 Confidence Intervals
Learn how confidence intervals quantify uncertainty, how to select the correct interval formula, how to plan sample sizes, and how to check assumptions and interpretations.
From point estimates to intervals
A combines a sample-based estimate with information about how much that estimate may vary across random samples.
The basic structure
A has the form
The identifies the center of the interval, while the determines how far the endpoints extend from that center. For example, a sample mean estimates a population mean, but another random sample would usually produce a different sample mean. An interval communicates this uncertainty more fully than a single number.
A is a single sample statistic used to estimate a population parameter:
The sample mean estimates the population mean .
The sample proportion estimates the population proportion .
The sample standard deviation estimates the population standard deviation .
Takeaway: A gives the best single-number estimate from the sample; a adds a range that reflects sampling uncertainty.
Confidence levels and critical values
The describes the long-run performance of an interval-producing method. If the same population were sampled repeatedly and the same method were applied each time, approximately the chosen percentage of the resulting intervals would contain the true parameter.
For a of , is the total probability placed in the two tails of the sampling distribution. Common standard-normal critical values are:
confidence: and .
confidence: and .
confidence: and .
A confidence procedure does not mean that a particular completed interval has a probability of containing the fixed parameter. After the interval is calculated, it either contains the parameter or it does not. The describes the success rate of the method over repeated samples.
Holding sample size and variability constant, increasing the increases the critical value and makes the interval wider.
Takeaway: Higher confidence provides a more reliable long-run procedure but requires a wider interval.
How interval width is determined
The is the distance between the and either endpoint of a symmetric interval. It is determined by the critical value and the standard error:
Thus, the interval can be written as
For many estimators, the standard error decreases in proportion to . This has an important planning consequence: reducing the by half generally requires about four times as many observations.
The tends to decrease when:
the sample size increases;
population variability decreases; or
a lower is selected.
Do not confuse the with the full interval width. If the is , the width of a symmetric interval is .
Takeaway: Sample size controls precision through a square-root relationship, while and variability also affect interval width.
Intervals for a population mean
For a population mean, the correct formula depends on whether the population standard deviation is known.
Known population standard deviation
When is known, use
The is
This method requires independent observations and an approximately normal sampling distribution for the mean. A normally distributed population makes the method exact; for large samples, the Central Limit Theorem often supports the normal approximation.
Unknown population standard deviation
When is unknown, replace it with the sample standard deviation and use the :
The degrees of freedom are
The heavier tails of the account for the additional uncertainty caused by estimating . As increases, the becomes increasingly similar to the standard normal distribution.
Worked example
Suppose , , and . For a interval with , take . Then
Therefore,
The appropriate interpretation is: using this sampling method, a for the population mean is from to units.
Takeaway: Use when the population standard deviation is known; use with degrees of freedom when it is unknown.
Intervals for a population proportion
A population proportion is estimated by the sample proportion
where is the number of successes and is the sample size. For a sufficiently large sample, the introductory normal-approximation interval is
Its is
The approximation is most reliable when the expected numbers of successes and failures are both adequately large. For small samples or proportions near or , a Wilson or exact binomial interval is generally preferable to the simple symmetric normal interval.
Worked example
In a random sample of voters, support a proposal. The sample proportion is
Using for a interval,
The interval is
Thus, the estimated population proportion is , and the approximate interval extends from to .
Takeaway: Check the success and failure counts before relying on the normal approximation for a proportion.
Planning sample size
Sample size must be planned before data collection by specifying the , the largest acceptable , and an estimate of population variability. Always round the result up, because rounding down could make the actual larger than requested.
Planning for a mean
When a planning value for the population standard deviation is available, use
For example, with confidence, , and ,
Rounding up gives .
Planning for a proportion
With a prior planning estimate , use
If no prior estimate is available, use . This is conservative because is largest at , producing the largest required sample size.
For confidence and with no prior estimate,
Rounding up gives .
When sampling without replacement from a relatively small, known population, a can reduce the required sample size:
Here, is the sample size calculated as if the population were very large, and is the population size.
Takeaway: Set the desired precision first, calculate the required size, and round up.
Assumptions and error checks
Before interpreting an interval, verify that the method matches the design and data.
Conditions to check
Randomness: The sample should reasonably represent the target population.
Independence: Observations should be independent, or the sampling design must account for dependence.
Distributional conditions: The normal or t-based method should be appropriate for the sample size and population shape.
Proportion counts: A normal approximation should not be used automatically when success or failure counts are very small.
Target parameter: An interval for a mean estimates a population mean, not the range of individual observations.
Frequent interpretation and calculation errors
Treating a as a probability for one fixed interval.
Using when is unknown and the sample is small, instead of using .
Rounding a required sample size down.
Confusing the with the full interval width.
Assuming that a higher produces a narrower interval.
A useful final check is to identify the parameter, the , the critical value, the standard error, and the assumptions supporting the method. Together, these determine whether the interval is both numerically correct and appropriately interpreted.
Final takeaway: A is meaningful only when its sampling method, formula, assumptions, and interpretation all agree.