05 Common Statistical Distributions
A structured guide to recognizing, interpreting, calculating with, and approximating common discrete and continuous probability distributions.
Foundations: Discrete and Continuous Models
Probability distributions are mathematical models for random variables. Begin by identifying whether the variable is discrete or continuous.
A discrete random variable takes countable values, such as the number of defective items. Its probabilities are assigned to individual outcomes with a .
A continuous random variable can take any value in an interval, such as a waiting time. Its probabilities are areas over intervals under a .
For a continuous variable, the probability of one exact value is zero; meaningful probabilities concern ranges of values.
A distribution is commonly described by its:
Mean or expected value, which indicates the center;
Variance and standard deviation, which describe spread;
Shape, including symmetry, skewness, and tail behavior;
Support, the set of values the variable can take.
Takeaway: First determine the type of random variable and the event being measured. Then check whether the distribution's assumptions fit the process.
Counting Successes:
Use the binomial model when you count successes across a predetermined number of trials. If counts successes in trials with constant success probability , write
The four required conditions are:
The number of trials is fixed.
Each trial has two possible outcomes, designated success and failure.
The trials are independent.
The success probability is the same on every trial.
The probability of exactly successes is
The mean, variance, and standard deviation are
For example, if each of 50 components independently has defect probability , the number of defects follows . The probability of exactly two defects is
and the expected number of defects is .
Do not use this model when the number of trials is not fixed, trials are dependent, or the success probability changes. Sampling without replacement from a small population may instead require a hypergeometric model.
Takeaway: “How many successes in a fixed number of independent, identical trials?” points to the .
Waiting for the First Success
Use the geometric model when the experiment continues until the first success. If each trial has success probability , then
The probability that the first success occurs on trial is
The first trials must fail, followed by a success on trial . Useful cumulative probabilities are
Its mean and variance are
For example, if a customer makes a purchase independently with probability on each visit, the probability of a first purchase on the fourth visit is
and the expected number of visits is .
The has the :
Thus, after any number of failures, the chance of needing more than an additional number of trials is unchanged.
Takeaway: “How many trials until the first success?” points to the .
Equal Likelihood Across an Interval
A continuous variable follows a uniform model on when all subintervals of equal length are equally likely. Its probability density is
For , the cumulative probability is
For an interval ,
The mean and variance are
For a waiting time uniformly distributed from zero to 10 minutes, the probability of waiting at most three minutes is
Takeaway: Uniform models are appropriate when probability is proportional to interval length throughout a finite range.
Symmetric Continuous Measurements
The normal model is continuous, symmetric, and bell-shaped. It is determined by its mean and standard deviation :
Its density is
For this distribution, the mean, median, and mode are equal. Approximately 68.27 percent of observations lie within one standard deviation of the mean, 95.45 percent lie within two standard deviations, and 99.73 percent lie within three standard deviations. These are useful benchmarks, while exact probabilities require the normal CDF or statistical software.
For example, if adult height is modeled by , the probability of a height no greater than 71 inches is found by standardizing:
Takeaway: The is a natural model for symmetric, bell-shaped continuous measurements, but its fit must be assessed rather than assumed.
Standardization and Normal Probabilities
The has mean zero and standard deviation one:
Its density and cumulative distribution function are
To convert into a standard normal variable, calculate the :
Then
For an interval,
Because the standard normal curve is symmetric about zero,
For instance,
A reliable workflow is to identify the mean and standard deviation, standardize each boundary, obtain CDF values from a table or calculator, and subtract for an interval. Use complements or symmetry for upper-tail probabilities.
Takeaway: Standardization puts different normal variables on the same reference scale, making probability calculations manageable.
Approximating Binomial Probabilities
A can sometimes be approximated by a when both expected success and expected failure counts are sufficiently large. A common guideline is
For , use a normal variable with
so its standard deviation is . The approximation is less dependable when is very close to zero or one or when is small.
Because the binomial variable is discrete and the normal variable is continuous, apply a :
;
;
.
The interval from to represents the discrete value . After making the correction, standardize the normal boundary or boundaries and use .
Takeaway: Check the approximation conditions, match the binomial event to the correct half-unit boundary, and then use normal-probability methods.
Selecting the Right Distribution
Choose a distribution by matching the question to the random mechanism:
Count successes in a fixed number of independent trials: .
Count trials until the first success: .
Model a value that is equally likely anywhere in a finite interval: .
Model a continuous, symmetric, bell-shaped measurement: .
Work with a standardized normal value: .
Before calculating, ask:
Is the variable discrete or continuous?
What is being counted or measured?
Is the number of trials fixed, or does the process stop at the first success?
Are trials independent and is the success probability constant?
What are the support, center, spread, and shape?
Does the model's mechanism match the data-generating process?
Probability distributions are models, not guarantees. Their assumptions should be checked before use. They provide foundations for confidence intervals, hypothesis tests, simulation, and broader statistical modeling.
Final summary: Discrete models assign probability to countable outcomes, while continuous models assign probability to intervals. Binomial and geometric distributions describe success-based trials; uniform and normal distributions describe continuous measurements. Standardization and connect these models to practical probability calculations.