8.4 Standard Deviation & The Normal Distribution (Empirical Rule)
Key Takeaways
Sample variance () and sample standard deviation () use Bessel's correction () to provide an unbiased estimate of population dispersion, measuring the typical distance of data points from the mean.
The standard normal distribution is a continuous, unimodal, bell-shaped density curve perfectly symmetric about mean where and the total area under the curve equals ().
The Empirical Rule (68–95–99.7% Rule) states that for bell-shaped distributions: approximately of data falls within , falls within , and falls within .
Exact regional tail slice percentages partitioned by standard deviations are: ( to ), ( to ), ( to ), and ().
The standardized z-score () measures the exact number of standard deviations a data value lies above () or below () the mean, enabling objective comparison across different distributions.
8.4 Standard Deviation & The Normal Distribution (Empirical Rule)
In statistical analysis, standard deviation is the universal benchmark metric for quantifying numerical dispersion. When data values follow a symmetric, bell-shaped distribution, the standard deviation pairs with the mean to completely define the Normal Distribution. On the CLEP College Mathematics exam, questions extensively test sample vs. population standard deviation, the Empirical Rule (68–95–99.7% Rule), regional slice calculations, and standardized z-scores.
1. Variance & Standard Deviation
Standard deviation measures the typical or average distance that data points deviate from their arithmetic mean. Because simple deviations always sum to zero, deviations are squared before averaging.
Sample vs. Population Formulas
| Statistical Parameter | Population () | Sample () |
|---|---|---|
| Variance | ||
| Standard Deviation | ||
| Units of Measure | Squared original units (e.g., ) | Identical to original data units (e.g., ) |
Why Divide by ? (Bessel's Correction)
When computing sample variance from a subset of a larger population, using the sample mean instead of the true population mean tends to underestimate the true dispersion. Dividing by the degrees of freedom (Bessel's correction) slightly inflates the variance, providing an unbiased estimator of the population variance .
Step-by-Step Manual Calculation Protocol
- Calculate the sample mean: .
- Compute the deviation of each value from the mean: .
- Square each deviation: .
- Sum all squared deviations: .
- Divide by to obtain the sample variance .
- Take the positive square root to obtain the sample standard deviation .
2. Properties of the Normal Distribution
The Normal Distribution (Gaussian distribution) is a continuous probability distribution defined by its mean () and standard deviation ().
+-----------------------------------------------------------------------------+
| PROPERTIES OF THE NORMAL DISTRIBUTION CURVE |
| |
| 1. BELL-SHAPED & CONTINUOUS: Smooth, symmetric curve over (-inf, +inf). |
| 2. PERFECT SYMMETRY: Mean = Median = Mode at the exact central peak (μ). |
| 3. TOTAL PROBABILITY AREA = 1.0 (100%): Exactly 50% lies below μ and |
| 50% lies above μ. |
| 4. ASYMPTOTIC: The tails extend infinitely in both directions, approaching|
| the horizontal axis closely but never touching or crossing it. |
| 5. INFLECTION POINTS: The curve transitions from concave downward to |
| concave upward at exactly μ - σ and μ + σ. |
+-----------------------------------------------------------------------------+
3. The Empirical Rule (68–95–99.7% Rule)
For any dataset that is normally distributed (or approximately bell-shaped), the Empirical Rule provides exact percentage approximations based on standard deviation intervals away from the mean:
+-----------------------------------------------------------------------------+
| THE EMPIRICAL RULE (68 - 95 - 99.7%) |
| |
| [------ 68.2% ------] |
| [------------- 95.4% -------------] |
| [-------------------- 99.7% --------------------] |
| |
| 34.0% 34.0% |
| 13.5% 13.5% |
| 2.35% 2.35% |
| 0.15% 0.15% |
| --+-------+-------+-------+-------+-------+-------+-------+-- |
| μ-3σ μ-2σ μ-1σ μ μ+1σ μ+2σ μ+3σ |
+-----------------------------------------------------------------------------+
Regional Slice Breakdown Table
Dividing the normal curve into standardized one-standard-deviation slices yields the universal values tested on the CLEP:
| Interval from Mean | Standard Deviation Range | Slice Percentage | Cumulative Area from |
|---|---|---|---|
| Extreme Lower Tail | Below | ||
| Outer Lower Slice | to | (Lower Cutoff) | |
| Inner Lower Slice | to | (Lower Cutoff) | |
| Central Lower Slice | to | (Median) | |
| Central Upper Slice | to | (Upper Tail) | |
| Inner Upper Slice | to | (Upper Tail) | |
| Outer Upper Slice | to | ||
| Extreme Upper Tail | Above |
4. Standardized Z-Scores
A standardized z-score indicates how many standard deviations a particular observation () lies above or below the mean.
Z-Score Formula
Interpretation of Z-Scores
- : The data value is exactly equal to the mean.
- (Positive): The data value lies above the mean (e.g., is standard deviations above average).
- (Negative): The data value lies below the mean (e.g., is standard deviations below average).
- Standard Normal Distribution: Converting all raw scores to -scores transforms any normal distribution into the standard normal distribution with mean and standard deviation .
Comparing Relative Standing Across Different Distributions
Z-scores allow direct, objective comparisons between individuals scored on completely different scales (e.g., comparing an ACT score of with an SAT score of ).
5. Step-by-Step Worked Examples
Worked Example 1: Calculating Sample Variance and Standard Deviation
Problem: A quality control technician measures the weight (in grams) of sample metal components: . Calculate the sample mean , sample variance , and sample standard deviation .
Solution:
- Sample mean: .
- Compute squared deviations:
| Value () | Deviation () | Squared Deviation |
|---|---|---|
| Total |
- Sample variance ():
- Sample standard deviation:
Worked Example 2: Empirical Rule Analysis
Problem: Scores on a national mathematics exam are normally distributed with a mean of and a standard deviation of . Out of test takers:
- What percentage of students scored between and ?
- Approximately how many students scored above ?
- What score corresponds to the percentile?
Solution:
- Identify standard deviation milestones:
- Sum the corresponding regional slices:
- Score above ():
- Tail percentage
- Number of students
- percentile:
- Cumulative area up to
- Score
Worked Example 3: Comparative Z-Score Analysis
Problem: Jessica scored on an English literature exam (class mean , standard deviation ). On her Chemistry exam, she scored (class mean , standard deviation ). On which exam did Jessica perform better relative to her peers?
Solution:
- Calculate Jessica's English z-score:
- Calculate Jessica's Chemistry z-score:
- Comparison: Jessica scored standard deviations above average in English, but standard deviations above average in Chemistry. Because , Jessica performed better relative to her peers in Chemistry.
6. Common CLEP Traps & Strategic Checkpoints
- Trap 1: Applying the Empirical Rule to Non-Normal Distributions: The 68–95–99.7% Rule applies only to bell-shaped, symmetric normal distributions. It cannot be used for uniform or skewed data.
- Trap 2: Dividing by Instead of for Samples: For sample standard deviation (), always divide the sum of squared deviations by .
- Trap 3: Forgetting Negative Signs on Z-Scores: If a raw score is below the mean (), the z-score MUST be negative (e.g., ). Omitting the negative sign reverses the percentile ranking.
- Trap 4: Confusing Variance with Standard Deviation: Variance is in squared units (). To find the standard deviation, you must take the square root ().
The distribution of heights of adult males in a region is normally distributed with a mean of μ = 70 inches and a standard deviation of σ = 3 inches. Using the Empirical Rule (68–95–99.7% Rule), what percentage of adult males in this region have heights between 64 inches and 76 inches?
68%
81.5%
95%
99.7%
A student scores 84 on a biology exam where the class mean is 76 with a standard deviation of 4. On a chemistry exam, the student scores 88 where the class mean is 80 with a standard deviation of 5. On which exam did the student perform better relative to the class, and what are the corresponding z-scores?
Chemistry (z = 2.00 vs. Biology z = 1.60)
Biology (z = 2.00 vs. Chemistry z = 1.60)
Both performances were identical (z = 1.80)
Biology (z = 1.60 vs. Chemistry z = 2.00)
When calculating the sample standard deviation s from a data set of n = 6 observations, why is the sum of squared deviations divided by (n - 1) = 5 instead of n = 6?
To convert the variance from a population parameter to a standardized z-score.
To ensure the standard deviation is always smaller than the interquartile range.
To eliminate the effect of any negative numbers that were squared during the calculation.
To correct for Bessel's bias and provide an unbiased estimator of the true population variance.
Sections you finish are checked off in the contents.