Quick Read
An average can be calculated correctly and still describe the data badly.
The mean, median and mode answer different questions. Outliers can pull the mean away from what is typical. Two data sets can have the same average but very different spread. A useful statistical summary therefore asks not only “what is the average?” but also which average, how variable is the data, and what decision are we trying to make?
The repair is to teach statistics as interpretation rather than as a sequence of calculator commands.
An average compresses many values into one number.
Compression is useful because it makes a data set easier to discuss.
But every compression discards information.
The question is whether the information being discarded matters.
Mean, median and mode are not interchangeable
Students often learn all three under the heading “averages”.
That can make them feel like three alternative procedures for the same task.
They describe different features.
- Mean: distributes the total equally across all observations.
- Median: identifies the middle position after ordering the values.
- Mode: identifies the most frequently occurring value or category.
Which one is useful depends on the shape and purpose of the data.
Why the mean is powerful
The mean uses every numerical observation.
If the data represent quantities where every value should contribute to the summary, this can be desirable.
It is also algebraically useful and connects naturally to total quantities.
If 10 students have a mean score of 72, their total score is 720.
The mean therefore contains information about the whole data set, not only its centre.
Why the mean can be misleading
Consider monthly incomes:
3,000; 3,200; 3,300; 3,500; 30,000.
The very large value pulls the mean upward.
The mean remains mathematically correct, but it may not describe the experience of the typical person in that small group very well.
Correct calculation does not guarantee useful interpretation.
Why the median can be more robust
The median depends on order rather than the exact size of every value.
Extreme values therefore have less influence on it.
In skewed distributions such as income or property prices, the median can sometimes give a clearer picture of the middle observation.
But the median also discards information about how far values sit above or below that middle.
No measure is automatically best.
Why the mode answers a different kind of question
The mode is useful when frequency itself matters.
A shoe shop may care about the most commonly purchased shoe size.
A survey may care about the most frequently selected category.
The mean may not even make sense for non-numerical categories.
Difficulty 1: students calculate before asking what the data represent
Averages are often taught as procedures:
- add and divide for mean;
- order and find the middle for median;
- count frequencies for mode.
The stronger question comes first:
What feature of this data are we trying to summarise?
Difficulty 2: outliers are treated as mistakes automatically
An outlier is an unusually distant value.
It may be an error, but it may also be a genuine observation.
Students should not delete an outlier simply because it is inconvenient.
Ask:
- Could this value plausibly occur?
- Was there a measurement or recording error?
- How strongly does it affect the chosen summary?
- Does it reveal something important about the population?
Statistics requires judgement about data quality as well as arithmetic.
Difficulty 3: two data sets can have the same mean but feel completely different
Consider:
- Set A: 50, 50, 50, 50, 50.
- Set B: 10, 30, 50, 70, 90.
Both have a mean of 50.
But the distributions are very different.
Set A has no spread at all. Set B is widely dispersed.
Centre without spread can hide the shape of the data.
Why range matters
The range gives a simple measure of spread:
maximum − minimum.
It is easy to calculate and useful for comparing overall width.
But because it depends only on the two extreme values, it can itself be strongly affected by outliers.
Later statistical measures provide richer descriptions of variability.
Why graphs should be read together with averages
A histogram, box plot or other suitable graph can reveal information a single average hides.
- skewness;
- clusters;
- gaps;
- outliers;
- spread;
- multiple peaks.
Statistics becomes stronger when numerical summaries and visual representations agree with one another.
This connects to Why Do Graphs Feel Hard in Secondary Mathematics?.
Weighted averages are not ordinary averages
If different components contribute different proportions, the simple mean of the component scores may be wrong.
For example, an examination worth 70% should influence the final score more than a quiz worth 10%.
The weights represent how much each value contributes to the whole.
This links averages to proportional reasoning. See Why Are Ratio, Rate and Percentage So Easy to Confuse?.
Why grouped data introduces approximation
When raw values are grouped into intervals, the exact observations may no longer be visible.
Using class midpoints to estimate a mean creates an estimate rather than an exact reconstruction of the original data.
Students should understand where approximation entered the process instead of treating every calculator result as exact.
Why sample size matters
An average from three observations and an average from 3,000 observations may have very different stability and usefulness.
Small samples can be highly sensitive to unusual observations.
The average alone does not tell us how much evidence lies behind it.
Why an average can answer the wrong question
Suppose a parent asks, “What score does a typical student get?”
Calculating a mean may or may not answer that well depending on the distribution.
Suppose a school asks, “What was the total performance across all students?”
The mean may now be more directly useful.
Statistical methods should follow the decision being made.
How misleading presentations arise
Even correct statistics can be presented selectively.
- choosing mean instead of median because it tells a more favourable story;
- reporting an average without spread;
- removing inconvenient observations without justification;
- using a graph scale that exaggerates small differences;
- comparing averages from populations of very different size.
Students should learn that statistical literacy includes asking what has not been shown.
A practical data-reading routine
- Identify the variable. What was measured?
- Inspect the raw values or graph.
- Look for skew, clusters and outliers.
- Choose an appropriate measure of centre.
- Add a measure or description of spread.
- Interpret both in the original context.
Use “same average, different data” practice
Give students several data sets with the same mean and ask what differs.
Then give data sets with similar spread but different centres.
This teaches them to see summaries as partial descriptions rather than complete portraits.
Ask students to choose the average before calculating it
Present a context and ask which measure would be most informative.
Only then perform the calculation.
This reverses the common habit of calculating every available statistic and deciding afterward what it might mean.
Statistics and probability are related but not identical
Probability begins with a model of uncertainty and reasons about possible outcomes.
Statistics begins with observed data and tries to summarise or infer from it.
Both require careful attention to the reference set, assumptions and how much information is being compressed.
See Why Does Probability Feel So Counterintuitive in Mathematics?.
How parents can diagnose statistics difficulty
- Can the child explain the difference between mean, median and mode?
- Can they predict which measure an outlier will affect most?
- Can they compare two data sets with the same mean but different spread?
- Do they understand when a weighted mean is needed?
- Can they interpret an average in a sentence rather than merely calculate it?
- Can they identify when a graph or summary is telling only part of the story?
When tuition can help
Tuition can help when statistics has become button-pressing without interpretation, when students choose mean, median or mode mechanically, or when they cannot connect numerical summaries to the shape and context of the data.
The aim should be to turn statistical calculations into evidence for a reasoned description.
Frequently Asked Questions
Is the mean the “real” average?
The mean is one important measure of centre, but median and mode answer different questions and can be more useful in particular distributions or contexts.
Why do outliers matter?
Extreme values can strongly affect some summaries, especially the mean and range. They may also contain important information about the population or reveal data-quality problems.
Can two groups have the same average and still be very different?
Yes. Their spread, shape, outliers and distribution can differ substantially even when a measure of centre is identical.
Final Thought: an average is a summary, not the data itself
The strongest student does not stop after calculating one central number.
Inspect the data → choose the right centre → examine spread → notice outliers → interpret in context → ask what the summary leaves out.
That is how averages become useful evidence rather than convenient but misleading answers.
Diagnostic routes: Find My Mathematics State · Mathematics Diagnosis · complete Mathematics directory.

