Business Mathematics and Statistics · Ch 8 — Descriptive Statistics and Probability
Measures of Central Tendency — Mean, Median and Mode
Measures of Central Tendency — Mean, Median and Mode
Business decisions are constantly made from data — average sales, typical wages, the most common size of order — and the first job of statistics is to reduce a whole mass of raw numbers to a single representative figure. A measure of central tendency is exactly such a figure: a single value that best represents the entire data set.
Arithmetic Mean. For ungrouped observations : . For data grouped into classes with frequencies and mid-points : . The mean uses every single value in the data set, which makes it the most commonly used average, but also the one most distorted by an extreme value (an outlier).
Median. The middle value of the data when arranged in ascending (or descending) order — for ungrouped data with observations, the median is the value if is odd, or the average of the and values if is even. Unlike the mean, the median is not affected by how extreme the highest or lowest values are, only by how many observations lie above and below it — this makes it the preferred average for skewed data, such as income or wealth figures with a few very large values.
Mode. The value that occurs most frequently in the data set. A data set can have one mode (unimodal), two (bimodal), or more; it is the only one of the three averages that can meaningfully describe purely categorical/qualitative data (e.g. 'the most commonly purchased size').
A Useful Empirical Relationship
For a moderately asymmetrical (skewed) distribution: — a quick way to estimate whichever of the three is not directly available, and a useful cross-check on all three when all are computed.
Computing and choosing between mean, median and mode is a universal first step of descriptive statistics, identical in method across every Indian commerce board's statistics syllabus.
The sum of all observations divided by the number of observations: (ungrouped) or (grouped).
The middle value of a data set arranged in order; for even, the average of the two middle values. Unaffected by the size of extreme values.
The value occurring with the highest frequency in a data set; the only central-tendency measure applicable to purely categorical data.