Measures of the Center of the Data Study Pack

Kibin's free study pack on Measures of the Center of the Data includes a 7-section study guide, 25 quiz questions, 30 flashcards, and 5 open-ended Explain review questions. Sign up free to track your progress toward mastery, plus upload your own notes and recordings to create personalized study packs organized by course.

Last updated May 28, 2026

Topic mastery0%

Measures of the Center of the Data Study Guide

Master the three measures of center — mean, median, and mode — and learn how each responds to outliers, skewness, and weighted values. Understand when to use the median over the mean for skewed data like income, and how symmetric vs. skewed distributions shift these measures apart.

Key Takeaways

  • The three primary measures of center — mean, median, and mode — each describe where data tends to cluster, but they respond differently to skewness and outliers.
  • The mean is calculated by summing all values and dividing by the count; it is the most mathematically precise measure but is pulled toward extreme values.
  • The median is the middle value when data are ordered and is resistant to outliers, making it the preferred measure of center for skewed distributions such as household income.
  • The mode identifies the most frequently occurring value and is the only measure of center applicable to purely categorical (nominal) data.
  • In a perfectly symmetric distribution, the mean, median, and mode coincide; in a right-skewed distribution the mean exceeds the median, and in a left-skewed distribution the mean falls below the median.
  • The weighted mean accounts for values that contribute unequally to a dataset, multiplying each value by its assigned weight before summing and dividing by total weight.
  • Measures of center alone do not fully describe a distribution — spread statistics such as standard deviation are needed to capture variability around the center.

What Measures of Center Represent

A measure of center is a single number that summarizes where the bulk of a dataset is located on a number line, giving a concise description of the data's typical value.

Purpose of a Central Value

  • A measure of center condenses an entire distribution into one representative number, enabling quick comparisons between datasets.
  • No single measure is universally best — the appropriate choice depends on the data's level of measurement and the shape of its distribution.

Levels of Measurement and Applicability

  • The mode can describe any data type, including nominal categories such as favorite color or political affiliation.
  • The median requires at least ordinal data — values must be rankable in order — but does not require meaningful arithmetic differences between values.
  • The mean requires interval or ratio data because it depends on the actual numerical distances between values.

The Mean: Arithmetic Average and Its Variants

The mean is the most commonly used measure of center and is computed by adding every value in a dataset and dividing by the total number of values.

Calculating the Arithmetic Mean

  • For a sample of n values, the sample mean (written x̄) equals the sum of all observations divided by n: x̄ = (Σx) / n.
  • For a population of N values, the population mean is denoted μ (mu) and uses the same formula with N in the denominator.
  • Every data point contributes equally to the arithmetic mean, so a single extreme outlier can shift the mean substantially.

The Weighted Mean

  • When different data points carry different levels of importance or frequency, a weighted mean assigns each value xᵢ a weight wᵢ.
  • The weighted mean equals (Σ wᵢxᵢ) / (Σ wᵢ), ensuring that more heavily weighted values exert proportionally greater influence on the result.
  • Grade point averages are a classic example: a 4-credit course contributes more to the GPA calculation than a 1-credit course.

Estimating the Mean from Grouped Data

  • When only a frequency table is available rather than raw data, each class midpoint is used as a representative value for that interval.
  • Multiplying each midpoint by its frequency, summing those products, and dividing by total frequency yields an approximate mean.

The Median: The Middle of Ordered Data

The median divides a rank-ordered dataset into two equal halves, with 50 percent of values falling at or below it and 50 percent at or above it.

Finding the Median

  • Sort all values from smallest to largest before locating the median — order is essential.
  • If the dataset has an odd number of values, the median is the single middle value at position (n + 1) / 2.
  • If the dataset has an even number of values, the median is the arithmetic mean of the two middle values at positions n/2 and (n/2) + 1.

Resistance to Outliers

  • Because the median depends only on rank position rather than actual magnitude, extreme high or low values do not pull it toward them.
  • This resistance makes the median the standard measure of center for income and home price data, which are typically right-skewed with very high outliers.

Median in Frequency Distributions

  • In a grouped frequency distribution, the median falls in the class interval that contains the cumulative 50th percentile, though its exact value within that interval requires interpolation.

The Mode: Most Frequent Value

The mode identifies the value or category that appears more often than any other in a dataset, making it particularly useful when frequency of occurrence is the central question.

Identifying the Mode

  • The mode is found by tallying how often each distinct value appears and selecting the one with the highest count.
  • A dataset can have one mode (unimodal), two modes of equal frequency (bimodal), or more than two (multimodal); a dataset where every value appears exactly once has no mode.

When to Use the Mode

  • The mode is the only appropriate measure of center for nominal data — for example, identifying the most commonly purchased product color in a retail dataset.
  • In continuous numerical data, the mode is less informative because values rarely repeat exactly; it is more meaningful when data are grouped into intervals.

Mode in Frequency Distributions

  • When data are organized into a frequency table, the modal class is the interval with the highest frequency, even if no single repeated value can be named.

Unlock the rest of this study guide

  • Access the full study pack
  • Track your mastery and be test-day ready
  • Upload your own notes to build personalized study guides, quizzes, flashcards, and more
Sign up free →

About this Study Pack

Created by Kibin to help students review key concepts, prepare for exams, and study more effectively. This Study Pack was checked for accuracy and curriculum alignment using authoritative educational sources. See sources below.

Sources

More in Statistics

See all topics →

Browse other courses

See all courses →