Master10
General Science20 Concepts & Facts

What Is the Interquartile Range (IQR)? Box Plots & Outlier Detection

Reviewed by the Master10 Editorial Board for accuracy, clarity and competitive-exam relevance.Editorial Policy
In descriptive statistics, the interquartile range, commonly abbreviated as IQR, is a measure of statistical dispersion that quantifies the spread of the middle fifty percent of an ordered dataset. To compute the IQR, a sample of numerical observations is arranged in ascending order and divided into four equal parts using three cut-off points called quartiles. The first quartile, designated as Q1 or the twenty-fifth percentile, separates the lowest quarter of the data from the rest. The third quartile, designated as Q3 or the seventy-fifth percentile, separates the lowest three-quarters from the top quarter. The interquartile range is calculated by subtracting Q1 from Q3.

Unlike traditional measures of dispersion such as the range, variance, or standard deviation, the interquartile range is a robust non-parametric statistic. Standard deviation and variance depend heavily on the arithmetic mean, making them vulnerable to extreme values or skewed data distributions. A single distorted measurement can artificially inflate the sample variance. In contrast, the interquartile range has a breakdown point of twenty-five percent, meaning that up to a quarter of the most extreme observations can be altered or corrupted without shifting the interquartile spread. Consequently, statisticians, economists, and data analysts prefer using the median and IQR over the mean and standard deviation when analyzing skewed distributions such as household wealth or national income.

The interquartile range forms the mathematical backbone of the box-and-whisker plot, an exploratory data visualization technique introduced by American statistician John Tukey in 1977. In a box plot, a central rectangle spans from Q1 to Q3, with an internal vertical line marking the median, or Q2. Tukey also devised an objective numerical threshold for detecting outliers, known as Tukey's fences. Any observation falling more than one point five times the IQR below Q1 or above Q3 is classified as an outlier. If a value falls beyond three times the IQR, it is flagged as an extreme outlier, assisting scientists in automated data screening and quality verification.

Key Concepts & Self-Assessment20 Key Facts

Review key What Is the Interquartile Range (IQR)? Box Plots & Outlier Detection exam facts and rate your mastery to track revision.

Progress: 0/20 Rated 0 Mastered 0 Review Later
#1
The interquartile range is a statistical measure of spread defined as the mathematical difference between the third quartile and the first quartile.
#2
The formula for computing the interquartile range is written as IQR equals Q3 minus Q1, representing the middle fifty percent of observations.
#3
Quartiles divide an ordered dataset into four equal segments, with Q1 at the twenty-fifth percentile and Q3 at the seventy-fifth percentile.
#4
The second quartile, or Q2, marks the fiftieth percentile and represents the median of the entire ordered data set.
#5
Because it focuses entirely on the central half of observations, the interquartile range remains resistant to extreme values and severe skewness.
#6
The breakdown point of the interquartile range is twenty-five percent, indicating high resistance to contaminated or anomalous sample values.
#7
American statistician John Tukey popularized the use of quartiles and the interquartile range through the invention of box-and-whisker plots in 1977.
#8
In a standard box plot, the central box represents the interquartile range, while whiskers extend outward to show distribution extremes.
#9
Tukey’s rule identifies an observation as an outlier if it lies more than one point five times the IQR beyond either quartile boundary.
#10
The lower inner fence for outlier detection is computed as Q1 minus one point five times the interquartile range.
#11
The upper inner fence for outlier detection is computed as Q3 plus one point five times the interquartile range.
#12
Values located beyond three times the interquartile range past Q1 or Q3 are formally designated as extreme statistical outliers.
#13
Semi-interquartile range, also called quartile deviation, is calculated by dividing the interquartile range by two.
#14
The coefficient of quartile deviation is a relative dispersion metric calculated as Q3 minus Q1 divided by the sum of Q3 and Q1.
#15
In a perfectly symmetric normal distribution, the interquartile range is approximately equal to one point three four nine multiplied by the standard deviation.
#16
For skewed economic data such as household income or real estate values, the median and IQR provide far more realistic summaries than the mean.
#17
Unlike sample range, which considers only the single maximum and minimum data values, IQR ignores extreme outer tails completely.
#18
Machine learning practitioners utilize IQR thresholds to detect and eliminate sensor errors and data entry anomalies during data preprocessing.
#19
If a dataset has identical values for Q1, Q2, and Q3, the interquartile range equals zero, reflecting concentrated distributions.
#20
In clinical pharmacology and laboratory blood testing, reference ranges for non-normally distributed biomarkers are routinely established using interquartile metrics.

Subject Specialist Commentary

Analytical perspective & practical exam advice from the Master10 academic board

Educator's Insight
Think of a classroom where nine students earn ordinary exam scores and one billionaire's child earns a zero because they skipped the test entirely. While the ordinary range and average score get heavily distorted by that single unusual score, the interquartile range simply measures the middle fifty percent of students, ignoring the extremes at both ends. It gives a genuine picture of typical student performance without being misled by outliers.
In UPSC CSAT, SSC CGL, and State PSC examinations, statistics questions regularly evaluate Tukey's fences and quartile deviations. Keep in mind the standard outlier multiplier: 1.5 times the IQR identifies standard outliers, while 3 times the IQR marks extreme outliers. Watch out for exam traps claiming that standard deviation is preferable for skewed data: whenever data is skewed or contains outliers, always choose median and IQR over mean and standard deviation.

Related Knowledge Topics to Discover

General Science
Standard Error vs Standard Deviation: Population Dispersion vs Sample Mean Variability

Learn the difference between standard deviation and standard error in statistics. Understand sample dispersion (s) vs sampling distribution variability (SE = s / sqrt(n)) for exams.

Explore Topic
Indian Economy
Pareto Efficiency (Pareto Optimality): Vilfredo Pareto, Edgeworth Box & Welfare Theorems

Master Pareto Efficiency in welfare economics: Vilfredo Pareto (1906), Pareto Improvement vs Pareto Optimality, Production Possibility Frontier (PPF), Edgeworth Box contract curve, First/Second Welfare Theorems, and Kaldor-Hicks efficiency.

Explore Topic
General Science
What Is the Pareto Principle? The 80/20 Rule & Power Law Distributions

Learn the Pareto principle and 80/20 rule. Discover power law distributions, mathematical alpha parameters, economics, quality control, and system designs.

Explore Topic

Looking for more GK practice?

Explore 52,789+ questions across 65 General Knowledge categories.

Open Interactive Search