Follow Us
Select Medium / माध्यम चुनें:
Eng (English) Beng (বাংলা) Hindi (हिन्दी)
WBB • Class XI • Economics • Ch 9
Estimated Time: 45 Mins
Study Progress: In Progress

Measures of Dispersion

Measures of Dispersion constitute one of the most critical analytical pillars of quantitative statistical economics, providing a rigorous methodology to evaluate, measure, and interpret the degree of scatter, spread, or variation among individual observations around a central average. Prescribed under the West Bengal Council of Higher Secondary Education (WBCHSE) Class 11 Economics curriculum within the Quantitative Economics and Statistics syllabus, this chapter systematically bridges descriptive univariate summary measures—namely, measures of central tendency—with advanced statistical modeling. While averages such as the arithmetic mean, median, and mode locate the central focal point of a distribution, they remain completely silent regarding the structural dispersion of the data. Two entirely distinct economic distributions can possess identical arithmetic means yet exhibit radically divergent degrees of internal stability: for instance, a company paying each of its five employees ₹20,000 has a mean wage of ₹20,000 with zero variability, whereas another firm paying ₹5,000, ₹10,000, ₹15,000, ₹20,000, and ₹50,000 also has a mean wage of ₹20,000, but exhibits extreme wage inequality and instability. Consequently, dispersion is universally termed 'averages of the second order'. The curriculum provides an exhaustive examination of both absolute measures (expressed in concrete physical measurement units such as rupees, kilograms, or quintals) and relative measures or coefficients (pure, dimensionless ratios or percentages that enable direct comparative analysis across disparate economic series). The syllabus covers positional measures including the Range and Quartile Deviation (Semi-Interquartile Range), mathematical deviation metrics including the Mean Deviation from Mean and Median, the definitive master measure of Standard Deviation (σ) and Variance (σ²) formulated by Karl Pearson, the Coefficient of Variation (CV) for comparative stability testing, and the visual geometric methodology of the Lorenz Curve for measuring national income and wealth inequalities.

Why This Chapter Matters

In empirical macroeconomic policy, corporate finance, industrial quality control, and social welfare analysis, dispersion metrics serve as the indispensable diagnostic for assessing risk, uniformity, and economic inequality. National economic planners cannot judge the true living standards of a population solely on the basis of Per Capita Income (Mean Income): a high national average can conceal devastating poverty if wealth is concentrated in the hands of a fractional elite. Economists employ the Gini Coefficient and Lorenz Curve—direct extensions of dispersion analysis—to quantify socioeconomic disparities, design progressive taxation structures, and monitor targeted poverty alleviation programs. In modern financial economics and investment management, Standard Deviation serves as the universal metric for financial risk and stock market volatility: higher volatility implies greater risk, directly determining portfolio allocation, derivative pricing, and capital adequacy requirements under international banking regulations. In manufacturing and industrial management, dispersion metrics under statistical quality control (SQC) determine whether automated assembly lines operate within permissible engineering tolerances. For higher secondary economics students, mastering the algebraic properties, calculation procedures, and comparative criteria of dispersion establishes the quantitative foundation required for correlation analysis, linear regression, econometric inference, and contemporary data-driven macroeconomic research.

Chapter Roadmap & Progression

1 Conceptual Foundations, Meaning & O...
2 Positional Measures: Range & Quarti...
3 Mean Deviation (Average Deviation)...
4 Karl Pearson's Standard Deviation (...
5 Mathematical Properties of Standard...
6 Coefficient of Variation (CV), Lore...

Complete Concept Guide (100% Curriculum Coverage)

Conceptual Foundations, Meaning & Objectives of Dispersion

1. Meaning & Limitations of Central Tendency Alone

Measures of central tendency (Arithmetic Mean, Median, Mode) are designed to provide a single representative figure that summarizes an entire distribution. However, an average by itself fails to convey the complete picture of a statistical series. It indicates the location of the center, but reveals nothing about how individual observations are scattered or clustered around that center.

Consider three distinct factories employing 5 workers each, with the following daily wage distributions (in ₹):

  • Factory A: 50, 50, 50, 50, 50 $\implies \text{Mean } (\bar{X}) = ₹50$
  • Factory B: 45, 48, 50, 52, 55 $\implies \text{Mean } (\bar{X}) = ₹50$
  • Factory C: 10, 20, 50, 80, 90 $\implies \text{Mean } (\bar{X}) = ₹50$

All three factories share the identical average wage of ₹50. Yet, in Factory A there is zero variation (perfect uniformity); in Factory B there is small variation (high consistency); and in Factory C there is extreme variation (severe wage inequality). Without measuring the spread of the data, the average alone is misleading. This spread or variation is termed Dispersion.

2. Authoritative Definitions of Dispersion
  • Arthur L. Bowley: "Dispersion is the measure of the variations of the items."
  • Brooks & Dick: "Dispersion or spread is the degree of the scatter or variation of the variables about a central value."
  • Simpson & Kafka: "The measurement of the scatterness of the mass of figures in a series about an average is called measure of variation or dispersion."

Because measures of central tendency are called averages of the first order, measures of dispersion—which measure the average deviation from those central values—are logically designated as averages of the second order.

3. Objectives of Measuring Dispersion
  1. Assessing the Reliability of an Average: If dispersion is small, individual values cluster tightly around the average, making it highly representative and reliable. If dispersion is large, the average is unrepresentative.
  2. Comparing Variability, Consistency, or Stability: Comparing two or more distributions (e.g., scoring consistency of two cricket batsmen, or production consistency of two machines).
  3. Controlling Variability: In industrial manufacturing, keeping product dimensions within acceptable tolerance limits through Statistical Quality Control (SQC).
  4. Facilitating Advanced Statistical Analysis: Dispersion serves as the foundational component for computing correlation coefficients, regression equations, testing hypotheses, and constructing economic index numbers.
4. Absolute vs. Relative Measures of Dispersion
Analytical Criterion Absolute Measure of Dispersion Relative Measure of Dispersion (Coefficient)
Definition Measures the actual amount of variation expressed in the physical units of the original data. Measures the variation as a pure, dimensionless ratio or percentage relative to a central average.
Units of Measurement Expressed in concrete units: rupees (₹), kilograms (kg), metres (m), quintals. Pure numbers without any units (e.g., 0.25 or 25%).
Comparability Cannot be used to compare two series having different units (e.g., height in cm vs weight in kg) or vastly different means. Ideal for comparing variability across disparate series regardless of measurement units or scales.
Standard Measures Range, Quartile Deviation, Mean Deviation, Standard Deviation. Coefficient of Range, Coefficient of QD, Coefficient of MD, Coefficient of Variation (CV).
5. Yule's Characteristics of an Ideal Measure of Dispersion

According to statistician G. Udny Yule, a satisfactory measure of dispersion must satisfy six criteria:

  1. It should be rigidly defined by a definitive mathematical formula.
  2. It should be based on all observations in the distribution.
  3. It should be readily comprehensible and simple to interpret.
  4. It should be simple to compute without undue algebraic labor.
  5. It should be capable of further algebraic manipulation (e.g., combined dispersion of multiple groups).
  6. It should possess sampling stability, being least affected by random sampling fluctuations.

Positional Measures: Range & Quartile Deviation (QD)

1. The Range: Definition & Formulation

The Range is the simplest and crudest measure of dispersion. It is defined as the absolute difference between the largest (maximum) value and the smallest (minimum) value in a series.

$$\text{Absolute Range} = L - S$$ $$\text{Coefficient of Range} = \frac{L - S}{L + S}$$

Where $L$ is the largest value and $S$ is the smallest value.

  • Individual & Discrete Series: Identify the maximum and minimum values directly from the observations $X$.
  • Continuous Frequency Distribution: Computed either as:
    • Upper limit of the highest class interval minus Lower limit of the lowest class interval; or
    • Mid-point of the highest class minus Mid-point of the lowest class.

Evaluation:

  • Merits: Exceptionally simple to compute; readily understood; widely used in daily weather reporting (maximum and minimum temperature) and industrial quality control charts (X-bar and R charts).
  • Demerits: Highly unstable because it depends entirely on two extreme outlier values; completely ignores the intermediate 98% of observations; cannot be computed for open-ended frequency distributions.
2. Quartile Deviation (Semi-Interquartile Range)

To eliminate the extreme outlier sensitivity of the Range, statisticians developed positional measures based on partition values called Quartiles.

Quartiles divide an ordered distribution into four equal parts: $Q_1$ (First/Lower Quartile, 25th percentile), $Q_2$ (Median, 50th percentile), and $Q_3$ (Third/Upper Quartile, 75th percentile).

$$\text{Interquartile Range (IQR)} = Q_3 - Q_1$$ $$\text{Quartile Deviation (QD)} = \frac{Q_3 - Q_1}{2}$$ $$\text{Coefficient of Quartile Deviation} = \frac{Q_3 - Q_1}{Q_3 + Q_1}$$

The Interquartile Range represents the absolute span of the central 50% of the distribution, while the Quartile Deviation (QD) is half of this span, hence termed the Semi-Interquartile Range.

3. Calculation of Quartiles Across Series

Step-by-step procedure for Continuous Series:

  1. Construct the cumulative frequency ($c.f.$) column.
  2. Locate the $Q_1$ class corresponding to cumulative frequency containing $\frac{N}{4}$, and the $Q_3$ class containing $\frac{3N}{4}$.
  3. Apply the interpolation formulas: $$Q_1 = L_1 + \frac{\frac{N}{4} - c.f.}{f_1} \times i_1$$ $$Q_3 = L_3 + \frac{\frac{3N}{4} - c.f.}{f_3} \times i_3$$ Where $L$ is the lower boundary of the quartile class, $c.f.$ is the cumulative frequency of the preceding class, $f$ is the simple frequency of the quartile class, and $i$ is the class width.
4. Merits & Limitations of Quartile Deviation
  • Merits:
    • Superior to the Range because it excludes the extreme 25% lower and 25% upper observations, making it completely immune to extreme outliers.
    • Can be easily computed for open-ended distributions (e.g., "Below 10" or "Above 80") because $Q_1$ and $Q_3$ lie safely in the middle ranges.
  • Limitations:
    • Ignores the actual values of 50% of the data (the bottom quarter and top quarter).
    • Not amenable to algebraic manipulation; combined quartile deviations of multiple distributions cannot be calculated.
    • More sensitive to sampling variations than standard deviation.

Mean Deviation (Average Deviation) & The Minimal Property

1. Theoretical Concept & Definition

The Mean Deviation (or Average Deviation) is defined as the arithmetic mean of the numerical deviations of observations from an average (Mean, Median, or Mode), with the critical mathematical stipulation that all algebraic signs (+ and -) are ignored (taking absolute values $|d|$).

If signs were not ignored, the sum of deviations taken from the arithmetic mean would be identically zero: $\sum (X - \bar{X}) = 0$. By taking absolute values $|X - \bar{X}|$, the measure quantifies the true average distance of observations from the center.

2. Computational Formulas Across Series
Series Type Mean Deviation from Mean ($MD_{\bar{X}}$) Mean Deviation from Median ($MD_M$)
Individual Series $$MD_{\bar{X}} = \frac{\sum |X - \bar{X}|}{N}$$ $$MD_M = \frac{\sum |X - M|}{N}$$
Discrete & Continuous Series $$MD_{\bar{X}} = \frac{\sum f |X - \bar{X}|}{N}$$
(where $X$ is the mid-value in continuous series)
$$MD_M = \frac{\sum f |X - M|}{N}$$
3. Coefficients of Mean Deviation

To convert absolute Mean Deviation into relative measures for comparative analysis:

$$\text{Coefficient of } MD_{\bar{X}} = \frac{MD_{\bar{X}}}{\bar{X}} \qquad \text{and} \qquad \text{Coefficient of } MD_M = \frac{MD_M}{M}$$
4. The Minimal Property of Mean Deviation
Crucial Board Theorem: Minimal Property of Median Deviations
The sum of absolute deviations of observations is minimum when deviations are taken from the Median: $$\sum |X - M| < \sum |X - A| \quad \text{for any } A \neq M$$ Consequently, $MD_M \le MD_{\bar{X}}$. Therefore, statistically and theoretically, Mean Deviation is always best calculated from the Median.
5. Merits & Limitations of Mean Deviation
  • Merits:
    • Based on all observations in the dataset.
    • Simple to comprehend and intuitively meaningful as the average distance from the center.
    • Less distorted by extreme outliers compared to the standard deviation.
  • Limitations:
    • Mathematically flawed: Artificially ignoring algebraic signs ($|-d| = +d$) violates basic algebraic principles and makes algebraic manipulation impossible.
    • Incapable of being combined algebraically across multiple groups.
    • Rarely used in advanced inferential statistics and hypothesis testing for this reason.

Karl Pearson's Standard Deviation (σ) & Variance (σ²)

1. Genesis & Mathematical Definition

Introduced by the eminent British statistician Karl Pearson in 1893, the Standard Deviation (denoted by the lowercase Greek letter sigma, $\sigma$) is the master measure of dispersion in statistical science. It completely overcomes the mathematical flaw of Mean Deviation (dropping negative signs) by squaring the deviations before averaging them, and then taking the positive square root.

Formal Definition: Standard Deviation is the positive square root of the arithmetic mean of the squares of deviations of observations taken strictly from their arithmetic mean.

$$\sigma = \sqrt{\frac{\sum (X - \bar{X})^2}{N}} = \sqrt{\frac{\sum x^2}{N}}$$

Because it computes the root of the mean of squared deviations, it is also classically termed the Root-Mean-Square Deviation.

The square of the standard deviation is called the Variance:

$$\text{Variance} = \sigma^2 = \frac{\sum (X - \bar{X})^2}{N}$$
2. Computational Pathways for Frequency Distributions

In WBCHSE examinations, candidates can compute $\sigma$ using three standard techniques:

Computation Method Condition for Use Standard Deviation Formula
Direct Method (Actual Mean $\bar{X}$) When $\bar{X}$ is an exact whole integer without decimals. $$\sigma = \sqrt{\frac{\sum f(X - \bar{X})^2}{N}} = \sqrt{\frac{\sum f x^2}{N}}$$
Short-cut Method (Assumed Mean $A$) When $\bar{X}$ is a fraction/decimal; eliminates decimal squaring. $$\sigma = \sqrt{\frac{\sum f d^2}{N} - \left(\frac{\sum f d}{N}\right)^2} \quad (d = X - A)$$
Step-Deviation Method (Scale Factor $c$) When class intervals have a uniform width $c$ in continuous series. $$\sigma = c \times \sqrt{\frac{\sum f d'^2}{N} - \left(\frac{\sum f d'}{N}\right)^2} \quad \left(d' = \frac{X - A}{c}\right)$$
3. Critical Computational Trap
Frequent Student Error in Short-Cut Formula:
In the formula $\sigma = \sqrt{\frac{\sum f d^2}{N} - \left(\frac{\sum f d}{N}\right)^2}$:
  • The first term $\frac{\sum f d^2}{N}$ requires squaring each deviation $d$, multiplying by frequency $f$, summing them, and dividing by $N$.
  • The second term $\left(\frac{\sum f d}{N}\right)^2$ requires finding the mean deviation $\frac{\sum f d}{N}$ first, and then squaring that entire quotient.
  • Since the second term is subtracted, $\frac{\sum f d^2}{N}$ is always greater than or equal to $\left(\frac{\sum f d}{N}\right)^2$. If you obtain a negative number under the radical, an arithmetic error has occurred!

Mathematical Properties of Standard Deviation & Empirical Rules

1. Cardinal Mathematical Properties of Standard Deviation

Standard Deviation possesses unique algebraic properties that make it the bedrock of inferential statistics:

Property I: Minimum Sum of Squared Deviations

The sum of squares of deviations of observations is minimum when taken from the arithmetic mean:

$$\sum (X - \bar{X})^2 < \sum (X - A)^2 \quad \text{for any } A \neq \bar{X}$$

This guarantees that the standard deviation is the unique minimum root-mean-square deviation of any dataset.

Property II: Effect of Change of Origin (Invariance)

The Standard Deviation is independent of change of origin. If a constant $a$ is added to or subtracted from each observation ($Y = X \pm a$), the standard deviation remains completely unchanged:

$$\sigma(X \pm a) = \sigma(X)$$
Property III: Effect of Change of Scale (Dependence)

The Standard Deviation is dependent on change of scale. If each observation is multiplied or divided by a constant $c$ ($Y = c \cdot X$), the standard deviation is multiplied by the absolute value of that constant:

$$\sigma(c \cdot X) = |c| \cdot \sigma(X)$$
Property IV: Combined Standard Deviation of Two Groups

If two independent samples have sizes $N_1, N_2$, means $\bar{X}_1, \bar{X}_2$, and standard deviations $\sigma_1, \sigma_2$, their combined standard deviation $\sigma_{12}$ is given by:

$$\sigma_{12} = \sqrt{\frac{N_1(\sigma_1^2 + d_1^2) + N_2(\sigma_2^2 + d_2^2)}{N_1 + N_2}}$$

Where:

  • Combined Mean: $\bar{X}_{12} = \frac{N_1\bar{X}_1 + N_2\bar{X}_2}{N_1 + N_2}$
  • Mean deviations: $d_1 = \bar{X}_1 - \bar{X}_{12}$ and $d_2 = \bar{X}_2 - \bar{X}_{12}$
2. Empirical Rules for Symmetrical (Normal) Distributions

For a moderately symmetrical or bell-shaped normal distribution, well-established mathematical relationships exist among the three primary dispersion measures:

$$QD = \frac{2}{3} \sigma = 0.6745 \sigma \qquad \text{and} \qquad MD = \frac{4}{5} \sigma = 0.7979 \sigma$$ $$6 \cdot QD \approx 5 \cdot MD \approx 4 \cdot \sigma$$ $$\text{Ratio: } QD : MD : \sigma = 10 : 12 : 15$$

Area Spread of the Normal Curve:

  • $\bar{X} \pm 1\sigma$ encompasses approximately 68.27% of all observations.
  • $\bar{X} \pm 2\sigma$ encompasses approximately 95.45% of all observations.
  • $\bar{X} \pm 3\sigma$ encompasses approximately 99.73% of all observations (virtually the entire population).
  • $\bar{X} \pm 1 QD$ encompasses exactly 50.00% of all observations.

Coefficient of Variation (CV), Lorenz Curve & Economic Inequality

1. Karl Pearson's Coefficient of Variation (CV)

When comparing the variability, consistency, or stability of two or more distributions with different measurement units or widely differing arithmetic means, absolute standard deviation is inadequate. In 1895, Karl Pearson introduced the Coefficient of Variation (CV) as a standardized percentage metric of relative dispersion:

$$CV = \frac{\sigma}{\bar{X}} \times 100$$
2. Decision Rule for Comparative Consistency & Stability
Interpretation Criteria for Higher Secondary Examinations:
  • Higher CV: Indicates greater variability, lesser consistency, lesser stability, or lesser uniformity.
  • Lower CV: Indicates lesser variability, greater consistency, greater stability, or greater uniformity.
  • Exam Application: If a question asks: "Which factory is more consistent in paying wages?" or "Which batsman is more dependable?", calculate the $CV$ for both. The entity with the LOWER CV is the correct answer!
3. Graphic Method: The Lorenz Curve

Formulated by American economic statistician Max O. Lorenz in 1905, the Lorenz Curve is a cumulative percentage graphical technique designed to measure and display economic inequality in the distribution of national income, wealth, profits, or landholdings.

Construction Principles:

  1. Convert both the frequency (e.g., number of income earners) and the variable (e.g., total income earned) into cumulative percentages running from 0% to 100%.
  2. Take the cumulative percentage of income earners on the horizontal axis ($X$-axis) from 0 to 100%.
  3. Take the cumulative percentage of total income on the vertical axis ($Y$-axis) from 0 to 100%.
  4. Draw the 45° diagonal line connecting $(0, 0)$ to $(100, 100)$. This is the Line of Equal Distribution (or Line of Absolute Equality). On this line, bottom 20% of earners receive 20% of income, bottom 50% receive 50% of income, etc.
  5. Plot the actual cumulative percentage coordinates and connect them with a smooth curve. This is the Lorenz Curve.
4. Economic Interpretation & Gini Coefficient
  • In any real-world unequal society, the Lorenz Curve always sags below the 45° Line of Equal Distribution.
  • The Distance Principle: The greater the distance/curvature of the Lorenz Curve away from the 45° diagonal line, the greater the degree of economic inequality in the distribution of income.
  • If Country A's Lorenz Curve lies closer to the diagonal than Country B's curve, Country A enjoys a more equitable distribution of wealth.
  • Gini Coefficient ($G$): Defined as the ratio of the area between the 45° line and the Lorenz Curve to the total area under the 45° line. $G = 0$ signifies perfect equality, and $G = 1$ signifies absolute inequality (one person owns all national income).

Key Economic Identities, Formulas & Business Principles

Range and Quartile Deviation Formulas
Mean Deviation & Standard Deviation Formulas
Coefficient of Variation & Combined Standard Deviation

Conceptual Solved Examples & Case Studies

Example 1
Step-by-Step Solution:
Part (a): Individual Series
Arranging observations: 110, 120, 150, 170, 180, 200, 250.
Largest value ($L$) = ₹250, Smallest value ($S$) = ₹110.
$$\text{Absolute Range} = L - S = 250 - 110 = ₹140$$ $$\text{Coefficient of Range} = \frac{L - S}{L + S} = \frac{250 - 110}{250 + 110} = \frac{140}{360} = 0.3889 \approx 0.39$$
Part (b): Continuous Frequency Distribution
Upper limit of the highest class ($L$) = 60.
Lower limit of the lowest class ($S$) = 10.
$$\text{Absolute Range} = L - S = 60 - 10 = 50 \text{ marks}$$ $$\text{Coefficient of Range} = \frac{L - S}{L + S} = \frac{60 - 10}{60 + 10} = \frac{50}{70} = 0.7143 \approx 0.71$$
Example 2
Step-by-Step Solution:
Step 1: Construct Cumulative Frequency Table
Wages (₹) Frequency ($f$) Cumulative Frequency ($c.f.$)
30-4066
40-501016
50-601834
60-701044
70-80650
Total $N = 50$ -
Step 2: Calculate First Quartile ($Q_1$)
$$\frac{N}{4} = \frac{50}{4} = 12.5 \text{th item}$$ $12.5$ falls in cumulative frequency 16 $\implies Q_1$ class is 40-50.
Here $L_1 = 40$, $c.f. = 6$, $f_1 = 10$, $i = 10$. $$Q_1 = L_1 + \frac{\frac{N}{4} - c.f.}{f_1} \times i = 40 + \frac{12.5 - 6}{10} \times 10 = 40 + 6.5 = ₹46.50$$
Step 3: Calculate Third Quartile ($Q_3$)
$$\frac{3N}{4} = \frac{3(50)}{4} = 37.5 \text{th item}$$ $37.5$ falls in cumulative frequency 44 $\implies Q_3$ class is 60-70.
Here $L_3 = 60$, $c.f. = 34$, $f_3 = 10$, $i = 10$. $$Q_3 = L_3 + \frac{\frac{3N}{4} - c.f.}{f_3} \times i = 60 + \frac{37.5 - 34}{10} \times 10 = 60 + 3.5 = ₹63.50$$
Step 4: Compute Quartile Deviation & Coefficient
$$\text{Quartile Deviation (QD)} = \frac{Q_3 - Q_1}{2} = \frac{63.50 - 46.50}{2} = \frac{17.00}{2} = ₹8.50$$ $$\text{Coefficient of QD} = \frac{Q_3 - Q_1}{Q_3 + Q_1} = \frac{63.50 - 46.50}{63.50 + 46.50} = \frac{17}{110} = 0.1545 \approx 0.155$$
Example 3
Step-by-Step Solution:
Step 1: Compute Mean ($ar{X}$) and Median ($M$)
Total frequency $N = 1 + 4 + 6 + 4 + 1 = 16$.
$$\sum fX = (2)(1) + (4)(4) + (6)(6) + (8)(4) + (10)(1) = 2 + 16 + 36 + 32 + 10 = 96$$ $$ar{X} = rac{\sum fX}{N} = rac{96}{16} = 6$$ For Median: $ rac{N}{2} = rac{16}{2} = 8 ext{th item}$. Cumulative frequencies are 1, 5, 11, 15, 16. The 8th item falls in $c.f. = 11$, corresponding to $X = 6$. Thus $ ext{Median } M = 6$.
Since the distribution is perfectly symmetrical, $ar{X} = M = 6$.

Step 2: Construct Deviation Table
$X$ $f$ $|X - 6|$ $f|X - 6|$
2144
4428
6600
8428
10144
Total $N = 16$ - $\sum f|X - 6| = 24$
Step 3: Calculate Mean Deviation & Coefficient $$MD_{ar{X}} = MD_M = rac{\sum f|X - 6|}{N} = rac{24}{16} = 1.5$$ $$ ext{Coefficient of Mean Deviation} = rac{MD}{ar{X}} = rac{1.5}{6} = 0.25$$
Example 4
Step-by-Step Solution:
Method 1: Direct Method (from Actual Mean)
Number of items $N = 5$.
$$\sum X = 8 + 12 + 13 + 15 + 22 = 70 \implies ar{X} = rac{70}{5} = 14$$ Deviations from actual mean $x = X - 14$:
  • $X = 8: x = -6 \implies x^2 = 36$
  • $X = 12: x = -2 \implies x^2 = 4$
  • $X = 13: x = -1 \implies x^2 = 1$
  • $X = 15: x = +1 \implies x^2 = 1$
  • $X = 22: x = +8 \implies x^2 = 64$
$$\sum x^2 = 36 + 4 + 1 + 1 + 64 = 106$$ $$\sigma = \sqrt{ rac{\sum x^2}{N}} = \sqrt{ rac{106}{5}} = \sqrt{21.2} pprox 4.604$$ $$ ext{Variance } (\sigma^2) = 21.20$$
Method 2: Short-Cut Method (Assumed Mean $A = 13$)
Deviations $d = X - 13$:
  • $X = 8: d = -5 \implies d^2 = 25$
  • $X = 12: d = -1 \implies d^2 = 1$
  • $X = 13: d = 0 \implies d^2 = 0$
  • $X = 15: d = +2 \implies d^2 = 4$
  • $X = 22: d = +9 \implies d^2 = 81$
$$\sum d = -5 - 1 + 0 + 2 + 9 = +5, \qquad \sum d^2 = 25 + 1 + 0 + 4 + 81 = 111$$ Applying the short-cut formula: $$\sigma = \sqrt{ rac{\sum d^2}{N} - \left( rac{\sum d}{N} ight)^2} = \sqrt{ rac{111}{5} - \left( rac{5}{5} ight)^2} = \sqrt{22.2 - (1)^2} = \sqrt{21.2} pprox 4.604$$ $$ ext{Variance } (\sigma^2) = 21.20$$ Both methods produce identical results, confirming algebraic accuracy.
Example 5
Step-by-Step Solution:
Step 1: Set Class Mid-points and Step-Deviations
Let assumed mean $A = 25$, class interval width $c = 10$.
Step-deviation $d' = rac{m - 25}{10}$.

Step 2: Computation Table
Wages (₹) Mid-value ($m$) $f$ $d' = rac{m-25}{10}$ $f d'$ $d'^2$ $f d'^2$
0-1052-2-448
10-20155-1-515
20-302580000
30-40353+1+313
40-50452+2+448
Total - $N = 20$ - $\sum f d' = -2$ - $\sum f d'^2 = 24$
Step 3: Calculate Arithmetic Mean ($ar{X}$) $$ar{X} = A + \left( rac{\sum f d'}{N} ight) imes c = 25 + \left( rac{-2}{20} ight) imes 10 = 25 - 1.0 = ₹24.00$$
Step 4: Calculate Standard Deviation ($\sigma$) $$\sigma = c imes \sqrt{ rac{\sum f d'^2}{N} - \left( rac{\sum f d'}{N} ight)^2} = 10 imes \sqrt{ rac{24}{20} - \left( rac{-2}{20} ight)^2}$$ $$\sigma = 10 imes \sqrt{1.20 - (-0.1)^2} = 10 imes \sqrt{1.20 - 0.01} = 10 imes \sqrt{1.19} = 10 imes 1.09087 = ₹10.91$$
Step 5: Calculate Coefficient of Variation ($CV$) $$CV = rac{\sigma}{ar{X}} imes 100 = rac{10.91}{24.00} imes 100 = 45.46\%$$
Example 6
Step-by-Step Solution:

Part (a): Comparing Consistency Using Coefficient of Variation
For Plant A:

$$CV_A = rac{\sigma_1}{ar{X}_1} imes 100 = rac{9}{186} imes 100 = 4.839\%$$

For Plant B:

$$CV_B = rac{\sigma_2}{ar{X}_2} imes 100 = rac{10}{175} imes 100 = 5.714\%$$

Decision: Since $CV_A < CV_B$ (4.84% < 5.71%), Plant A has a smaller coefficient of variation, indicating that Plant A has greater uniformity and consistency in paying wages to its employees.

Part (b): Combined Mean and Combined Standard Deviation

  1. Combined Mean ($ar{X}_{12}$):

$$ar{X}_{12} = rac{N_1 ar{X}_1 + N_2 ar{X}_2}{N_1 + N_2} = rac{500(186) + 600(175)}{500 + 600} = rac{93000 + 105000}{1100} = rac{198000}{1100} = ₹180.00$$


2. Deviations of Group Means from Combined Mean:

$$d_1 = ar{X}_1 - ar{X}_{12} = 186 - 180 = +6 \implies d_1^2 = 36$$

$$d_2 = ar{X}_2 - ar{X}_{12} = 175 - 180 = -5 \implies d_2^2 = 25$$


3. Combined Standard Deviation ($\sigma_{12}$):

$$\sigma_{12} = \sqrt{ rac{N_1(\sigma_1^2 + d_1^2) + N_2(\sigma_2^2 + d_2^2)}{N_1 + N_2}}$$

$$\sigma_{12} = \sqrt{ rac{500(9^2 + 36) + 600(10^2 + 25)}{1100}} = \sqrt{ rac{500(81 + 36) + 600(100 + 25)}{1100}}$$

$$\sigma_{12} = \sqrt{ rac{500(117) + 600(125)}{1100}} = \sqrt{ rac{58500 + 75000}{1100}} = \sqrt{ rac{133500}{1100}} = \sqrt{121.3636} pprox ₹11.02$$

The combined mean wage is ₹180.00 and the combined standard deviation is ₹11.02.

Common Misconceptions & Examiner Traps

Common Misconception

Confusing (Σfd / N)² with Σfd² / N in the short-cut standard deviation formula.

Scientific Reality & Correction

Σfd² / N requires squaring deviation d first and multiplying by f, then summing and dividing by N. (Σfd / N)² requires finding the mean deviation first and squaring the final quotient.

Common Misconception

Believing that a higher Coefficient of Variation indicates greater consistency.

Scientific Reality & Correction

Higher CV means GREATER variability and LESSER consistency. The entity with the LOWER CV is always the more consistent and stable one.

Common Misconception

Assuming that adding a constant to all observations changes the standard deviation.

Scientific Reality & Correction

Standard deviation is independent of change of origin: σ(X ± a) = σ(X). Only multiplication or division (change of scale) alters σ.

Visual Learning & Conceptual Map

σ WBCHSE Class 11 Economics • Quantitative Statistics: Measures of Dispersion Range, Quartile Deviation (QD), Mean Deviation (MD), Standard Deviation (σ), CV & Lorenz Curve 1. Concept of Dispersion: Identical Mean, Different Spread Mean x̄ = 50 Series A: Low Dispersion (More Consistent) Series B: High Dispersion (More Variable) • Averages indicate the central value; Dispersion measures the extent of scatter around it. 2. Standard Deviation (σ) & Normal Distribution Karl Pearson's Root-Mean-Square Deviation (σ): σ = √[Σ(X - x̄)² / N] = √[Σd²/N - (Σd/N)²] Empirical Relationship Among Dispersion Measures: 6 QD ≈ 5 MD ≈ 4 σ (QD : MD : σ = 10 : 12 : 15) Empirical Rule in Symmetrical Normal Distributions: Normal Distribution Spread: x̄ ± 1σ = 68.27% | x̄ ± 2σ = 95.45% | x̄ ± 3σ = 99.73% Independent of origin: σ(X ± a) = σ(X) | Scale dependent: σ(c·X) = |c|·σ(X) 3. Coefficient of Variation (CV) & Lorenz Curve Coefficient of Variation (Relative Measure): CV = (σ / x̄) × 100 % • Higher CV ➔ Greater variability, less consistency, lower stability • Lower CV ➔ Lesser variability, higher consistency, greater stability Lorenz Curve (Graphic Measure of Economic Inequality): -- 45° Line of Equal Distribution __ Lorenz Curve (Actual Distribution) Greater the curvature away from diagonal, greater the economic inequality TargetExams Academic Series • WBCHSE Quantitative Statistics • Absolute & Relative Dispersion Analysis

Chapter Summary & 10 Key Takeaways

Takeaway 1
  1. Measures of Dispersion evaluate the degree of scatter, spread, or variation of individual observations around a central average; they are termed 'averages of the second order'.
Takeaway 2
  1. Central tendency locates the center of a distribution, but gives no information on stability or variability. Series with identical means can have vastly different degrees of dispersion.
Takeaway 3
  1. Absolute measures of dispersion are expressed in concrete physical units of data (₹, kg, m); Relative measures (Coefficients) are pure dimensionless percentages used for comparative analysis.
Takeaway 4
  1. Range is the difference between largest and smallest values (L - S). It is simple but excessively sensitive to extreme outliers and unusable for open-ended classes.
Takeaway 5
  1. Quartile Deviation (QD = [Q3 - Q1] / 2) measures the semi-interquartile range of the middle 50% of data. It is robust to extreme values and computable for open-ended distributions.
Takeaway 6
  1. Mean Deviation is the arithmetic mean of absolute deviations from an average (|d|). It is minimized when deviations are taken from the Median: Σ|X - M| is minimum.
Takeaway 7
  1. Standard Deviation (σ), formulated by Karl Pearson, is the root-mean-square deviation from the arithmetic mean: σ = √[Σ(X - X̄)² / N]. It is the most reliable, master measure of dispersion.
Takeaway 8
  1. Standard deviation is independent of change of origin [σ(X ± a) = σ(X)], but dependent on change of scale [σ(c · X) = |c| · σ(X)].
Takeaway 9
  1. Karl Pearson's Coefficient of Variation (CV = σ / X̄ × 100) evaluates consistency: lower CV indicates greater consistency, stability, and uniformity.
Takeaway 10
  1. The Lorenz Curve is a graphic method to depict economic inequality. The greater the curvature away from the 45° Line of Equal Distribution, the greater the income or wealth disparity.

Check Your Understanding (Diagnostic Practice Questions)

Diagnostic questions testing core conceptual clarity. Answers are hidden initially — solve each problem first, then click to reveal the step-by-step verified solution.

1
Reveal Answer & Explanation
Answer:
2
Reveal Answer & Explanation
Answer:
3
Reveal Answer & Explanation
Answer:
4
Reveal Answer & Explanation
Answer:
5
Reveal Answer & Explanation
Answer:
Finished Studying This Chapter?
READY TO PRACTICE?

Timed CBT Practice Tests (Exam Simulator)

Put your concepts to the test with official curriculum-aligned Foundation and Advanced practice tests. Get instant accuracy scores, time metrics, and step-by-step verified explanations.