Economic Monetary Aggregates: An Application of Index Number and Aggregation Theory
📄 Summarized from the full manuscript · Human-reviewed for faithfulness before publication
In brief
Is it legitimate to add cash and savings deposits together and call the total the money supply? This 1980 paper argues that the conventional simple sum is, in the author's words, severely defective, and derives a weighted index as the theoretically correct alternative. On United States data from 1970 to 1978, different kinds of passbook savings substitute closely for one another, but transaction balances and savings balances substitute only weakly, so summing them is inappropriate. The simple sum then misreads the 1970s shift toward higher-yielding accounts as a falling quantity of money. Why it matters: how money is measured changes what the data seem to say.
What this paper finds — and why it matters
This 1980 Journal of Econometrics paper by William A. Barnett lays out the theoretical foundation for treating monetary aggregates such as M2 as genuine economic quantities and derives the Divisia index as the theoretically correct way to measure them, arguing that the simple-sum index conventionally used by central banks “is severely defective.” Barnett shows that a monetary aggregate is meaningful as an economic variable only if the representative consumer’s utility function is blockwise weakly separable in the relevant monetary assets – utility can be written as a function of “other goods” and a subutility function over the assets to be aggregated – and that, given this separability plus linear homogeneity of the subutility function, Green’s (1964) theorem permits two-stage (or n-stage) budgeting in which total monetary expenditure is allocated first across blocks, such as transaction balances versus a passbook-savings aggregate, and then within each block across its component assets. Extending his own 1978 user-cost formula to incorporate taxation, Barnett prices each asset’s foregone return relative to a benchmark bond yield and estimates nested CES demand systems by full-information maximum likelihood on quarterly U.S. data, 1970Q1-1978Q1: across three types of passbook-savings institutions (commercial bank, savings and loan, mutual savings bank) the estimated elasticity of substitution is high, sigma = 2.66 (beta-hat = 0.62), so high that a near-linear approximation “may be a reasonable approximation” at that level, though unequal estimated weights (alpha-hat = 0.55, 0.26, 0.20) still statistically reject simple summation; but between transaction balances and the passbook aggregate substitutability is much lower, sigma = 0.28 (beta-hat = -2.53), which Barnett calls “too low to justify a linear approximation,” so simple-sum M2 is inappropriate at that level of aggregation. Barnett then shows that the Diewert-superlative Tornquist-Theil Divisia index and the Fisher Ideal index – both exact or near-exact for a broad class of well-behaved aggregator functions – agree to three decimal places when computed for M3 over 1968Q1-1978Q1, while simple-sum M3 velocity over the same period declines secularly (from 1.0 in 1968Q1 to about 0.906 in 1978Q1) even as Divisia M3 velocity stays comparatively stable (roughly 1.00-1.07); Barnett interprets this as evidence that simple-sum aggregation misreads 1970s substitution toward higher-yielding, less-regulated monetary assets as a decline in the “quantity of money,” when in fact the Divisia-measured aggregate is roughly unchanged.
Summary of a classic paper, AI-assisted and human-reviewed. See the linked original for the authoritative claims and full conditions.
Questions & answers
Q1. What is the paper’s central claim, and why does Barnett describe the conventional simple-sum monetary aggregates as “severely defective”?
Barnett argues that simple-sum aggregates like M2 – which add up the dollar balances of different monetary assets with equal weight – are “badly designed as an economic monetary quantity index” and that “the use of simple sum monetary quantity aggregates as economic indices of the quantity of monetary services should be discontinued” (Section 8, p. 34). His diagnosis is that simple summation implicitly assumes every component asset is a perfect substitute for every other in a fixed one-to-one ratio; it therefore cannot distinguish a genuine change in the quantity of monetary services from a reallocation among assets driven by relative yield changes.
Q2. Under what condition is a monetary aggregate a meaningful economic variable in the first place?
A monetary aggregate is meaningful only if the representative consumer’s utility function is blockwise weakly separable in the assets being aggregated – formally, overall utility over the monetary components and other goods can be written as u(f(monetary components), other goods) for some subutility function f (Section 2, pp. 12-13). Barnett states the logic directly: “If the concept of money has meaning, then it follows that an aggregate of monetary assets must exist which is treated by the economy as if it were a single good, which we thereby can call ‘money.’” Without this separability, he argues, any aggregate “is inherently arbitrary and spurious and does not define an economic variable” (p. 13).
Q3. Where does the simple-sum index fit into this framework, and what does it assume about substitutability?
The simple-sum index is the degenerate limiting special case of the framework in which components are perfect substitutes in identical ratios – indifference curves that are straight lines at 45 degrees – with a Leontief (minimum-cost) dual price index. This requires what Barnett calls the “Linearity Conditions”: infinite elasticities of substitution between every pair of components within the aggregate (Sections 6.2 and 8, pp. 28, 34). Because real monetary assets pay different yields and are held for different purposes, this is a strong and, Barnett argues, generally false assumption.
Q4. How does multi-stage budgeting let the paper decompose the aggregation problem, and what nesting does it use for U.S. monetary assets?
Given blockwise weak separability and linear homogeneity of the subutility functions, Green’s (1964) Theorem 4 lets the consumer’s decision decentralize into nested two-stage (or n-stage) budgeting: a first stage allocates total monetary expenditure between transaction balances and a passbook-savings aggregate, and a second stage allocates passbook expenditure across the three institution types offering passbook accounts (Sections 3.4 and 4, pp. 21-24). This nested structure is what lets Barnett estimate separate CES demand systems at each level rather than one large system over every asset at once.
Q5. How is the “price” of holding a monetary asset defined, and how does it enter the aggregation?
Extending his own 1978 formula to incorporate taxation, Barnett defines the current-period user cost of asset i as pi_it = p_t(R_t - r_it)(1 - tau_t) / [1 + R_t(1 - tau_t)]*, where R_t is the benchmark bond yield, r_it is the asset’s own nominal yield, tau_t is the marginal tax rate, and p*_t is the aggregate price index (Eq. 3.5, Section 3.2, pp. 18-20). The real user cost, pi*_it = (R_t - r_it)(1 - tau_t)/[1 + R_t(1 - tau_t)], is independent of p*_t. An asset becomes a “free good” (zero user cost) when its own yield equals the benchmark rate, and user costs of monetary assets rise relative to durables as expected inflation rises (pp. 19-20, fn. 18). This user cost, not the dollar balance alone, is the expenditure-share weight later used in the Divisia index.
Q6. What do the estimated CES demand systems show about substitutability among passbook accounts at different institutions?
FIML estimates of the CES system over passbook accounts at three institution types (Table 1) give beta-hat = 0.62, implying an elasticity of substitution of sigma = 1/(1-beta) = 2.66 – high enough that a near-linear (Laspeyres-type) quantity index “may be a reasonable approximation to the theoretical quantity index” at this level. Even so, simple summation is still statistically rejected because the estimated intensity weights, alpha-hat = (0.55, 0.26, 0.20), are unequal (tail area for equal weights < 0.00001); commercial-bank passbook accounts contribute the most to the aggregate (Section 6.3, Table 1, pp. 28-30).
Q7. What do the estimates show at the next level up – between transaction balances and passbook savings – and what does that imply for M2?
Between transaction balances and the passbook-savings aggregate, substitutability is far lower: beta-hat = -2.53, giving an elasticity of substitution of only sigma = 1/(1-(-2.53)) = 0.28, which Barnett calls “far lower than between passbook accounts at different institution types” and “too low to justify a linear approximation.” The corresponding weights are (alpha-hat_1, alpha-hat_2) = (0.77, 0.23) with an estimated first-order autocorrelation coefficient rho-hat = 0.96 (Section 7.2, Table 2, pp. 32-33). Because simple summation requires near-infinite substitutability, this result implies the conventional simple-sum M2 index is inappropriate at the transaction-passbook level even though a near-linear approximation might be tolerable within the passbook block alone.
Q8. What is the Törnquist-Theil Divisia index, and what empirical evidence does Barnett offer for preferring it (and how does it compare to the Fisher Ideal index)?
The Törnquist-Theil Divisia index defines the aggregate’s growth rate as a share-weighted average of the growth rates of its components, log Q_t - log Q_{t-1} = sum_i sit (log m_it - log m{i,t-1}), where s_it is the average of current and lagged expenditure shares computed from user costs (Eq. 10.1, p. 38).** Following Hulten (1973), the continuous-time Divisia index is exact for any consistent aggregator function; Diewert (1976) shows both the Törnquist-Theil Divisia and the Fisher Ideal index are “superlative,” meaning they are exact for aggregator functions that can second-order-approximate any linearly homogeneous function. Computed for the M3 aggregate over 1968Q1-1978Q1, the two superlative indices’ implied velocities are identical to three decimal places, so “the choice between those two indices is of no importance” (Section 10.2, Table 3, p. 39); Barnett favors the Törnquist-Theil Divisia specifically because it admits a natural interpretation as a weighted average of component growth rates and because it is the natural discrete-time approximation to the continuous-time Divisia line integral (pp. 38-39).
Q9. What does the velocity evidence show, and what mechanism does Barnett propose to explain why simple-sum and Divisia velocity diverge?
Simple-sum M3 velocity declines secularly from 1.0 in 1968Q1 to about 0.906 in 1978Q1, a range of 0.201, while Divisia (and Fisher Ideal) M3 velocity is comparatively stable, ranging roughly 1.00-1.07 – a range of only 0.089, less than half as wide – and extending the aggregate to M3+ (adding repurchase agreements, dealer commercial paper, bankers’ acceptances, and short-dated negotiable Treasury securities) further stabilizes Divisia velocity while simple-sum velocity keeps trending the “wrong” way (Section 10.2, Fig. 1, Table 3, pp. 39-43). Barnett’s proposed mechanism: during the 1970s, as market rates on unregulated assets rose relative to Regulation Q-constrained deposits, consumers substituted toward higher-yielding, less-regulated assets. Simple summation reads this substitution as a decline in the “quantity of money” (falling velocity), but because the Divisia index weights each component by its user-cost expenditure share, it registers the same substitution as an internal reallocation that leaves the true monetary aggregate roughly unchanged – so its velocity stays stable (pp. 12, 41-43).
Key terms in this paper
Definitions below follow the paper's own usage.
- Blockwise weak separability
- the condition, central to this paper's definition of a "meaningful" monetary aggregate, that the representative consumer's utility function over all goods can be written as u(f(monetary components), other goods) for some subutility function f -- i.e., the monetary components can be aggregated into a single argument of utility without loss of information about the consumer's preferences over everything else. Barnett treats this as necessary and sufficient for a set of monetary assets to be aggregable into a single economic "money" variable at all (Section 2, pp. 12-13).
- User cost of a monetary asset
- in this paper, the opportunity cost of holding one dollar of asset i for one period rather than the benchmark bond, defined (with taxation) as pi_it = p*_t(R_t - r_it)(1-tau_t)/[1+R_t(1-tau_t)], where R_t is the bond yield and r_it is the asset's own yield (Eq. 3.5, Section 3.2). It is the "price" that enters both the CES demand systems and the expenditure-share weights of the Divisia index, and it falls to zero when an asset's own yield equals the benchmark rate.
- Simple-sum index (as a degenerate case)
- in Barnett's framework, not an alternative aggregation method but the specific limiting case of the general theory in which all components are perfect substitutes in fixed ratios (linear indifference curves at 45 degrees), requiring infinite elasticities of substitution among all components -- the "Linearity Conditions." Barnett treats this as a testable (and, in his data, rejected or nearly-rejected) special case rather than a free-standing definition of money (Sections 6.2, 8, pp. 28, 34).
- Diewert-superlative index number
- an index number is superlative if it is exact for an aggregator function flexible enough to provide a second-order approximation to any linearly homogeneous function (Diewert 1976). Barnett uses this concept to justify treating both the Törnquist-Theil Divisia and the Fisher Ideal index as theoretically preferred alternatives to simple summation, and as the basis for his empirical finding that the two agree almost exactly in this paper's data (Section 10, pp. 37-39).
- Törnquist-Theil Divisia index
- the specific superlative index Barnett advocates, defined so that the aggregate's log growth rate equals a weighted average of component log growth rates, with weights equal to the average of current and lagged user-cost expenditure shares (Eq. 10.1, p. 38). Barnett favors it over the (numerically near-identical) Fisher Ideal index because of its natural interpretation as a share-weighted average of growth rates and because it is the discrete-time approximation to the continuous-time Divisia line integral shown exact by Hulten (1973).