EconBase
← Back to paper

Capital and Labor Income Pareto Exponents across Time and Space

Extracted main text — title through conclusion, appendix excluded. This is what our citation measures are computed over, published so the extraction can be checked by eye.

67,106 characters · 12 sections · 70 citation commands

Rendered from LaTeX for readability, not typeset faithfully. Citation keys are highlighted; maths is left as source; figures, tables and equation environments are summarised rather than reproduced; unrecognised commands are greyed out so nothing is silently dropped. Email addresses are removed.

Capital and Labor Income Pareto Exponents across Time and Space

abstractWe estimate capital and labor income Pareto exponents across 475\xspace country-year observations that span 52\xspace countries over half a century (1967--2018). We document two stylized facts: (i) capital income is more unequally distributed than labor income in the tail; namely, the capital exponent (1--3, median 1.46\xspace) is smaller than labor (2--5, median 3.35\xspace), and (ii) capital and labor exponents are nearly uncorrelated. To explain these findings, we build an incomplete market model with job ladders and capital income risk that gives rise to a capital income Pareto exponent smaller than but nearly unrelated to the labor exponent. Our results suggest the importance of distinguishing income and wealth inequality. {\bf Keywords:} income fluctuation problem, inequality, power law. {\bf JEL codes:} C46, D15, D31, D52.

Introduction

The purpose of this paper is to estimate and document the Pareto exponents for capital and labor income separately for as many countries and years as possible. We say that a positive random variable $X$ obeys a power law with Pareto exponent $\alpha>0$ if the tail probability decays like a power function: $\operatorname{P}(X>x)\sim x^{-\alpha}$ for large $x$.\footnote{More precisely, we say that a random variable $X$ has a Pareto upper tail with exponent $\alpha>0$ if $\operatorname{P}(X>x)=x^{-\alpha}\ell(x)$ for some slowly varying function $\ell$. A function $\ell:(0,\infty)\to \R$ is said to be slowly varying (at infinity) if it is nonzero for sufficiently large $x$ and $\lim_{x\to\infty}\ell(tx)/\ell(x)=1$ for each $t>0$. See BinghamGoldieTeugels1987 for a comprehensive treatment of the theory of regular variation.} In the context of the income distribution, the Pareto exponent characterizes the tail heaviness of high incomes and hence top tail inequality. We remain agnostic about the shape of the income distribution away from the tail. Our study is motivated by the following two observations. First, we are not aware of a comprehensive study that documents the capital and labor income Pareto exponents separately for many countries and years, despite their importance. Second, the Pareto exponent has desirable properties relative to other popular inequality measures such as the Gini coefficient or top income shares.

Consider the first point. Conceptually, capital and labor income are different entities. While the former is the return for providing capital (wealth), the latter is the return for providing labor services, and there is no particular reason to expect a relation between the two. Although these two forms of income are conceptually distinct, it is often put together as just “income” and discussed in the context of inequality and related policies. If capital and labor income are quantitatively different, a policy design based on total income may be misleading. To give one example, consider the theory of optimal taxation Saez2001, where the income Pareto exponent plays an important role. SaezStancheva2018 carefully distinguish capital and labor income and apply the theory of optimal taxation in the United States. They find that with an income elasticity of $e=0.5$, the optimal top marginal tax rate is about 50% for labor and 60% for capital (see their Figure 5). This difference directly comes from the fact that capital and labor income Pareto exponents are distinct. Thus, distinguishing capital and labor income inequality is potentially important for policy designs.

Consider the second point about the desirable properties of Pareto exponents. In the applied literature such as Piketty2003 and piketty-saez2003, top income shares (such as the top 1% income share) are more commonly reported than the income Pareto exponent, perhaps because top shares are summary statistics that can be computed without specifying functional forms or can be understood by non-experts without special knowledge of statistics. However, atkinson2005comparing documents methodological problems regarding the cross-country comparison of top income shares, citing the differences in tax units (\eg, individual or household) and legislation (\eg, whether social security benefits are taxable). One of the reasons such issues arise is because it is not always clear how to define the population and measure small units.\footnote{Imagine how to formally distinguish cities, towns, villages, and settlements; continents, islands, and islets; and inland seas, lakes, and ponds. How to define units and how to measure small units matter for the size distribution of population, land mass, and water surface area.} Because common inequality measures such as the Gini coefficient and top income shares require the knowledge of the entire distribution, these quantities are greatly affected by the definition and measurement of small units. Using the Pareto exponents significantly alleviates these definition and measurement issues because the Pareto distribution is scale invariant (see JessenMikosch2006 for a summary) and its exponent depends only on the tail behavior, not the entire distribution. For example, doubling the income of all households in the top 1% of the income distribution makes the top 1% income share (roughly) twice as large, but the Pareto exponent is unaffected. A similar comment applies to any inequality measure that depends on the entire distribution, such as the Gini coefficient. While we do not claim that the Pareto exponent is the only interesting inequality measure, it is certainly a robust (detail-independent) measure for top tail inequality. See gabaix2009,Gabaix2016JEP for more discussion on the robustness of the Pareto distribution.

In this paper, we use the harmonized Luxembourg Income Study database (hereafter \citetalias{LIS2021}) to document the capital and labor income Pareto exponents across all available 475\xspace country-year observations that span 52\xspace countries over half a century. We document two empirical findings. First, we find that the capital income Pareto exponent is roughly in the range 1--3 (with median 1.46\xspace) and is smaller than the labor income Pareto exponent, which ranges between 2--5 (with median 3.35\xspace). This implies that capital income is more unequally distributed than labor income. This fact is unsurprising and well known for a specific country or year (see, for example, the Lorenz curve in Figure 1 of SaezStancheva2018). However, we are not aware of a comprehensive study that systematically analyzes datasets from many countries and years, and therefore our finding suggests that capital income is generally more unequal than labor income. More specifically, using a statistical test recently developed by hoga2018detecting, we formally test the equality of capital and labor income Pareto exponents and the null is rejected in 86%\xspace of samples. In every single case of rejection, the capital exponent is smaller than the labor exponent. Second, we find that the capital income Pareto exponent is nearly unrelated to the labor exponent. In particular, the correlation between the two exponents across countries is close to zero.

To explain our empirical findings, we build a simple incomplete market model with job ladders and capital income risk. In the model, agents get randomly promoted to the next job ladder. Because individual income follows a random multiplicative process, we obtain a Pareto-tailed labor income distribution. The agents also save assets and face idiosyncratic investment risk, which generates a Pareto-tailed wealth (hence capital income) distribution. Because the capital income Pareto exponent is mainly determined by the asset return distribution, while the labor income Pareto exponent is mainly determined by the income growth distribution, the relation between the two is weak. Furthermore, we analytically characterize the capital and labor income Pareto exponents and show that the former tends to be smaller than the latter for common parametrization. Our results suggest the importance of distinguishing income and wealth inequality.

\paragraph{Related literature}

The power law behavior of income was first recognized by Pareto1895LaLegge,Pareto1896LaCourbe,Pareto1897Cours, who used tabulation data of tax returns in many European countries. More recent research that employs micro data include reed2001 for U.S., reed2003 for U.S., Canada, Sri Lanka, and Bohemia, nirei-souma2007 for Japan, Toda2011PRE,Toda2012JEBO for U.S., Jenkins2017 for U.K., and IbragimovIbragimov2018 for Russia. These papers all concern specific countries and years. BandourianMcDonaldTurley2003 estimate eleven parametric distributions (some of which exhibit Pareto tails) using 82 household labor income datasets from Luxembourg Income Study \citepalias{LIS2021} as we do, though they neither focus on the Pareto exponent nor consider capital income. gabaix2009 mentions “[the] tail exponent of income seems to vary between 1.5 and 3”, citing AtkinsonPiketty2007, though without providing specific details. AtkinsonPiketty2010 document income Pareto exponents across many countries and years estimated from top income share data based on tax returns. However, these estimates are computed from total income, and since (as we document in Section (ref)) the capital income Pareto exponents tend to be smaller than labor exponents, their estimates are best understood as capital income (hence wealth) Pareto exponents. BenhabibBisinLuo2017 make the point that wealth is more skewed than income, citing a few Pareto exponent estimates from BadelDalyHuggettNybom2018 for income and Vermeulen2018 for wealth. Using the 2013-2014 individual income tax data from Romania, Oancea_2018 document that the capital income Pareto exponent (1.44) is smaller than the labor exponent (2.53). Using time series regressions for 21 countries, Bengtsson_2018 document a positive correlation between the gross capital share in national accounts and the top 1% income shares. Their finding can be explained if top income earners tend to have higher capital income share, which is the case if the capital income Pareto exponent is smaller than the labor exponent as we document in this paper. As mentioned in the introduction, there seems to be no comprehensive studies that document the capital and labor income Pareto exponents separately for many countries and years.

Data

In this section we describe the dataset that we use and discuss its limitations.

The LIS database

We use the data from the Luxembourg Income Study \citepalias{LIS2021}, which is a large, harmonized database of micro-level income data that covers over 50 countries worldwide and many years since the late 1960s. In many countries, the data derive from government surveys (for example, the U.S.\ data is based on the Current Population Survey). The \citetalias{LIS2021} data are available at both individual and household level. We focus on the household labor and capital income because

enumerate*• it is reasonable to assume that economic decisions such as financial planning are made at the household level, and • incomes among couples are likely correlated due to assortative matching in the marriage market Siow_2015, which invalidates statistical estimation.\footnote{In our data, we find an average correlation of 0.22 between labor income of husband and wife, which underpins the conjectured dependency.}

The \citetalias{LIS2021} defines labor income as “cash payments and value of goods and services received from dependent employment, as well as profits/losses and value of goods from self-employment, including own consumption”. Capital income is defined as “cash payments from property and capital (including financial and non-financial assets), including interest and dividends, rental income and royalties, and other capital income from investment in self-employment activity”. Together these two categories make up total factor income. See the \citetalias{LIS2021} 2019 USER GUIDE\footnote{https://www.lisdatacenter.org/wp-content/uploads/files/data-lis-guide.pdf} for a detailed summary on how these data are retrieved and calculated.

Data limitations

Our analysis draws upon datasets from many different countries that are harmonized into a common framework by the \citetalias{LIS2021}. However, many details about the collection of data in the different countries are omitted. For example, we find evidence of top-coding in some countries and years, as the largest income order statistic is equal to the second largest.\footnote{Among all 475\xspace country-year observations, the first and second order statistics are equal in 6 cases for labor income and 11 cases for capital income. Therefore we conjecture that the top-coding issue is not severe.} Top-coding induces an upward bias in the estimation of the Pareto exponent. This issue is not necessarily resolved if, instead, one relies on administrative tax income data, for similar biases arise such as rich households trying to understate their taxable income atkinson2011top. burkhauser2012recent detail a method that can be used to overcome the bias due to top-coding, however at the end of their paper they show that the results are robust even if estimates are based on the top-coded series. For these reasons we treat the datasets as not being top-coded in our analysis.

Another limitation of the \citetalias{LIS2021} database is that it is based on government surveys and the measurement error may be larger compared to administrative data based on tax returns. The fact that the income distribution in administrative data is often reported as tabulations, not micro data, causes no problem for estimating Pareto exponents, as TodaWang2021JAE provide an efficient estimation method for such data. In fact, AtkinsonPiketty2010 document income Pareto exponents across countries and years estimated from top income share data. However, their table is based on total income, and since (as we document in Section (ref)) the capital income Pareto exponents tend to be smaller than labor exponents, the estimates in AtkinsonPiketty2010 are best understood as capital income (hence wealth) Pareto exponents. Since we are not aware of a comprehensive income database that distinguishes capital and labor income, we decided to use the \citetalias{LIS2021} database.

Pareto exponents across countries and years

In this section we estimate the capital and labor income Pareto exponents for all countries and years that are available in the \citetalias{LIS2021} database, which spans 52\xspace countries over half a century (475\xspace country-year observations in total). We then formally test for the equality of the Pareto exponents of capital and labor income.

Estimation method

For each country and year, we suppose that the (capital or labor) income observations $\set{X_n}_{n=1}^N$ are independent and identically distributed (\iid) with cumulative distribution function (CDF) $F(x)=\operatorname{P}(X_n \le x)$. The assumption that the upper tail of income obeys a power law with Pareto exponent $\alpha>0$ translates into the regular variation condition

equation[equation omitted — 55 chars of source]

for some slowly varying function $\ell$ (see Footnote (ref)). Note that the assumption on $\ell$ involves only the limit as $x\to\infty$; we are assuming a power law behavior in the upper tail without taking a stance on the shape of the entire distribution.

We are interested in estimating the Pareto exponent $\alpha$ for each country and year. For this purpose, we employ the hill1975simple (maximum likelihood) estimator

equation[equation omitted — 139 chars of source]

Here $X_{(n)}$ denotes the $n$-th largest order statistic from the sample $\set{X_n}_{n=1}^N$ and $k\in \set{1,\dots,N}$ denotes the number of tail observations used to estimate the Pareto exponent. Hall1982 shows that the standard error of $\widehat{\alpha}(k)$ is $\widehat{\alpha}(k)/\sqrt{k}$ under flexible assumptions on the CDF (see (ref) below). We use the Hill estimator because we are interested in formally testing the equality of the capital and labor income Pareto exponents using the method of hoga2018detecting, for which the Hill estimator is required.\footnote{If we are only interested in estimating the Pareto exponents, then there are many alternative methods available. GomesGuillou2015 review 13 commonly used estimators. Fedotenkov2020 reviews more than 100.}

When the population distribution is known to be exactly Pareto (so $\ell$ in (ref) is zero below the minimum size $x_{\min}$ and constant above this threshold), it is well known that the Hill estimator for the full sample ($k=N$) is consistent, asymptotically normal, and asymptotically efficient because it is a maximum likelihood estimator. In practice, the CDF is not exactly Pareto and the researcher needs to select an appropriate value of $k$. For instance, if $F(x)$ satisfies

equation[equation omitted — 86 chars of source]

with some $\beta>0$, then Hall1982 shows that choosing $k = o(N^{2\beta/(2\beta + \alpha)})$ together with $k \to \infty$ as $N \to \infty$ is sufficient for consistency and asymptotic normality (see also embrechts2013modelling).\footnote{Recall that we write $f(x)=o(g(x))$ if for all $\epsilon>0$, we have $\abs{f(x)}\le \epsilon\abs{g(x)}$ for large enough $x$.} Notice that this choice puts a bound on the growth rate of $k$. The expansion (ref) covers a wide range of distributions of interest, such as the $t$-distribution and the type II extreme value distribution danielsson1997tail.

Despite these asymptotic results, it is notoriously difficult to pick $k$ optimally in finite samples Hall1990,ResnickStarica1997,Danielsson_2001. In practice, researchers often plot the Hill estimator (ref) over a range of $k$ to find a flat region or plot the log rank $\log 1,\dots,\log N$ against the log size $\log X_{(1)},\dots,\log X_{(N)}$ to find a region that exhibits a straight line pattern and choose a size threshold to run the log-rank regression.\footnote{gabaix-ibragimov2011 study the asymptotic behavior of log rank regression and show that the standard error is larger by a factor of $\sqrt{2}$ than the Hill estimator. However, they do not discuss how to select the threshold. In their empirical application, they consider the size distribution of population in U.S.\ metropolitan statistical areas, which are already far into the tail and hence the threshold selection is less of an issue. IbragimovIbragimov2018 apply the same methodology to Russian household income data and consider the top 5% and 10% thresholds.} Unfortunately, this graphical approach is not feasible in our setting because \citetalias{LIS2021} does not allow researchers to download the micro data for confidentiality concerns (researchers are required to submit their execution files to conduct statistical analyses) and there is little scope for exploratory graphical data analysis. In this paper we simply use the largest 5% observations, so

equation[equation omitted — 49 chars of source]

which is standard in the literature.\footnote{An alternative approach is to estimate a parametric distribution $F$ that admits a Pareto upper tail by maximum likelihood using the entire sample. The double Pareto-lognormal distribution proposed by reed2003 and reed-jorgensen2004 often performs best. See Toda2012JEBO for a horse race across several parametric distributions in the context of U.S.\ labor income.} Unreported simulations show that our results are robust to using other thresholds such as the largest 1% and 10% observations, or using a data-driven procedure using the Kolmogorov-Smirnov distance as in danielsson2016tail.

Capital and labor income Pareto exponents

We estimate the capital and labor income Pareto exponents for all countries and years available in the \citetalias{LIS2021} database. The database spans 52\xspace countries across the years 1967--2018, with a total of 475\xspace country-year observations. The point estimates of the capital and labor income Pareto exponents for each country and year as well as their standard errors can be found in Table (ref) in Appendix (ref). To avoid small sample issues, we restrict our analysis to countries with at least 1,000\xspace positive observations for income, resulting in (all) 475\xspace country-year pairs for labor income and 342\xspace for capital income. For visibility, Figure (ref) shows the histogram and scatter plot of the capital and labor income Pareto exponents.

figure[figure omitted — 433 chars of source]

Figure (ref) shows the histogram of the capital and labor income Pareto exponents pooled across all available countries and years. The capital and labor income Pareto exponents are generally in the range 1--3 and 2--5 with medians 1.46\xspace and 3.35\xspace, respectively. This suggests that

enumerate*• capital income is generally more unequally distributed than labor income, but • there is significant heterogeneity in both capital and labor income inequality across countries and years.

Figure (ref) shows the scatter plot of the Pareto exponents together with the 45 degree line. The confidence interval of the correlation coefficient $\rho$ is computed assuming all country-year observations are independent and there is no sampling error in the point estimates of the Pareto exponents. Although it is not obvious how to account for these issues, doing so will only widen the confidence interval. Therefore the fact that the na\"{\i}ve confidence interval almost contains zero suggests that the correlation between the two Pareto exponents is weak. Furthermore, for the vast majority of countries and years, the capital income Pareto exponent is smaller than the labor exponent, again suggesting that capital income is more unequal than labor income.

How do the Pareto exponents evolve over time? Because many countries appear only sporadically in the \citetalias{LIS2021} database, we only consider six countries for which the cross-sectional sample size is large and the time series is long, namely Canada, Germany, Switzerland, Taiwan, United Kingdom, and United States. Figure (ref) plots the capital and labor Pareto exponents and their 95% confidence intervals for these countries. The capital exponent tends to be stable at around 1--2 and is smaller than the labor exponent.\footnote{An exception may be Germany before 1983. However, this may be an artifact of the sudden change in sample size, which was over 40,000 until 1983 and about 5,000 since 1984. Thus the 5% rule for capital income may be including too many observations in the body of the distribution before 1983.} Again, capital income appears to be more unequally distributed than the labor income.

figure[figure omitted — 489 chars of source]

Testing equality of capital and labor Pareto exponents

We now formally test whether the capital and labor Pareto exponents are equal. In particular our test is

equation*[equation* omitted — 132 chars of source]

where $\alpha_\text{cap},\alpha_\text{lab}$ denote the capital and labor income Pareto exponents. Testing the null hypothesis $H_0$ is complicated by the fact that there is dependency between labor income $\set{X_{\text{lab},n}}_{n=1}^N$ and capital income $\set{X_{\text{cap},n}}_{n=1}^N$, because individuals who are rich (receive high labor income) tend to be wealthy and receive high capital income. Thus we cannot use the 95% confidence intervals in Figure (ref) to test the equality of Pareto exponents. Instead, we apply the test recently developed by hoga2018detecting, which allows for dependence in the data but assumes restrictions on the growth rate of the tail dependence (see hoga2018detecting). The test is based upon the inverse of the Hill estimator (ref), which we denote by $\widehat{\gamma} \coloneqq 1/\widehat{\alpha}$. The test statistic is defined by

equation[equation omitted — 290 chars of source]

where $t_0 \in (0,1)$ is a tuning parameter and $\widehat{\gamma}(t)$ is the inverse Hill estimator

equation[equation omitted — 161 chars of source]

Using the Hill estimator based on the subsample with only $\floor{kt}$ observations leads to self-normalization of the test statistic $T_N$ and renders a test that is asymptotically pivotal. The limiting distribution is

equation[equation omitted — 98 chars of source]

where $W(t)$ is a standard Brownian motion. Since the test statistic (ref) can be computed using only the Hill estimator and conducting numerical integration, there is no need to estimate the (potentially difficult) tail covariance. The tuning parameter $t_0$ affects the size of the test in finite samples: high values of $t_0$ make the integral in (ref) based on too few differences of $\widehat{\gamma}$, and low values of $t_0$ yield volatile $\widehat{\gamma}$ in (ref) when $t$ is close to $t_0$. Both of these effects may cause size distortions. Therefore we set $t_0 = 0.2$ following the recommendation of hoga2018detecting, who finds that this choice leads to favorable size properties.\footnote{The choice of $t_0$ in hoga2018detecting comes out using an automated selection procedure to choose $k$, which is different than our 5% rule. An earlier version of our paper employs the same automated procedure, which leads to very similar results.} We reject the null $H_0$ when the test statistic $T_N$ is large. According to Table I of hoga2018detecting, the 95 percentile of (ref) for $t_0=0.2$ is 55.44, which we use as the critical value for testing $H_0$ at 5% significance level.

One issue with the test statistic (ref) is that it requires the same number of tail observations $k$ for both cross-sections of capital and labor income. Hence the 5% rule (ref) discussed in Section (ref) becomes problematic as the number of people with capital income in our data set is rather small. Many households do not hold liquid financial wealth and hence have no capital income. The resulting test is thus not feasible since $k_\text{lab}$ based on our 5% rule could be wildly different from $k_\text{cap}$. To overcome this issue, we only test the equality of Pareto exponents for countries that have more than 1,000\xspace positive capital income observations and set $k = \floor{0.05N_\text{cap}}$, where $N_\text{cap}$ is the number of positive capital income observations (these households always have positive labor income), resulting in 342\xspace country-year observations out of 475\xspace. In practice this means that we estimate the Pareto exponent of labor income further in the tail because the sample size of labor income $N_\text{lab}$ tends to be larger than $N_\text{cap}$. However, this is acceptable since the Pareto approximation tends to fit better for smaller $k$. Figure (ref) presents a scatter plot of the estimated labor income Pareto exponent $\widehat{\alpha}_\text{lab}$ using 5% of the full sample ($N_\text{lab}$) and 5% of the sample with positive capital income ($N_\text{cap}$). The fact that most points are close to the 45 degree supports our claim.

figure[figure omitted — 196 chars of source]

Table (ref) in Appendix (ref) shows the test results of the null hypothesis $H_0: \alpha_\text{lab} = \alpha_\text{cap}$. We reject the null in 294\xspace country-year observations out of 342\xspace (86%\xspace) that meet our sample selection criterion. In every single case of rejection, we have $\widehat{\alpha}_\text{lab}>\widehat{\alpha}_\text{cap}$, and therefore we formally confirm the observation in Section (ref) that capital income is more unequally distributed than labor income.

Model of capital and labor Pareto exponents

Our empirical analysis in Section (ref) suggests that

enumerate*• the capital income Pareto exponent is smaller than the labor one (\ie, capital income is more unequally distributed than labor income), and • the correlation between capital and labor income Pareto exponents is weak.

To explain these empirical findings, we present a simple dynamic model of consumption and savings, which builds on one of the authors' prior works MaStachurskiToda2020JET,MaToda2021JET. Our model is more specialized but the characterizations are sharper. The proofs of propositions in this section are deferred to Appendix (ref).

Income fluctuation problem

Time is discrete and denoted by $t=0,1,2,\dotsc$. Let $a_t$ be the financial wealth of a typical agent at the beginning of period $t$ including current income. The agent chooses consumption $c_t\ge 0$ and saves the remaining wealth $a_t-c_t$. The period utility function is $u:(0,\infty)\to \R$, the discount factor is $\beta>0$, the gross return on wealth between time $t-1$ and $t$ is $R_t>0$, and non-financial income at time $t$ is $Y_t>0$. Thus the agent solves

subequations\begin{align} &\maximize && \E_0\sum_{t=0}^\infty \beta^tu(c_t) \\ &\st && a_{t+1}=R_{t+1}(a_t-c_t)+Y_{t+1}, \\ &&& 0\le c_t\le a_t, \end{align}

where the initial wealth $a_0=a>0$ is given, (ref) is the budget constraint, and (ref) implies that the agent cannot borrow (which is without loss of generality according to the discussion in ChamberlainWilson2000). Throughout the rest of the paper we maintain the following assumptions.

asmp[CRRA utility] The utility function exhibits constant relative risk aversion (CRRA) with coefficient $\gamma>0$, so $u(c)=\frac{c^{1-\gamma}}{1-\gamma}$ if $\gamma\neq 1$ and $u(c)=\log c$ if $\gamma=1$.
asmp[\iid shocks] Let $G_{t+1}\coloneqq Y_{t+1}/Y_t$ be gross growth rate of income. The sequence $\set{R_{t+1},G_{t+1}}_{t=0}^\infty$ is independent and identically distributed (\iid).

These assumptions are similar to Carroll2020, except that we allow for stochastic returns on savings. Note that the asset return $R_{t+1}$ and income growth $G_{t+1}$ are potentially mutually dependent. Due to the \iid assumption, the state variables of the income fluctuation problem (ref) are financial wealth $a_t>0$ and current income $Y_t>0$. Exploiting homotheticity (Assumption (ref)), we can reduce the number of state variables to just one, namely the wealth-income ratio (normalized wealth) $\tilde{a}_t\coloneqq a_t/Y_t$. To see this, letting $\tilde{c}_t\coloneqq c_t/Y_t$ be the consumption-income ratio (normalized consumption), dividing the borrowing constraint (ref) by $Y_t$, we obtain $0\le \tilde{c}_t\le \tilde{a}_t$. Similarly, dividing the budget constraint (ref) by $Y_{t+1}$, we obtain

align[align omitted — 224 chars of source]

where $\tilde{R}_{t+1}\coloneqq R_{t+1}/G_{t+1}$ is the asset return relative to income growth. As for the utility function, since

equation*[equation* omitted — 79 chars of source]

(here we interpret $\prod_{s=1}^0\bullet=1$), assuming $Y_0=1$ (which is without loss of generality) and $\gamma\neq 1$, it follows from (ref) that

align[align omitted — 302 chars of source]

where $\tilde{\beta}_t\coloneqq \beta G_t^{1-\gamma}$. The discussion for $\gamma=1$ is similar. Therefore the problem reduces to an income fluctuation problem with CRRA utility, random discount factors $\set{\tilde{\beta}_t}_{t=1}^\infty$, stochastic returns $\set{\tilde{R}_t}_{t=1}^\infty$ on wealth, and constant income ($\tilde{Y}_t\equiv 1$). The general theory of income fluctuation problems with stochastic discounting, returns, and income in a Markovian setting was developed by MaStachurskiToda2020JET. Therefore we immediately obtain the following result. In what follows, we drop the time subscript when no confusion arises.

propSuppose Assumptions (ref), (ref) hold and \begin{equation} \beta\E[ G^{1-\gamma}]<1 \quad and \quad \beta\E [RG^{-\gamma}]<1. \end{equation} Then the income fluctuation problem (ref) has a unique solution. The consumption function can be expressed as $c(a,Y)=Y\tilde{c}(a/Y)$, where $\tilde{c}:(0,\infty)\to (0,\infty)$ is the consumption function of the detrended problem (maximizing (ref) subject to (ref)), which can be computed by policy function iteration.\footnote{See LiStachurski2014 and MaStachurskiToda2020JET for details on policy function iteration.}

Tail behavior of income and wealth

We now characterize the tail behavior of income and wealth in the context of the income fluctuation problem in Section (ref).

To make the model stationary, suppose that agents survive to the next period with probability $v\in (0,1)$ (perpetual youth model as in yaari1965). Whenever agents die, they are replaced by newborn agents. For simplicity, assume that the discount factor $\beta$ in (ref) already accounts for survival probability and that there is no market for life insurance (allowing for life insurance only changes $R$ to $R/v$ and is thus mathematically equivalent after reparametrization). Without loss of generality, suppose that newborn agents start with income $Y_0=1$. Then the income of a randomly selected agent is $Y_T$, where $T$ is a geometric random variable with mean $\frac{1}{1-v}$. By the assumption on income growth, the log income of a randomly selected agent

equation*[equation* omitted — 61 chars of source]

is a geometric sum of \iid random variables, for which we can characterize the tail behavior as follows.

prop[Income Pareto exponent] Suppose that $\operatorname{P}(G>1)>0$ and $1<v\E [G^z]<\infty$ for some $z>0$. Then the cross-sectional income distribution has a Pareto upper tail, whose exponent is the unique positive solution $z=\alpha_Y$ to \begin{equation} v\E [G^z]=1. \end{equation}
proofSee BeareToda-dPL.

To characterize the tail behavior of wealth, we first note that the normalized consumption function $\tilde{c}$ in Proposition (ref) is concave and asymptotically linear with a specific slope.\footnote{CarrollKimball1996 showed the sufficiency of hyperbolic absolute risk aversion (HARA, which includes CRRA) for the concavity of the consumption function. Toda2021JME proved the necessity.}

prop[Concavity and asymptotic linearity] Let everything be as in Proposition (ref). Then $\tilde{c}$ is concave and \begin{equation} \lim_{a\to\infty}\frac{\tilde{c}(a)}{a}=\begin{cases*} 1-(\E[\beta R^{1-\gamma}])^{1/\gamma} & if $\E [\beta R^{1-\gamma}]<1$,\\ 0 & otherwise. \end{cases*} \end{equation}

Using Proposition (ref) and setting $\rho=\min\set{(\E [\beta R^{1-\gamma}])^{1/\gamma},1}$, for high enough asset level, the detrended budget constraint (ref) becomes approximately

equation*[equation* omitted — 73 chars of source]

which is a random multiplicative process (kesten1973 process). Under specific assumptions, MaStachurskiToda2020JET prove that the upper tail of the stationary distribution of normalized wealth $\tilde{a}_t$ has a Pareto lower bound. Although a sharp characterization of the tail behavior is generally difficult, in our setting it is possible to obtain an exact characterization due to concavity and the \iid assumption.

propLet $\rho=\min\set{(\E [\beta R^{1-\gamma}])^{1/\gamma},1}$ and $H=\rho\tilde{R}$. Suppose that \begin{enumerate*} • $R$ is thin-tailed (meaning $\E[R^z]<\infty$ for all $z>0$), • $\log H$ is non-lattice (not supported on an evenly spaced grid), and • $\operatorname{P}(H>1)>0$, and $1<v\E [H^z]<\infty$ for some $z>0$. \end{enumerate*} Then the cross-sectional normalized wealth distribution is either bounded or has a Pareto upper tail, in which case the exponent is the unique positive solution $z=\tilde{\alpha}$ of \begin{equation} v\E [H^z]=1. \end{equation} The Pareto exponent for wealth and capital income is then $\alpha=\min\set{\tilde{\alpha},\alpha_Y}$.

Proposition (ref) is significant despite its simplicity. According to the model, we always have $\alpha\le \alpha_Y$, typically with a strict inequality as we see in the numerical example below. This is in sharp contrast to canonical incomplete market general equilibrium models such as aiyagari1994, where agents can save using only a risk-free asset. In such models, the impossibility theorem of StachurskiToda2019JET implies that the tail behavior of income and wealth is the same, implying $\alpha=\alpha_Y$ in our setting.\footnote{The original proof in StachurskiToda2019JET contained an error; it has been corrected in StachurskiToda2020Corrigendum.} Therefore, unlike canonical incomplete market models, our model can explain the empirical fact that capital income is more unequal than labor income. The key assumption leading to this conclusion is the presence of stochastic returns.

We discuss an analytically solvable example to build intuition.

exmpLet $\Delta>0$ be the length of time of one period and the discount factor be $\beta=\e^{-\delta\Delta}$, where $\delta>0$ is the discount rate. Suppose income grows at a constant rate $g>0$, so $G=\e^{g\Delta}$. Suppose asset return is risk-free, so $R=\e^{r \Delta}$ with $r>0$. Finally, let the survival probability be $v=\e^{-\eta\Delta}$, where $\eta$ is the death rate. Then (ref) becomes \begin{equation*} 1=\e^{-\eta\Delta}\e^{zg\Delta}\iff z=\eta/g, \end{equation*} so the income Pareto exponent is $\alpha_Y=\eta/g$. (This is the classical result of WoldWhittle1957 in discrete-time.) Suppose in addition that $-\eta+r(1-\gamma)<0$ so that $\beta R^{1-\gamma}<1$. Since \begin{equation*} H=(\E [\beta R^{1-\gamma}])^{1/\gamma}\tilde{R}=(\beta R)^{1/\gamma}/G=\e^{(\frac{r-\eta}{\gamma}-g)\Delta}, \end{equation*} solving (ref) the normalized wealth Pareto exponent is \begin{equation*} \tilde{\alpha}=\frac{\eta\gamma}{r-\eta-g\gamma} \end{equation*} assuming $r-\eta-g\gamma>0$. Therefore \begin{equation*} \tilde{\alpha}<\alpha_Y\iff \frac{\eta\gamma}{r-\eta-g\gamma}<\frac{\eta}{g}\iff r>\eta+2g\gamma, \end{equation*} so the wealth (hence capital income) Pareto exponent is smaller than the labor income Pareto exponent if the return on wealth $r$ is sufficiently large. In summary, we obtain the following result: suppose $-\eta+r(1-\gamma)<0$ and let $\alpha_\text{cap},\alpha_\text{lab}$ be the capital and labor income Pareto exponents. Then \begin{equation} \begin{cases*} \alpha_cap=\alpha_lab=\frac{\eta}{g} & if $r\le \eta+2g\gamma$,\\ \alpha_cap=\frac{\eta\gamma}{r-\eta-g\gamma}<\frac{\eta}{g}=\alpha_lab& if $r>\eta+2g\gamma$. \end{cases*} \end{equation} Note that the labor income Pareto exponent $\alpha_\text{lab}=\eta/g$ is highly sensitive to the income growth rate $g$. However, provided that $r>\eta+2g\gamma$, the capital income Pareto exponent $\alpha_\text{cap}$ is not very sensitive to the value of $g$ because the denominator is $r-\eta-g\gamma$. This example is consistent with our result in Section (ref) that the capital Pareto exponent is smaller than the labor Pareto exponent but the two values are only weakly related.

Numerical example

We further examine the tail behavior of income and wealth using a numerical example of the income fluctuation problem (ref). Suppose that asset return is \iid lognormal, so $\log R\sim N((\mu-\sigma^2/2)\Delta,\sigma^2\Delta)$, where $\Delta>0$ is the length of one period, $\mu$ is the expected return, and $\sigma$ is volatility. Suppose every period the agent is “promoted” with some probability, so the income growth rate is

equation*[equation* omitted — 123 chars of source]

where $p\in (0,1)$ is the promotion probability and $g$ is the log income growth rate conditional on promotion. We parametrize the promotion probability as $p=1-\e^{-\Delta/L}$, where $L$ is the expected length of time until a promotion. Using (ref), the labor income Pareto exponent is determined such that

equation[equation omitted — 123 chars of source]

We set the parameter values as in Table (ref). One unit of time corresponds to a year and one period is a quarter, so $\Delta=1/4$. The preference parameters (discount rate and risk aversion) are standard. The death rate of $\eta=0.025$ implies an average (economically active) age of $1/\eta=40$ years. The expected return and volatility roughly correspond to the stock market. We set the labor income Pareto exponent to $\alpha_Y=3$, which is roughly the median value in Figure (ref). Using the survival probability $v=\e^{-\eta\Delta}$ and (ref), the implied value of income growth upon promotion is $g=0.0403$. The wealth Pareto exponent determined by (ref) is then $\alpha=1.201$.

table[table omitted — 462 chars of source]

To numerically solve the income fluctuation problem (ref), we discretize the log asset return $\log R$ using a 7-point Gauss-Hermite quadrature and apply policy function iteration (see Appendix (ref)). After solving the individual problem, we apply the Pareto extrapolation algorithm developed in Gouin-BonenfantTodaParetoExtrapolation to accurately compute the stationary (normalized) wealth distribution. Finally, we also simulate an economy with $10^5$ agents. Figure (ref) shows the results.

figure[figure omitted — 421 chars of source]

Figure (ref) shows the normalized consumption function $\tilde{c}(\tilde{a})$ in the range $\tilde{a}\in [0,100]$. Consistent with Proposition (ref), the consumption function is roughly linear for high asset level. Figure (ref) shows the size distributions of income $Y$ normalized wealth $\tilde{a}=a/Y$ in a log-log plot, both from the theoretical model and the simulation. The fact that the tail probability $\operatorname{P}(X>x)$ exhibits a straight line pattern in a log-log plot suggests that the size distributions have Pareto upper tails, consistent with theory. Furthermore, the slope for income is steeper than that of normalized wealth, so wealth (hence capital income) is more unequally distributed than labor income.

Finally, Figure (ref) shows the income and wealth Pareto exponents when we change the income growth rate $g$ in the range $g\in [0.02,0.1]$, fixing other parameters. Because the income Pareto exponent is inversely proportional to income growth by (ref), the labor income Pareto exponent is highly sensitive to income growth. On the other hand, the wealth (capital income) Pareto exponent does not depend much on income growth by the same intuition as in Example (ref). Thus our model is consistent with our empirical findings in Section (ref) that capital and labor income Pareto exponents are only weakly related.

figure[figure omitted — 175 chars of source]
thebibliography{61} \expandafter\ifx\csname urlstyle\endcsname\relax \else \fi \bibitem[Aiyagari(1994)]{aiyagari1994} S. Rao Aiyagari. \newblock Uninsured idiosyncratic risk and aggregate saving. \newblock Quarterly Journal of Economics, 109\penalty0 (3):\penalty0 659--684, August 1994. \newblock doi: \begingroup \urlstyle{rm}\Url{10.2307/2118417}. \bibitem[Atkinson(2005)]{atkinson2005comparing} Anthony B. Atkinson. \newblock Comparing the distribution of top incomes across countries. \newblock Journal of the European Economic Association, 3\penalty0 (2-3):\penalty0 393--401, May 2005. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1162/jeea.2005.3.2-3.393}. \bibitem[Atkinson and Piketty(2007)]{AtkinsonPiketty2007} Anthony B. Atkinson and Thomas Piketty, editors. \newblock Top Incomes over the Twentieth Century. \newblock Oxford University Press, New York, NY, 2007. \bibitem[Atkinson and Piketty(2010)]{AtkinsonPiketty2010} Anthony B. Atkinson and Thomas Piketty, editors. \newblock Top Incomes: A Global Perspective. \newblock Oxford University Press, New York, NY, 2010. \bibitem[Atkinson et al.(2011)Atkinson, Piketty, and Saez]{atkinson2011top} Anthony B. Atkinson, Thomas Piketty, and Emmanuel Saez. \newblock Top incomes in the long run of history. \newblock Journal of Economic Literature, 49\penalty0 (1):\penalty0 3--71, March 2011. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1257/jel.49.1.3}. \bibitem[Badel et al.(2018)Badel, Daly, Huggett, and Nybom]{BadelDalyHuggettNybom2018} Alejandro Badel, Moira Daly, Mark Huggett, and Martin Nybom. \newblock Top earners: Cross-country facts. \newblock Federal Reserve Bank of St. Louis Review, 100\penalty0 (3):\penalty0 237--257, 2018. \newblock doi: \begingroup \urlstyle{rm}\Url{10.20955/r.100.237-57}. \bibitem[Bandourian et al.(2003)Bandourian, McDonald, and Turley]{BandourianMcDonaldTurley2003} Ripsy Bandourian, James B. McDonald, and Robert S. Turley. \newblock A comparison of parametric models of income distributions across countries and over time. \newblock \emph{Estadist\'{\i}ca}, 55\penalty0 (164):\penalty0 135--152, 2003. \newblock URL \texttt{https://ssrn.com/abstract=324900}. \bibitem[Beare and Toda(2017)]{BeareToda-dPL} Brendan K. Beare and Alexis Akira Toda. \newblock Geometrically stopped {M}arkovian random growth processes and {P}areto tails. \newblock 2017. \newblock URL \texttt{https://arxiv.org/abs/1712.01431}. \bibitem[Bengtsson and Waldenstr\"om(2018)]{Bengtsson_2018} Erik Bengtsson and Daniel Waldenstr\"om. \newblock Capital shares and income inequality: Evidence from the long run. \newblock \emph{Journal of Economic History}, 78\penalty0 (3):\penalty0 712--743, September 2018. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1017/S0022050718000347}. \bibitem[Benhabib et al.(2017)Benhabib, Bisin, and Luo]{BenhabibBisinLuo2017} Jess Benhabib, Alberto Bisin, and Mi Luo. \newblock Earnings inequality and other determinants of wealth inequality. \newblock \emph{American Economic Review: Papers and Proceedings}, 107\penalty0 (5):\penalty0 593--597, May 2017. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1257/aer.p20171005}. \bibitem[Bingham et al.(1987)Bingham, Goldie, and Teugels]{BinghamGoldieTeugels1987} Nicholas H. Bingham, Charles M. Goldie, and Jozef L. Teugels. \newblock \emph{Regular Variation}, volume 27 of \emph{Encyclopedia of Mathematics and Its Applications}. \newblock Cambridge University Press, 1987. \bibitem[Burkhauser et al.(2012)Burkhauser, Feng, Jenkins, and Larrimore]{burkhauser2012recent} Richard V. Burkhauser, Shuaizhang Feng, Stephen P. Jenkins, and Jeff Larrimore. \newblock Recent trends in top income shares in the {U}nited {S}tates: Reconciling estimates from {M}arch {CPS} and {IRS} tax return data. \newblock \emph{Review of Economics and Statistics}, 94\penalty0 (2):\penalty0 371--388, May 2012. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1162/REST_a_00200}. \bibitem[Carroll(2020)]{Carroll2020} Christopher D. Carroll. \newblock Theoretical foundations of buffer stock saving. \newblock \emph{Quantitative Economics}, 2020. \newblock URL \texttt{http://qeconomics.org/ojs/forth/354/354-15.pdf}. \newblock Forthcoming. \bibitem[Carroll and Kimball(1996)]{CarrollKimball1996} Christopher D. Carroll and Miles S. Kimball. \newblock On the concavity of the consumption function. \newblock \emph{Econometrica}, 64\penalty0 (4):\penalty0 981--992, July 1996. \newblock doi: \begingroup \urlstyle{rm}\Url{10.2307/2171853}. \bibitem[Chamberlain and Wilson(2000)]{ChamberlainWilson2000} Gary Chamberlain and Charles A. Wilson. \newblock Optimal intertemporal consumption under uncertainty. \newblock \emph{Review of Economic Dynamics}, 3\penalty0 (3):\penalty0 365--395, July 2000. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1006/redy.2000.0098}. \bibitem[Dan{\'{\i}}elsson and de Vries(1997)]{danielsson1997tail} J{\'{o}}n Dan{\'{\i}}elsson and Casper G. de Vries. \newblock Tail index and quantile estimation with very high frequency data. \newblock \emph{Journal of Empirical Finance}, 4\penalty0 (2-3):\penalty0 241--257, June 1997. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1016/S0927-5398(97)00008-X}. \bibitem[Danielsson et al.(2001)Danielsson, de Haan, Peng, and de Vries]{Danielsson_2001} Jon Danielsson, Laurens de Haan, Liang Peng, and Casper G. de Vries. \newblock Using a bootstrap method to choose the sample fraction in tail index estimation. \newblock \emph{Journal of Multivariate Analysis}, 76\penalty0 (2):\penalty0 226--248, February 2001. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1006/jmva.2000.1903}. \bibitem[Danielsson et al.(2016)Danielsson, Ergun, de Haan, and de Vries]{danielsson2016tail} Jon Danielsson, Lerby M. Ergun, Laurens de Haan, and Casper G. de Vries. \newblock Tail index estimation: Quantile driven threshold selection. \newblock 2016. \newblock URL \texttt{https://ssrn.com/abstract=2717478}. \bibitem[Embrechts et al.(2013)Embrechts, Kl{\"u}ppelberg, and Mikosch]{embrechts2013modelling} Paul Embrechts, Claudia Kl{\"u}ppelberg, and Thomas Mikosch. \newblock \emph{Modelling Extremal Events: For Insurance and Finance}, volume 33. \newblock Springer Science & Business Media, 2013. \bibitem[Fedotenkov(2020)]{Fedotenkov2020} Igor Fedotenkov. \newblock A review of more than one hundred {P}areto-tail index estimators. \newblock \emph{Statistica}, 80\penalty0 (3):\penalty0 245--299, 2020. \newblock doi: \begingroup \urlstyle{rm}\Url{10.6092/issn.1973-2201/9533}. \bibitem[Gabaix(2009)]{gabaix2009} Xavier Gabaix. \newblock Power laws in economics and finance. \newblock \emph{Annual Review of Economics}, 1:\penalty0 255--293, 2009. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1146/annurev.economics.050708.142940}. \bibitem[Gabaix(2016)]{Gabaix2016JEP} Xavier Gabaix. \newblock Power laws in economics: An introduction. \newblock \emph{Journal of Economic Perspectives}, 30\penalty0 (1):\penalty0 185--206, Winter 2016. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1257/jep.30.1.185}. \bibitem[Gabaix and Ibragimov(2011)]{gabaix-ibragimov2011} Xavier Gabaix and Rustam Ibragimov. \newblock Rank$-1/2$: A simple way to improve the {OLS} estimation of tail exponents. \newblock \emph{Journal of Business and Economic Statistics}, 29\penalty0 (1):\penalty0 24--39, January 2011. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1198/jbes.2009.06157}. \bibitem[Gomes and Guillou(2015)]{GomesGuillou2015} M. Ivette Gomes and Armelle Guillou. \newblock Extreme value theory and statistics of univariate extremes: A review. \newblock \emph{International Statistical Review}, 83\penalty0 (2):\penalty0 263--292, August 2015. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1111/insr.12058}. \bibitem[Gouin-Bonenfant and Toda(2018)]{Gouin-BonenfantTodaParetoExtrapolation} {\'E}milien Gouin-Bonenfant and Alexis Akira Toda. \newblock {P}areto extrapolation: An analytical framework for studying tail inequality. \newblock 2018. \newblock URL \texttt{https://ssrn.com/abstract=3260899}. \bibitem[Hall(1982)]{Hall1982} Peter Hall. \newblock On some simple estimates of an exponent of regular variation. \newblock \emph{Journal of the Royal Statistical Society, Series B}, 44\penalty0 (1):\penalty0 37--42, September 1982. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1111/j.2517-6161.1982.tb01183.x}. \bibitem[Hall(1990)]{Hall1990} Peter Hall. \newblock Using the bootstrap to estimate mean squared error and select smoothing parameter in nonparametric problems. \newblock \emph{Journal of Multivariate Analysis}, 32\penalty0 (2):\penalty0 177--203, February 1990. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1016/0047-259X(90)90080-2}. \bibitem[Hill(1975)]{hill1975simple} Bruce M. Hill. \newblock A simple general approach to inference about the tail of a distribution. \newblock \emph{Annals of Statistics}, 3\penalty0 (5):\penalty0 1163--1174, September 1975. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1214/aos/1176343247}. \bibitem[Hoga(2018)]{hoga2018detecting} Yannick Hoga. \newblock Detecting tail risk differences in multivariate time series. \newblock \emph{Journal of Time Series Analysis}, 39\penalty0 (5):\penalty0 665--689, March 2018. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1111/jtsa.12292}. \bibitem[Ibragimov and Ibragimov(2018)]{IbragimovIbragimov2018} Marat Ibragimov and Rustam Ibragimov. \newblock Heavy tails and upper-tail inequality: The case of {R}ussia. \newblock \emph{Empirical Economics}, 54\penalty0 (2):\penalty0 823--837, March 2018. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1007/s00181-017-1239-0}. \bibitem[Jenkins(2017)]{Jenkins2017} Stephen P. Jenkins. \newblock {P}areto models, top incomes and recent trends in {UK} income inequality. \newblock \emph{Economica}, 84\penalty0 (334):\penalty0 261--289, April 2017. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1111/ecca.12217}. \bibitem[Jessen and Mikosch(2006)]{JessenMikosch2006} Anders Hedegaard Jessen and Thomas Mikosch. \newblock Regularly varying functions. \newblock \emph{Publications de l'Institut Math{\'e}matique}, 80\penalty0 (94):\penalty0 171--192, 2006. \bibitem[Kesten(1973)]{kesten1973} Harry Kesten. \newblock Random difference equations and renewal theory for products of random matrices. \newblock \emph{Acta Mathematica}, 131\penalty0 (1):\penalty0 207--248, 1973. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1007/BF02392040}. \bibitem[Li and Stachurski(2014)]{LiStachurski2014} Huiyu Li and John Stachurski. \newblock Solving the income fluctuation problem with unbounded rewards. \newblock \emph{Journal of Economic Dynamics and Control}, 45:\penalty0 353--365, August 2014. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1016/j.jedc.2014.06.003}. \bibitem[{Luxembourg Income Study (LIS) Database}(2021)]{LIS2021} {Luxembourg Income Study (LIS) Database}. \newblock \texttt{http://www.lisdatacenter.org} (multiple countries; September 2019--June 2021), 2021. \newblock Luxembourg: LIS. \bibitem[Ma and Toda(2021)]{MaToda2021JET} Qingyin Ma and Alexis Akira Toda. \newblock A theory of the saving rate of the rich. \newblock \emph{Journal of Economic Theory}, 192:\penalty0 105193, March 2021. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1016/j.jet.2021.105193}. \bibitem[Ma et al.(2020)Ma, Stachurski, and Toda]{MaStachurskiToda2020JET} Qingyin Ma, John Stachurski, and Alexis Akira Toda. \newblock The income fluctuation problem and the evolution of wealth. \newblock \emph{Journal of Economic Theory}, 187:\penalty0 105003, May 2020. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1016/j.jet.2020.105003}. \bibitem[Mirek(2011)]{Mirek2011} Mariusz Mirek. \newblock Heavy tail phenomenon and convergence to stable laws for iterated {L}ipschitz maps. \newblock \emph{Probability Theory and Related Fields}, 151\penalty0 (3-4):\penalty0 705--734, December 2011. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1007/s00440-010-0312-9}. \bibitem[Nirei and Souma(2007)]{nirei-souma2007} Makoto Nirei and Wataru Souma. \newblock A two factor model of income distribution dynamics. \newblock \emph{Review of Income and Wealth}, 53\penalty0 (3):\penalty0 440--459, September 2007. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1111/j.1475-4991.2007.00242.x}. \bibitem[Oancea et al.(2018)Oancea, Pirjol, and Andrei]{Oancea_2018} Bogdan Oancea, Dan Pirjol, and Tudorel Andrei. \newblock A {P}areto upper tail for capital income distribution. \newblock \emph{Physica A: Statistical Mechanics and its Applications}, 492:\penalty0 403--417, February 2018. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1016/j.physa.2017.09.034}. \bibitem[Pareto(1895)]{Pareto1895LaLegge} Vilfredo Pareto. \newblock La legge della demanda. \newblock \emph{Giornale degli Economisti}, 10:\penalty0 59--68, January 1895. \bibitem[Pareto(1896)]{Pareto1896LaCourbe} Vilfredo Pareto. \newblock \emph{La Courbe de la R\'epartition de la Richesse}. \newblock Imprimerie Ch. Viret-Genton, Lausanne, 1896. \bibitem[Pareto(1897)]{Pareto1897Cours} Vilfredo Pareto. \newblock \emph{Cours d'{\'E}conomie Politique}, volume 2. \newblock F. Rouge, Lausanne, 1897. \bibitem[Piketty(2003)]{Piketty2003} Thomas Piketty. \newblock Income inequality in {F}rance, 1901--1998. \newblock \emph{Journal of Political Economy}, 111\penalty0 (5):\penalty0 1004--1042, October 2003. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1086/376955}. \bibitem[Piketty and Saez(2003)]{piketty-saez2003} Thomas Piketty and Emmanuel Saez. \newblock Income inequality in the {U}nited {S}tates, 1913--1998. \newblock \emph{Quarterly Journal of Economics}, 118\penalty0 (1):\penalty0 1--41, February 2003. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1162/00335530360535135}. \bibitem[Reed(2001)]{reed2001} William J. Reed. \newblock The {P}areto, {Z}ipf and other power laws. \newblock \emph{Economics Letters}, 74\penalty0 (1):\penalty0 15--19, December 2001. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1016/S0165-1765(01)00524-9}. \bibitem[Reed(2003)]{reed2003} William J. Reed. \newblock The {P}areto law of incomes---an explanation and an extension. \newblock \emph{Physica A}, 319\penalty0 (1):\penalty0 469--486, March 2003. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1016/S0378-4371(02)01507-8}. \bibitem[Reed and Jorgensen(2004)]{reed-jorgensen2004} William J. Reed and Murray Jorgensen. \newblock The double {P}areto-lognormal distribution---a new parametric model for size distribution. \newblock \emph{Communications in Statistics---Theory and Methods}, 33\penalty0 (8):\penalty0 1733--1753, 2004. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1081/STA-120037438}. \bibitem[Resnick and St{\u{a}}ric{\u{a}}(1997)]{ResnickStarica1997} Sidney Resnick and C{\u{a}}t{\u{a}}lin St{\u{a}}ric{\u{a}}. \newblock Smoothing the {H}ill estimator. \newblock \emph{Advances in Applied Probability}, 29\penalty0 (1):\penalty0 271--293, March 1997. \newblock doi: \begingroup \urlstyle{rm}\Url{10.2307/1427870}. \bibitem[Saez(2001)]{Saez2001} Emmanuel Saez. \newblock Using elasticities to derive optimal income tax rates. \newblock \emph{Review of Economic Studies}, 68\penalty0 (1):\penalty0 205--229, January 2001. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1111/1467-937X.00166}. \bibitem[Saez and Stantcheva(2018)]{SaezStancheva2018} Emmanuel Saez and Stefanie Stantcheva. \newblock A simpler theory of optimal capital taxation. \newblock \emph{Journal of Public Economics}, 162:\penalty0 120--142, June 2018. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1016/j.jpubeco.2017.10.004}. \bibitem[Siow(2015)]{Siow_2015} Aloysius Siow. \newblock Testing {B}ecker's theory of positive assortative matching. \newblock \emph{Journal of Labor Economics}, 33\penalty0 (2):\penalty0 409--441, April 2015. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1086/678496}. \bibitem[Stachurski and Toda(2019)]{StachurskiToda2019JET} John Stachurski and Alexis Akira Toda. \newblock An impossibility theorem for wealth in heterogeneous-agent models with limited heterogeneity. \newblock \emph{Journal of Economic Theory}, 182:\penalty0 1--24, July 2019. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1016/j.jet.2019.04.001}. \bibitem[Stachurski and Toda(2020)]{StachurskiToda2020Corrigendum} John Stachurski and Alexis Akira Toda. \newblock Corrigendum to “{A}n impossibility theorem for wealth in heterogeneous-agent models with limited heterogeneity” [{J}ournal of {E}conomic {T}heory 182 (2019) 1--24]. \newblock \emph{Journal of Economic Theory}, 188:\penalty0 105066, July 2020. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1016/j.jet.2020.105066}. \bibitem[Toda(2011)]{Toda2011PRE} Alexis Akira Toda. \newblock Income dynamics with a stationary double {P}areto distribution. \newblock \emph{Physical Review E}, 83\penalty0 (4):\penalty0 046122, 2011. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1103/PhysRevE.83.046122}. \bibitem[Toda(2012)]{Toda2012JEBO} Alexis Akira Toda. \newblock The double power law in income distribution: Explanations and evidence. \newblock \emph{Journal of Economic Behavior and Organization}, 84\penalty0 (1):\penalty0 364--381, September 2012. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1016/j.jebo.2012.04.012}. \bibitem[Toda(2021)]{Toda2021JME} Alexis Akira Toda. \newblock Necessity of hyperbolic absolute risk aversion for the concavity of consumption functions. \newblock \emph{Journal of Mathematical Economics}, 94:\penalty0 102460, May 2021. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1016/j.jmateco.2020.102460}. \bibitem[Toda and Wang(2021)]{TodaWang2021JAE} Alexis Akira Toda and Yulong Wang. \newblock Efficient minimum distance estimation of {P}areto exponent from top income shares. \newblock \emph{Journal of Applied Econometrics}, 36\penalty0 (2):\penalty0 228--243, March 2021. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1002/jae.2788}. \bibitem[Vermeulen(2018)]{Vermeulen2018} Philip Vermeulen. \newblock How fat is the top tail of the wealth distribution? \newblock \emph{Review of Income and Wealth}, 64\penalty0 (2):\penalty0 357--387, June 2018. \newblock doi: \begingroup \urlstyle{rm}\Url{10.1111/roiw.12279}. \bibitem[Wold and Whittle(1957)]{WoldWhittle1957} Herman O. A. Wold and Peter Whittle. \newblock A model explaining the {P}areto distribution of wealth. \newblock \emph{Econometrica}, 25\penalty0 (4):\penalty0 591--595, October 1957. \newblock doi: \begingroup \urlstyle{rm}\Url{10.2307/1905385}. \bibitem[Yaari(1965)]{yaari1965} Menahem E. Yaari. \newblock Uncertain lifetime, life insurance, and the theory of the consumer. \newblock \emph{Review of Economic Studies}, 32\penalty0 (2):\penalty0 137--150, April 1965. \newblock doi: \begingroup \urlstyle{rm}\Url{10.2307/2296058}.