EconBase
← Back to paper

Fractional trends and cycles in macroeconomic time series

Extracted main text — title through conclusion, appendix excluded. This is what our citation measures are computed over, published so the extraction can be checked by eye.

56,581 characters · 10 sections · 81 citation commands

Rendered from LaTeX for readability, not typeset faithfully. Citation keys are highlighted; maths is left as source; figures, tables and equation environments are summarised rather than reproduced; unrecognised commands are greyed out so nothing is silently dropped. Email addresses are removed.

Fractional trends and cycles in macroeconomic time series

\thispagestyle{empty} \setcounter{page}{0} \paragraph{\bf Abstract.}

spacing{1.15} We develop a generalization of correlated trend-cycle decompositions that avoids prior assumptions about the long-run dynamic characteristics by modelling the permanent component as a fractionally integrated process and incorporating a fractional lag operator into the autoregressive polynomial of the cyclical component. The model allows for an endogenous estimation of the integration order jointly with the other model parameters and, therefore, no prior specification tests with respect to persistence are required. We relate the model to the Beveridge-Nelson decomposition and derive a modified Kalman filter estimator for the fractional components. Identification, consistency, and asymptotic normality of the maximum likelihood estimator are shown. For US macroeconomic data we demonstrate that, unlike $I(1)$ correlated unobserved components models, the new model estimates a smooth trend together with a cycle hitting all NBER recessions. While $I(1)$ unobserved components models yield an upward-biased signal-to-noise ratio whenever the integration order of the data-generating mechanism is greater than one, the fractionally integrated model attributes less variation to the long-run shocks due to the fractional trend specification and a higher variation to the cycle shocks due to the fractional lag operator, leading to more persistent cycles and smooth trend estimates that reflect macroeconomic common sense.

\paragraph{\bf Keywords.} unobserved components, fractional lag operator, long memory, trend-cycle decomposition, Kalman filter.

\paragraph{\bf JEL-Classification.}

C22, C51, E32

Introduction

Unobserved components (UC) models are widely applied in macroeconomic research, e.g.\ to decompose GDP and industrial production into trend and cycle Har1985, MorNelZi2003, Web2011, MorPig2012, to study cyclical consumption Mor2007, and to measure long-run investment HarTri2003. While empirical evidence supports a strong negative correlation of long- and short-run shocks for the aforementioned macroeconomic aggregates, the correlated UC model as proposed by BalWoh2002 and MorNelZi2003 frequently produces a volatile long-run component estimate together with a noisy cycle Web2011, KamMorWo2018, thereby contradicting macroeconomic common sense. In addition, the integration order of the long-run component is subject to debate. While all aforementioned papers model the long-run component as an $I(1)$ process, Cla1987 and OhZiv2006 suggest specifications with $I(2)$ trends that nest the HP filter Gom1999, Gom2001, and $I(0)$ specifications with structural breaks are considered in PerWad2009 and Wad2012.

In this paper, we argue that economically implausible cycle estimates are likely to result from a too restrictive specification of the integration order: If a process is in fact integrated of order greater than one, then misspecifying the long-run component to be $I(1)$ upward-biases the variance estimate of the long-run shocks, yielding a high signal-to-noise ratio. This results in a volatile trend estimate together with a noisy cycle. As will be shown, generalizing the persistence properties from integer integration orders to the fractional domain and estimating the integration order jointly with the other parameters of the model can solve the problem.

Focusing on key macroeconomic indicators, the assumption of integer integration orders (0, 1, or 2) has been contested for several variables. For real GDP, MueWat2017 find that the likelihood is flat around $d=1$, yielding a $90\%$ confidence interval for $d$ that is given by $[0.51, 1.44]$. Inference for $d>1$ is found in Cha1998 for low frequency transformations of income, consumption, investment, exports, and imports for the UK, and in Erg2019 for the GDP of several high-income OECD countries. Finally, DieRud1991 find evidence for $d>1$ for consumption and income.

We contribute to the methodological literature by deriving a fractional trend-cycle decomposition, where the long-run component is allowed to be fractionally integrated ($I(d)$), $d \in \mathbb{R}^+$, whereas the fractional lag operator $L_d = 1-\Delta_+^d$ of Joh2008, that is defined in (ref), enters the lag polynomial of the cyclical component. In UC models fractionally integrated processes allow for richer dynamics of both, trend and cycle components, and nest a broader class of data-generating mechanisms. Contrary to $I(1)$ UC models, they allow for mean-reversion of the long-run component for $d<1$, while $d>1$ assigns a higher persistence to long-run shocks. Particularly in the latter case the variance estimate of the long-run component in $I(1)$ UC models is upward-biased, yielding a high signal-to-noise ratio that causes volatile trend estimates together with noisy, implausible cycles. Conversely, fractionally integrated UC models are likely to adequately capture the true signal-to-noise ratio and produce reliable trend and cycle estimates, as they nest the data-generating mechanism.

In addition, the fractional lag operator $L_d$ yields a weighted sum of past realizations when multiplied to a contemporaneous random variable, and thus it qualifies as a lag operator Joh2008, TscWebWe2013. Since it preserves the integration order of a process for non-negative $d$, the fractional lag operator adds flexibility to the short-run properties of a process rather than influencing the integration order. While the standard lag operator $L = 1 - (1-L)$ subtracts an $I(-1)$ process from a contemporaneous variable, the fractional lag operator $L_d = 1- (1-L)^d_+$ subtracts an $I(-d)$ process. Thus, for equality of the variances of the two generic processes $L z_{1,t} = z_{1, t-1}$, $L_d z_{2,t} = (1-\Delta_+^d)z_{2,t}$, i.e.\ $\operatorname{Var}(z_{1,t-1}) = \operatorname{Var}((1-\Delta_+^d)z_{2,t}) $, with $z_{1,t}, z_{2,t}$ white noise, it must hold that $\operatorname{Var}(z_{1,t}) < \operatorname{Var}(z_{2,t})$ for $d>1$ and vice versa for $d<1$. Consequently for $d>1$, which will be the relevant case in our applications, the variance of the short-run shocks in the fractionally integrated UC model must be greater than in the $I(1)$ case to arrive at the same variance of the cyclical component, yielding a smaller signal-to-noise ratio and, therefore, a more persistent cycle.

Since $d$ is defined on a continuous support and enters the likelihood function as an unknown parameter, our model allows for an endogenous estimation of the integration order jointly with the other model parameters, avoids prior unit root testing and takes into account model selection uncertainty with respect to $d$.

The canonical (reduced) form of the fractional trend-cycle model exhibits a fractional ARIMA representation in the fractional lag operator $L_d$, and directly relates to a generalization of the Beveridge-Nelson decomposition to the fractional domain. We discuss identification and show that consistency and asymptotic normality carries over from the ARFIMA estimator of HuaRob2011 and Nie2015. Contrary to the correlated UC model, that requires an autoregressive cycle of order $p \geq 2$ to uniquely identify trend and cycle, our model is identified for any $p \geq 0$ whenever $d \neq 1$. Finally, we assess the state space representation of the fractional trend-cycle model and propose a computationally efficient modification of the Kalman filter for the estimation of the latent long- and short-run components.

When taking the new fractional UC framework to the data, we demonstrate that it makes an important difference for key questions in empirical macroeconomics. We contribute to the empirical literature by applying our fractional trend-cycle decomposition to US GDP, industrial production, private investment, and personal consumption. Our model nests a wide class of UC models and allows to draw inference on the proper specification of the long-run component for the four series under study. We contrast the trend and cycle estimates from integer-integrated UC models with the results from our fractional decomposition and state differences regarding shape, smoothness, variance and importance of the different components. Especially for industrial production we obtain a decomposition that is in line with economic theory, as the cyclical component captures all NBER recession periods, whereas the correlated UC model clearly fails to produce a plausible cycle. For all time series under study, we estimate a continuous increase of the cyclical components in periods of economic upswing, where the correlated UC model produces a noisy cycle that sharply increases before the economy is hit by a recession.

The structure of the paper is as follows. Section (ref) details the fractional UC model, relates it to well-established trend-cycle decompositions, derives the reduced form, and discusses identification. Section (ref) derives the state space representation, proposes a computationally efficient modified Kalman filter for the estimation of the latent fractional trend and cycle, and discusses parameter estimation. In section (ref) the model is applied to decompose macroeconomic aggregates. Section (ref) concludes.

A fractional trend-cycle decomposition

We define a fractional trend-cycle decomposition of a scalar time series $\{y_{t}\}_{t=1}^{n}$ as the sum of a long-run component $\tau_{t}$ and a cycle component $c_{t}$

align[align omitted — 73 chars of source]

The long-run component $\tau_{ t}$ is characterized by an autocovariance function that decays more slowly than with an exponential rate and, therefore, captures the long-run dynamics of a time series, whereas the cycle component $c_{t}$ is $I(0)$ and accounts for transitory fluctuations of a series around its trend. In contrast to the bulk of the literature on unobserved components models that specifies the stochastic trend component typically as a nonstationary process integrated of order 1 or 2 -- most often as a random walk -- we suggest a more general formulation. $\tau_t$ is specified as a combination of a linear deterministic process and a fractionally integrated series

align[align omitted — 103 chars of source]

where $\mu_0$ and $\mu_1$ are constants, $\eta_{t} \sim \mathrm{i.i.d. \ N}(0, \sigma_{\eta}^2)$, and $d \in \mathbb{R}^+$. The fractional difference operator $\Delta^{d}$ is defined as

align[align omitted — 209 chars of source]

and a $+$-subscript denotes a truncation of an operator at $t \leq 0$, e.g. for an arbitrary process $z_t$, $\Delta_+^d z_t= \sum_{j=0}^{t-1}\pi_j(d) z_{t-j}$ Joh2008. The fractional long-run component $x_t$ adds flexibility to the weighting of past shocks for $d \in \mathbb{R}^+$ and nests the classic integer integrated specifications for $d \in \mathbb{N}$. The memory parameter $d$ determines the rate at which the autocovariance function of $x_t$ decays, and a higher $d$ implies a slower decay. For $d<1$ $x_t$ is mean-reverting, while $d \in [1, 2)$ yields the aggregate of a mean-reverting process. Throughout the paper, we adopt the type II definition of fractional integration MarRob1999 that assumes zero starting values for all fractional processes, and, as a consequence, allows for a smooth treatment of the asymptotically stationary ($d < 0.5$) and the nonstationary ($d \geq 0.5$) case. Due to the type II definition of fractional integration the inverse of the fractional difference operator $\Delta_+^{-d} = (1-L)^{-d}_+$ exists for all $d$, such that we can write

align*[align* omitted — 213 chars of source]

Turning to the transitory component, we allow for an AR($p$) process in the fractional lag operator

align[align omitted — 61 chars of source]

where $\phi(L_{d})=1-\phi_{1}L_{d} - ... - \phi_{p} L_{d}^p$, $L_{d} = 1-\Delta_+^{d}$ is the fractional lag operator Joh2008, and $\varepsilon_{t} \sim \mathrm{i.i.d. \ N}(0, \sigma_{\varepsilon}^2)$. For stability of the fractional lag polynomial $\phi(L_{d})$ the condition of Joh2008 is required to hold. It implies that the roots of $\vert \phi(z) \vert = 0$ lie outside the image $\mathbb{C}_{d}$ of the unit disk under the mapping $z \mapsto 1-(1-z)^{d}$. In fractional models $L_d$ plays the role of the standard lag operator $L_1 = L$, since $(1-L_d) x_t= \Delta_+^d x_t \sim I(0)$. For an arbitrary process $z_t$, $L_d z_t = - \sum_{j=1}^{t-1}\pi_j(d)z_{t-j}$ is a weighted sum of past $z_t$, and hence $L_d$ qualifies as a lag operator. Furthermore, by definition the filter $\phi(L_{d})$ preserves the integration order of a series since $d > 0$.

We do not exclude contemporaneous correlation between trend and cycle innovations. Hence, we allow $\rho = \operatorname{Corr}(\eta_t, \varepsilon_t) \neq 0$, which directly implies $\mathrm{E}(\eta_{t} \varepsilon_{t}) = \sigma_{\eta \varepsilon}\neq 0$. For different time indexes we restrict the cross-correlation to be zero, $\mathrm{E}(\eta_{t} \varepsilon_{s}) = 0 \ \forall t \neq s$. Thus, the long- and short-run shocks are i.i.d.\ N distributed with non-diagonal variance $Q$

align*[align* omitted — 255 chars of source]

Our model is very general in terms of its long-run dynamic characteristics, as it nests the well-known framework of Har1985 for $d=1$, where the long-run component is a random walk with drift, and $c_{t}$ is an autoregressive process of finite order. Correlated shocks as in BalWoh2002, MorNelZi2003, and Web2011 are explicitly allowed. For $d=2$, one obtains the double-drift unobserved components model of Cla1987, and a fractional plus noise decomposition as proposed in Har2002 is obtained by setting $d \in \mathbb{R}^+$, $p=0$.

As shown in (ref) below, similar to the classic UC-ARMA model that exhibits an ARIMA representation MorNelZi2003 the canonical form of our fractional trend-cycle model is an ARIMA($p, d, n-1$) model in the fractional lag operator. To see this, note that $(1- L_{d})[\phi(L_{d})^{-1}]_+ \varepsilon_{t} = \theta^\varepsilon_+(L_d) \varepsilon_t$ is a stable moving average process in the fractional lag operator Joh2008, such that we can write

align[align omitted — 302 chars of source]

where $\psi_+(L_d) = \phi(L_d)\theta_+^u(L_d)$ is a truncated moving average polynomial of infinite order that results from the aggregation of $\phi(L_{d} )\eta_{t} + (1- L_{d})\varepsilon_{t}$. Its existence together with a recursive formula for the coefficients $\psi_j$ is shown in appendix (ref). $u_{t} \sim \mathrm{N}(0, \sigma_u^2)$ holds the disturbances and is Gaussian white noise with $\sigma_u^2 = \sigma_\eta^2 + \sigma_\varepsilon^2 + 2\sigma_{\eta \varepsilon}$, which follows from GraMor1976 for contemporaneously dependent $\varepsilon_t$, $\eta_t$, and $\theta_+^u(L_d) = \sum_{i=0}^{t-1}\theta_i^u L_d^i$, $\theta_0 = 1$, $\theta_i^u = \frac{\sigma_\varepsilon}{\sigma_u}\theta_i^\varepsilon $ for all $i > 0$. While aggregating MA processes in the standard lag operator yields an MA process whose lag length equals the maximum lag order of its aggregates, this does not hold in general for the aggregation of MA processes in the fractional lag operator $L_d$, since $L_d^i u_t$, $L_d^j u_t$ are not independent for $i, j > 1$, $i \neq j$. Only for $p=1$ equation (ref) becomes an ARIMA($1, d, 1$) in the fractional lag operator, since $\eta_t + \varepsilon_t$ and $L_d (\phi_1 \eta_t + \varepsilon_t)$ are independent. For $d \in \mathbb{N}$ the model in (ref) nests the integer-integrated ARIMA models. Due to the inclusion of the fractional lag operator (ref) differs from the standard ARFIMA model. Nonetheless, (ref) exhibits an ARFIMA($n-1$, $d$, $n-1$) representation as $\phi(L_d)$ can be written as an AR($n-1$) polynomial.

Our fractional trend-cycle model can be seen as a generalization of the decomposition of BevNel1981 to the fractional domain. To see this, consider (ref) from which one obtains directly

align*[align* omitted — 178 chars of source]

such that multiplication with $\Delta_+^{-b}$ yields the long- and short-run components

align[align omitted — 170 chars of source]

Equality of the decomposition of BevNel1981 and the UC model in (ref), (ref), and (ref) was shown in MorNelZi2003 for $d=1$. Note that $x_t^{BN}$ and $c_t^{BN}$ are identical to the unobserved components in (ref) and (ref) for any $d$, which follows immediately from plugging $L_d = 1$ in (ref). Consequently, the fractional trend-cycle decomposition generalizes the $I(1)$ Beveridge-Nelson decomposition to the class of ARIMA models in the fractional lag operator.

For $d = 1$ MorNelZi2003 demonstrate that the integer-integrated UC model is not identified for $p=1$, $d= 1$, $\sigma_{\eta \varepsilon}\neq 0$. In that case, imposing the restriction $\sigma_{\eta \varepsilon}=0$ yields a decomposition that is different to the one of BevNel1981. The same is shown by Web2011 for the simultaneous unobserved components model identified by heteroscedasticity and by TreWeb2016 for the multivariate UC model.

In fact, $d= 1$ is the only case where the unobserved components model is not identified for $p = 1$. In any other case, where $d \in \mathbb{R}^+$, $d\neq 1$, we show in the following that the model parameters $\phi(L_{d})$, $d$, $\sigma_{\eta}$, $\sigma_{\varepsilon}$, and $\sigma_{\eta \varepsilon}$ can be uniquely recovered from (ref), which is sufficient for identification. Since $\phi(L_{d})$ and $d$ are obtained directly from the model in its canonical form, we consider $\sigma_{\eta}$, $\sigma_{\varepsilon}$, and $\sigma_{\eta \varepsilon}$ on which identification of the unobserved components model crucially depends.

For $d \neq 1$ the parameters $\sigma_{\eta}$, $\sigma_{\varepsilon}$, and $\sigma_{\eta \varepsilon}$ are obtained from the autocovariance function of $\psi(L_d)u_{t}$ for any $p\geq 0$, whereas $p \geq 2$ is required for $ d=1$, as MorNelZi2003 demonstrate. To see this, we consider $p=2$, for which

align*[align* omitted — 400 chars of source]
align*[align* omitted — 688 chars of source]

Note that $d = 1$ implies $\pi_j(d)=0 \ \forall j>1$ and $\pi_j(2d)=0 \ \forall j > 2$. If now $\phi_{2} = 0$, then $\gamma_j=0 \ \forall j >1$, and hence the model is not identified, as also MorNelZi2003 demonstrate. For $d \neq 1$ the model is identified for any $p\geq0$, since $\gamma_{2} \neq 0$. Contrary to the $I(2)$-model of OhZivCr2008 with three shocks that requires $p \geq 4$, our model is identified for any $p \geq 0$ when $d=2$, since $\gamma_0$, $\gamma_1$, $\gamma_2$ are different from $0$ which is sufficient for the identification of $\sigma_{\eta}$, $\sigma_{\varepsilon}$, and $\sigma_{\eta \varepsilon}$.

State space representation and estimation

In this section we derive a state space representation for the fractional trend-cycle decomposition together with a modified Kalman filter estimator for the unobserved components. Furthermore, we discuss the maximum likelihood estimator for the unknown model parameters and show that consistency carries over from the ARFIMA estimator of HuaRob2011 and Nie2015.

There exists a finite-order state space representation of (ref) and (ref) for fixed sample size $n$, since any type II fractionally integrated process exhibits an autoregressive representation of order $n-1$. Thus, an exact state space representation of the fractionally integrated system yields a state vector of dimension $k \geq 2n$ for $d \notin \mathbb{N}$. Estimation of $\tau_t$, $c_t$ via the Kalman filter then involves the inversion of the $k \times k$ conditional state variance for each $t=1,...,n$, which slows down the Kalman filter substantially for large $n$.

Consequently, different approximations for fractionally integrated processes in state space form have been considered ChaPal1998, Pal2007. HarWei2018 study ARMA($v$, $w$) approximations with $v, w \in \{2, 3, 4\}$ for fractional processes. HarTscWeb2019 then correct for the resulting approximation error of the ARMA approximations for fractionally integrated trends. Their approach keeps the state dimension manageable and is feasible from a computational perspective, as it only requires to calculate the inverse of $\operatorname{Var}[(y_{1}, ...., y_n)']$ once. Furthermore, it yields the identical likelihood function as the exact state space model but is computationally superior. We generalize their method to the fractional trend-cycle model in the following, and thereby provide a computationally feasible exact Kalman filter estimator for $\tau_t$, $c_t$.

To begin with, collect all model parameters in $\theta = (d, \phi_1, ..., \phi_p, \sigma_\eta, \sigma_{\eta \varepsilon}, \sigma_{\varepsilon})'$ and define $\operatorname{E}_\theta$, $\operatorname{Var}_\theta$, $\operatorname{Cov}_\theta$ as the moments given the parameter vector $\theta$. Define $\mathcal{F}_t$ as the $\sigma$-field generated by $y_1, ...., y_t$ and let $z_{t|s} = \operatorname{E}_\theta(z_t| \mathcal{F}_s )$ for any random variable $z_t$. Then the prediction error of the Kalman filter for the exact state space model of (ref), (ref), and (ref) is

align[align omitted — 153 chars of source]

Let $\tilde{x}_{t}$ and $\tilde{c}_t$ denote the approximate long-run and cyclical components defined in detail in (ref) and (ref) below. We will show that the following relationships for the conditional expectations hold

align[align omitted — 138 chars of source]

where $\epsilon_{t}^x$, $\epsilon_{t}^c$ denote the approximation errors of the Kalman filter estimates and are $\mathcal{F}_t$-measurable. Then (ref) can be rewritten as

align[align omitted — 212 chars of source]

with $\ddot{y}_{t+1}= y_{t+1} - \epsilon_{t}^x - \epsilon_{t}^c$ as the approximation-corrected $y_t$. Therefore, the prediction errors of the approximation-corrected model and the exact state space model are the same. Next we derive (ref) and (ref).

The long-run component $x_t$ in (ref) is approximated by an ARMA($v, w$) process $\tilde{x}_{t} = [a(L, d)^{-1}m(L, d)]_+\eta_t = \sum_{j=0}^{t-1} b_{j}(d)\eta_{t-j}$ where $a_0=m_0=1$ and, therefore, $b_0 = 1$. $a(L, d)$ is an AR polynomial of order $v$, whereas $m(L, d)$ is a MA polynomial of order $w$. This yields an approximation error

align[align omitted — 116 chars of source]

The ARMA coefficients are obtained beforehand by minimizing the mean squared error between the Wold representations $x_{t} = \sum_{j=0}^{t-1}\varphi_{j}(d)\eta_{t-j}$ and $\tilde{x}_{t} = \sum_{j=0}^{t-1} b_j(d)\eta_{t-j}$ for a fixed $d$ and sample size $n$. A continuous function that maps from the integration order $d$ to its ARMA coefficients is then obtained by optimizing over a grid of $d$ and smoothing the outcomes using splines. Hence, optimization of the likelihood for the fractional trend-cycle decomposition is conducted over the scalar fractional integration order $d$, and does not involve the estimation of any parameters in $a(L, d)$, $m(L, d)$, such that the dimension of the parameter vector $\theta$ is kept small during the optimization. Details together with a large simulation study are contained in HarWei2018.

The cyclical component $c_{t}$ can be expressed as an AR($n-1$) process in the standard lag operator $\phi(L_{d})c_{t} = \delta_+(L, d, \phi)c_{t}$ that is initialized deterministically with $c_{t}=0$ $\forall t \leq 0$ and where $\delta_+(L, d, \phi)$ results from

align*[align* omitted — 175 chars of source]

An approximation for the fractional cyclical component is obtained by truncating $\delta(L, d, \phi)$ after lag $l$, $\tilde{\delta}(L, d, \phi)=\sum_{j=0}^{l} {\delta}_{j}L^j$, $\tilde{\delta}(L, d, \phi) \tilde{c}_{t} = \varepsilon_{t}$. Note that $\delta(L, d, \phi)$, $\tilde{\delta}(L, d, \phi)$ solely depend on $d$ and $\phi_1,...,\phi_p$. Define $\omega_+(L, d, \phi)=[\delta(L, d, \phi)]^{-1}_+$, $\tilde{\omega}_+(L, d, \phi)=[\tilde{\delta}(L, d, \phi)]^{-1}_+$ as moving average lag polynomials of $c_{t}$, $\tilde{c}_{t}$ in the standard lag operator $L$. The approximation error is then given by

align[align omitted — 281 chars of source]

With these approximations and an expression for the resulting approximation error at hand, we are ready to derive the impact of the two approximations on the Kalman filter estimates of $\tau_{t}$ and $c_{t}$. Note that

align[align omitted — 604 chars of source]

and define $y_{1:t}=(y_{1}, ..., y_{t})' - (\operatorname{E}_\theta[y_{1}], ..., \operatorname{E}_\theta[y_{t}])' $, $\eta_{1:t} = (\eta_{1},...,\eta_{t})'$, and $\varepsilon_{1:t}=(\varepsilon_{1}, ... , \varepsilon_{t})'$. Then the joint distribution can be stated as

align[align omitted — 421 chars of source]

where $\Sigma_{\eta_{1:t}y_{1:t}}= \operatorname{Cov}_\theta(\eta_{1:t}, y_{1:t})$, $\Sigma_{\varepsilon_{1:t}y_{1:t}}= \operatorname{Cov}_\theta(\varepsilon_{1:t}, y_{1:t})$, and $\Sigma_{y_{1:t}}= \operatorname{Var}_\theta(y_{1:t})$ with entries from equations (ref), (ref), and (ref). \\ Let $e_{j}$ be a $t$-dimensional unit vector with a one at column j and zeros elsewhere. Then computing the conditional expectation for (ref) and (ref) respectively delivers

align*[align* omitted — 803 chars of source]

where $\epsilon_{t}^x$, $\epsilon_{t}^c$ are the approximation errors in (ref), (ref), and the last step follows from lemma 1 in DurKoo2012. It is easy to see that the approximation errors $\epsilon_t^x=\sum_{j=1}^{t}(\varphi_{j}(d) - b_{j}(d)) e_{ t+1-j} \Sigma_{\eta_{1:t}y_{1:t}} \Sigma_{y_{1:t}}^{-1} y_{1:t}$ and $\epsilon_t^c = \sum_{j=1}^{t}(\omega_{j} - \tilde{\omega}_{j}) e_{t+1-j} \Sigma_{\varepsilon_{1:t}y_{1:t}} \Sigma_{y_{1:t}}^{-1} y_{1:t}$ solely depend on the parameters $\theta$ and $y_1,...,y_t$. Hence, they are $\mathcal{F}_t$-measurable and can be computed precisely.

Define the approximation-corrected $\ddot{y}_{t+1}= y_{t+1} - \epsilon_{t}^x - \epsilon_{t}^c$. Then the prediction error of the exact state space model and of the approximation-corrected, truncated state space model are identical, which proves (ref). Thus they have the same conditional log likelihood given a set of parameters $\theta$. Consequently, maximization of the conditional log likelihood of the approximation-corrected truncated model solves the same optimization problem as for the exact state space representation but reduces the dimension of the state vector. The state space representation of the approximation-corrected truncated model is derived in appendix (ref).

The exact choice of $v$, $w$ for the ARMA approximation of the fractional trend component and $l$ for the truncation of the fractional lag operator does not affect the equality in (ref), since the approximation-correction yields the exact likelihood function that is identical with a non-truncated model but is computationally superior. Nonetheless, numerical optimization takes longer when $v$, $w$, $l$ are chosen too big, as the Kalman filter then has to invert high-dimensional covariance matrices. As a rule-of-thumb, we suggest $v=w=4$, which keeps the dimension of the state vector small and is found to resemble the dynamics of fractionally integrated Gaussian noise $\Delta_+^d \eta_t$ well, as it yields a better fit than autoregressive and moving average processes of order $50$ HarWei2018. For the cyclical component we suggest $l=10$. We use this specification in all empirical applications that follow.

Finally we comment on the estimation of $\theta$ via maximum likelihood (ML). Under the prerequisites derived in HuaRob2011 the conditional sum-of-squares estimator of (ref), that is asymptotically equivalent to the ML estimator, is consistent and asymptotically normally distributed. As they show, imposing stationarity and invertibility on $\theta_+^u(L_{d})$ together with $u_t$ being white noise is sufficient for consistency and asymptotic normality. Similar results are obtained by Nie2015. Since the ML estimator has the same limit distribution as the conditional sum-of-squares estimator, and since our model in its reduced form satisfies the conditions of HuaRob2011 and Nie2015, their asymptotic results hold for the ML estimator of the reduced form in (ref).

Under identification, the asymptotic results carry over from the reduced form in (ref) to the structural form in (ref), (ref), and (ref). As shown in section (ref), the parameters of the structural form $\theta$ can be uniquely recovered from (ref) for any $p$ if $d \neq1$ (and for any $p \geq 2$ if $d=1$ as shown in MorNelZi2003). Therefore, the maximum likelihood estimator of $\theta$ based on the probability density function of the prediction errors in (ref) is consistent.

Consequently, three different model formulations of the fractional trend-cycle decomposition yield consistent estimates for $\theta$ in (ref), (ref), and (ref) via maximum likelihood, namely the reduced form model in (ref), the exact state space representation based on $y_t$ and the approximation-corrected truncated model based on $\ddot{y}_t$. The latter model allows to estimate $\tau_t$ and $c_t$ directly via the Kalman filter and is computationally superior to the exact state space representation. Therefore, it forms the basis of our empirical analysis in the next section.

Empirical applications

We apply our fractional trend-cycle decomposition in (ref), (ref), and (ref) to extract long-run and transitory components from real GDP, industrial production, gross private domestic investment, and personal consumption expenditures for the US. Trend-cycle decompositions of real economic output are typically conducted to estimate the cyclical deviation of output from its long-run growth path. Examples are HarTri2003, GarRobWr2006, PerWad2009 for log US real GDP and Cla1987, StoWat1999, Web2011 for log US real industrial production. Mor2007 estimates the long-run component of personal consumption, whereas HarTri2003 also consider US investment. Hence, our results from the fractional trend-cycle model can easily be compared and checked against widely used alternatives.

In our application several advantages of the fractional trend-cycle decomposition become apparent. From a methodological perspective the endogenous treatment of the integration order neither requires assumptions about the persistence of a series nor prior unit root testing or differencing. Furthermore, fractional trends offer additional flexibility in modelling the permanent component, which directly affects the estimation of the transitory cycle. From an empirical perspective, we contribute to the literature by providing new insights on the persistence of long-run output, investment and consumption, when the trend component is not restricted to be $I(1)$. In addition, we study cyclical adjustments during economic recessions and comment on the correlation structure between permanent and transitory shocks. Finally, we investigate how establishing the fractional lag operator affects the estimate of the cyclical component. Since $d > 0$, $L_d \varepsilon_t$ is a weighted sum of past $\varepsilon_t$ that is $I(0)$. Consequently, the fractional lag operator allows for a more flexible way of modelling the short-run properties of a series while preserving the integration order.

The data was downloaded from the Federal Reserve Bank of St.\ Louis (mnemonics: GDPC1, INDPRO, PCECC96, GPDIC1), is in quarterly frequency and spans from 1961:1 to 2018:4. All series are seasonally and inflation adjusted and enter the dataset in logs.

To estimate the unknown parameters $\theta$ in (ref), (ref), and (ref) we draw $100$ combinations of starting values from uniform distributions with appropriate support and maximize the log likelihood of the fractional trend-cycle model via the Nelder-Mead algorithm up to a certain relative tolerance. We ignore the approximation-correction, that has a negligible impact on the performance of the ML estimator as shown in HarWei2018, for the estimation of the starting values to speed up the computations. Next, the parameters corresponding to the greatest log likelihood are set as starting values for a finer maximization via the exact approximation-corrected method discussed in section (ref). $p$ is chosen via the Bayesian Information Criterion (BIC).

To study the impact of fractional trends and cycles we introduce a benchmark model that restricts $d = 1$ in (ref) and (ref). Hence, we contrast the fractional trend-cycle model with the $I(1)$ correlated unobserved components model studied in MorNelZi2003 and Web2011. The restricted model is given in equation (ref) below and will be called T-C specification in the following. We will refer to the unrestricted model, that is given in (ref), as FT-FC specification

align[align omitted — 329 chars of source]

Both models allow for correlated permanent and transitory shocks, $\rho= \mathrm{Corr}(\eta_{ t}, \varepsilon_{ t})\neq 0$.

Estimation results together with the log likelihoods are reported in table (ref).

GDP: gradual cyclical upswing

For log GDP, empirical evidence for the exact value of the persistence parameter $d$ is mixed. DieRud1989 and TscWebWe2013 estimate $d$ to be slightly smaller than one, whereas MueWat2017 find that the likelihood is flat around $d = 1$, such that a $90\%$ confidence interval yields $d \in [0.51, 1.44]$. From the exact local Whittle estimator of ShiPhi2005 and the method of GewPor1983 we obtain $\hat{d}^{EW} = 1.24$ and $\hat{d}^{GPH} = 1.24$ with tuning parameter $\alpha = 0.65$ as in ShiPhi2005.

For the fractional trend-cycle model the ML estimator yields $\hat{d}^{FT-FC} = 1.32$, implying that log US real GDP is a non-stable, nonstationary fractional process. As figure (ref) shows, the log likelihood is considerably flat around $\hat{d}^{FT-FC}$, which explains the different results for the persistence parameter in the literature and confirms the findings in MueWat2017. Nonetheless, most of the probability mass clearly lies at $d \geq 1$. Contrary to the benchmark, the FT-FC specification attributes more volatility to the transitory shocks, whereas $\sigma_\eta$ is estimated to be smaller than in the T-C specification.

figure[figure omitted — 650 chars of source]

Figure (ref) plots the decompositions from the T-C and the FT-FC specification in (ref) and (ref). At first glance, it demonstrates that the T-C and the FT-FC decomposition for log US real GDP yield rather similar results, which may be due to the flat likelihood of the FT-FC model around $d=1.32$, and the results coincide with the literature MorNelZi2003, Sin2009. As economic theory suggests, both cyclical components decline during the NBER recession periods. The FT-FC specification suggests a gradual cyclical upswing in non-recession periods and therefore captures an important feature of the business cycle, contrary to the cyclical T-C component that exhibits a steep increase right before a recession period. Similar cycle estimates as from the FT-FC specification are obtained from the nonlinear regime-switching UC-FP-UR model of MorPig2012. Thereby, the parsimonious parametrization of the FT-FC model together with its ability to resemble nonlinear dynamics foster its generality. Furthermore, the fractional trend component is smoother than its $I(1)$ counterpart, as $\sigma_{\eta}$ in table (ref) shows. As KamMorWo2018 demonstrate, forcing the signal-to-noise ratio to be small can yield cycle estimates via correlated $I(1)$ UC models that are in line with economic theory. But as an inspection of their trend estimate shows, this comes with the cost of producing fractionally integrated long-run shocks that violate the white noise assumption. In contrast, fractionally integrated UC models directly estimate a small signal-to-noise ratio without restricting parameters to a certain interval and yield long-run shocks that are $I(0)$. Thus, a high signal-to-noise ratio in $I(1)$ UC models can indicate a violation of the $I(1)$ assumption for the long-run component. As explained in Web2011, the strong negative correlation between $\eta_t$ and $\varepsilon_t$ is typically interpreted as causal impact from long-run shocks to the transitory component, where a positive trend shift yields a negative cyclical adjustment that vanishes over time due to the stationary nature of the transitory component. However, Web2011 also finds significant negative effects in the reverse direction.

Industrial production: plausible cycles in recessions

For log US industrial production we find $\hat{d}^{EW}=1.18$ and $\hat{d}^{GPH}=1.26$ which indicates a violation of the I(1) assumption of the unobserved components model. This is confirmed by the fractional trend-cycle model, for which the ML estimator yields $\hat{d}^{FT-FC}=1.66$. As figure (ref) shows, the likelihood is steep around $\hat{d}$.

figure[figure omitted — 670 chars of source]

Figure (ref) plots the unobserved components estimates from the T-C and the FT-FC specification for log US industrial production. Since $\hat{d}^{FT-FC}=1.66$ is considerably large, whereas $\hat{\sigma}_\eta^{{FT-FC}} = 0.14$ is relatively small compared to the benchmark $\hat{\sigma}_\eta^{{T-C}} = 8.09$, the fractional trend-cycle decomposition yields a smooth trend that only slightly drops during economic recessions, whereas the $I(1)$ counterpart is more erratic. The small ratio $\hat{\sigma}_\eta^{FT-FC}/\hat{\sigma}_\varepsilon^{FT-FC}$ may serve as an explanation for the differences between $\hat{d}^{FT-FC}$ and the nonparametric estimates $\hat{d}^{EW}$, $\hat{d}^{GPH}$, since a small signal-to-noise ratio can downward-bias the latter estimators SunPhi2004.

The cyclical component from the FT-FC specification is in line with the one obtained for log US real GDP, as it captures the dynamics from the business cycle well. It sharply drops during the NBER recession periods and recovers continuously in the aftermath, whereas the T-C cycle tends to increase during economic recessions, thereby contradicting economic theory. The results of Web2011, who shows that a sufficiently long AR polynomial (in his case $p=10$ for monthly industrial production) can produce a more plausible cycle in an $I(1)$ correlated UC setup, are in line with the fractional cycle specification that can be interpreted as an autoregressive process of order $n-1$.

Investment: strong cyclical variation

Turning to log US real gross private domestic investment, the nonparametric estimators yield $\hat{d}^{EW}=1.12$ and $\hat{d}^{GPH}=1.15$, whereas the maximum likelihood estimator for the fractional trend-cycle decomposition returns a slightly larger $\hat{d}^{FT-FC}=1.28$ as shown in table (ref). The likelihood is relatively steep around $\hat{d}^{FT-FC}=1.28$, as figure (ref) indicates.

figure[figure omitted — 686 chars of source]

Figure (ref) shows that the FT-FC specification produces a smoother trend than the T-C benchmark and attributes a larger fraction of total variation to the cyclical component. More in line with economic theory, long-run investment from the FT-FC model is almost linear during economic upswings, whereas the T-C estimate peaks directly before the NBER recession periods. This is especially striking in the 2000s. There, the T-C model ascribes a permanent character to development before and in the great recession. The FT-FC model instead finds a strong cyclical upswing before the great recession, followed by a pronounced slump of the cycle. Regarding the debate on the nature and effects of the recession, this leads to clearly different conclusions.

Consumption: smooth trend

Finally, for log US real personal consumption we estimate $\hat{d}^{EW}=1.40$ and $\hat{d}^{GPH}=1.37$ via the nonparametric estimators. Similarly, the fractional trend-cycle model yields $\hat{d}^{FT-FC}=1.44$, as table (ref) shows.

figure[figure omitted — 689 chars of source]

Contrary to the results obtained for investment, the fractional decomposition attributes less variation to the cyclical component than the T-C benchmark. Hence, transitory consumption is estimated to be less volatile over the business cycle in the FT-FC framework. We find this more to be in line with economic theory than the results obtained from the T-C model, which indicate excessive overconsumption directly before a recession period.

Structural breaks and longer cycles

Since PerWad2009 find that the stochastic long-run component of US GDP is well described by an I(0) process when a trend break in 1973:1 is introduced, we check the impact of the PerWad2009 break on our fractional trend-cycle decomposition. DieIno2001 argue that structural breaks and fractional trends can easily be confused. Hence, the robustness check clarifies whether the better performance of the fractional trend-cycle decomposition results from an ignored trend break.

Table (ref) reports the parameter estimates when a trend break in 1973:1 is allowed. As it shows, neither the integration order estimates $\hat{d}$, nor the autoregressive parameters and variance parameters differ substantially. The correlation between long- and short-run shocks is estimated to be slightly weaker when a trend break is introduced. The likelihood ratio (LR) test suggests that introducing a structural break in 1973:1 does not significantly improve the goodness of fit for GDP (p-value: $0.05$), industrial production (p-value: $0.12$), investment (p-value: $0.38$), and personal consumption (p-value: $0.09$).

The trend-cycle decompositions in figures (ref) -- (ref) remain largely unaffected by the structural break, as figure (ref) in appendix (ref) shows.

As a second robustness check, we include further lags to the cyclical polynomial by setting $p=4$. In a non-fractional setting this implies that the cycle component contains lagged information from four quarters, which we consider as the maximum lag length of a cyclical component for quarterly data. By adding additional lags to the cyclical polynomial, we investigate if an increased flexibility of the cycle yields the same integration order estimates, or if the estimated fractional integration orders $\hat{d}$ are just an artifact from a too restrictive parametrization of $c_t$. Estimation results are given in table (ref) in appendix (ref). For industrial production, investment, and consumption additional lags have no significant impact. For GDP, slightly different autoregressive coefficients for the cycle are obtained, but they do not increase the overall fit of the model significantly, as a comparison of the likelihoods shows. Furthermore the estimated integration order is quite similar. The trend-cycle decompositions are sketched in figure (ref). Since differences between the decompositions presented above and those contained in figure (ref) are negligible, we conclude that our results are robust to additional lags of the cyclical lag polynomials.

Conclusion

We generalized unobserved components models to the fractional domain by modelling the long-run component as a fractionally integrated series together with a cyclical component where the fractional lag operator enters the lag polynomial. We derived the reduced form representation, related the model to the decomposition of BevNel1981, and showed that the model is uniquely identified independent of the lag length of the cyclical polynomial for $d \neq 1$. With the modified Kalman filter for the truncated, approximation-corrected state space representation of our fractional UC model we proposed a computationally feasible exact estimator for the latent components.

In an application to various macroeconomic series for the US, estimates for the cyclical component from the fractional trend-cycle model were often found to better capture the business cycle dynamics than those of a benchmark correlated unobserved components model with an $I(1)$ trend. E.g.\ for industrial production, the fractional trend-cycle model was shown to produce a cycle that is in line with economic theory. Furthermore, the fractional UC models estimated a smoother trend. The reason for the better performance of fractionally integrated UC models compared to $I(1)$ UC models is the smaller signal-to-noise ratio, i.e.\ the ratio of long- and short-run shock variances. For $d>1$, as in our four applications, a violation of the $I(1)$ assumption in $I(1)$ UC models causes an upward-biased estimate of the long-run shock variance, which results in a high signal-to-noise ratio and, therefore, in a volatile trend estimate together with a noisy cycle. In contrast, allowing for fractional trends adequately captures the long-run dynamics of the trend and yields a consistent estimate of the long-run shock variance. In addition, the fractional lag operator $L_d$ attributes a higher variance to the short-run shocks to arrive at the same cyclical variance as in the $I(1)$ benchmark for $d>1$, thereby lowering the signal-to-noise ratio in the fractionally integrated UC model. Thus, the relatively small signal-to-noise ratio in the fractional model produces smooth trend estimates together with persistent cycles that reflect macroeconomic common sense.

The fractional trend-cycle model offers a variety of opportunities for future research. The model may be generalized to the multivariate case, where fractional trends of different persistence with correlated innovations are allowed. A multivariate fractional trend-cycle model would then allow to estimate common fractional trends of cointegrated variables and test for polynomial cointegration. Furthermore, inferential methods that test for the number of common trends or the equality of integration orders could be established. As shown in DieIno2001, fractionally integrated processes and structural breaks are related, since the former class of processes can produce level shifts and since structural breaks can be misinterpreted as $I(d)$ processes. Hence, combining both concepts, e.g.\ in a fractional UC model with regime switching, can be a fruitful challenge for future research.

To applied researchers, the model offers a flexible data-driven method to treat permanent and transitory components in macroeconomic and financial applications. It provides a solution for many issues of model specification that caused uncertainty and debates about realistic trend-cycle decompositions and estimation of recessions. Based on that, also the interaction of trends and cycles can be analyzed.