EconBase
← Back to paper

Large structural VARs with multiple linear shock and impact inequality restrictions

Extracted main text — title through conclusion, appendix excluded. This is what our citation measures are computed over, published so the extraction can be checked by eye.

53,657 characters · 10 sections · 146 citation commands

Rendered from LaTeX for readability, not typeset faithfully. Citation keys are highlighted; maths is left as source; figures, tables and equation environments are summarised rather than reproduced; unrecognised commands are greyed out so nothing is silently dropped. Email addresses are removed.

Large structural VARs with multiple linear shock and impact inequality restrictions

abstract\begin{singlespace} We propose a high-dimensional structural vector autoregression framework that features a factor structure in the error terms and accommodates a large number of linear inequality restrictions on impact impulse responses, structural shocks, and their element-wise products. In particular, we demonstrate that narrative restrictions can be imposed via constraints on the structural shocks, which can be used to sharpen inference and disentangle structurally interpretable shocks. To estimate the model, we develop a highly efficient sampling algorithm that scales well with both the model dimension and the number of inequality restrictions on impact responses and structural shocks. It remains computationally feasible even in settings where existing algorithms may break down. To illustrate the practical utility of our approach, we identify five structural shocks and examine the dynamic responses of thirty macroeconomic variables, highlighting the model's flexibility and feasibility in complex empirical applications. We provide empirical evidence that financial shocks are the most important driver of business cycle dynamics. \end{singlespace}

Keywords: Structural Identification, Sign Restrictions, Narrative Restrictions, Large BVARs

JEL classification: C11, C32, C55, E50

\thispagestyle{empty}

\pagenumbering{arabic}

Introduction

uhlig2005effects popularizes the use of sign restrictions on impulse response functions, offering a flexible and theory-consistent approach to identifying structural shocks in Vector Autoregression (VAR) models. Instead of relying on often implausible zero restrictions, sign restrictions constrain the signs of impulse responses based on economic intuition. Building on this framework, antolin2018narrative introduce narrative restrictions, which incorporate external prior information about specific historical episodes-such as well-documented policy interventions or economic events-to further sharpen set identification.\footnote{The sign restrictions on shocks over distinct time periods may be interpreted as a more agnostic framework than impact sign restrictions, as noted in antolin2018narrative. Nonetheless, it is emphasized that sign restrictions must be grounded in a plausible narrative foundation. } By conditioning structural shocks on particular time periods, narrative restrictions enable researchers to embed economically meaningful prior knowledge directly into the model, enhancing interpretability and credibility of the identified shocks. Using narrative restrictions have become popular recently.\footnote{Important examples that use narrative restrictions include zeev2018can, altavilla2019loan, furlanetto2019immigration, cheng2020revisiting, kilian2020does, kilian2022oil, laumer2020government, redl2020uncertainty, zhou2020refining, antolin2021structural, caggiano2021financial, larsen2021components, ludvigson2021uncertainty, maffei2021does, berger2022unified, fanelli2022sovereign, inoue2022joint, badinger2023measuring, berthold2023macroeconomic, caggiano2023global, conti2023bank, harrison2023structural, herwartz2023point neri2023long, reichlin2023monetary, ascari2023endogenous, boer2024energy and ruth2023monetary.} However, in most applications researcher use rather small dimensional VARs with narrative restrictions. Furthermore, the implementation of narrative restriction rely on accept-and-reject algorithms (see, e.g., antolin2018narrative or giacomini2021identification), which may become computationally burdensome as the number of restrictions increases. Therefore, applications with narrative restriction are limited to VARs with a small number of narrative restrictions, typically for just a single structural shock.

Since the influential work of banbura2010large, large VARs have become a cornerstone of modern structural analysis and forecasting. By incorporating dozens-or even hundreds-of variables, large VARs offer a powerful solution to a key limitation of smaller models: omitted variable bias. This bias can lead to distorted forecasts, unreliable policy conclusions, and incorrect identification of structural shocks. Expanding the information set helps capture the economic environment more comprehensively and reduces the risk of excluding relevant dynamics. As emphasized by hansen2019two and lippi1993dynamic,lippi1994var, using a narrower information set than that available to economic agents can make a model non-fundamental, thus preventing accurate recovery of structural shocks. Building on this foundation, numerous studies-including carriero2009forecasting, giannone2015prior, jarocinski2017granger, huber2019adaptive, chan2024large-have explored various methodological and empirical extensions. Overall, large VARs provide a flexible framework for structural inference in high-dimensional settings to address informational deficiencies.

In this paper, we propose a high-dimensional Structural Vector Autoregression (SVAR) framework capable of accommodating a large number of linear inequality constraints on impact impulse responses, structural shocks, and their element-wise products. While previous literature has primarily focused on small narrative SVAR models, our approach extends these methods by enabling the inclusion of a much larger set of restrictions as well as a greater number of variables. The ability to incorporate a broader set of inequality restrictions enhances the structural set identification, allowing for more precise inference, see antolin2018narrative. Moreover, in cases where impact sign restrictions alone cannot easily distinguish between different structural shocks, additional sign restrictions directly on the structural shocks provide a flexible way to separate them.

Recently, the assumption that VAR errors follow a factor structure has gained popularity, with latent factors being interpreted as structural shocks (see, for instance, Korobilis2022, chan2022large, hauzenberger2022bayesian, banbura2023drives, gambetti2023agreed, pfarrhofer2024high, pruser2024large, korobilis2025exploring and chan2025large). We build on this literature, leveraging the factor structure for several key advantages. First, from a computational perspective, the factor structure enables equation-by-equation estimation, simplifying the estimation process and allowing for the inclusion of a rich set of variables in the model. Second, as the number of variables increases (including different measures of prices and output), it is reasonable to assume that the number of structural shocks is smaller than the number of variables, making a factor-based approach well-suited for high-dimensional settings. Third, and particularly important for our paper, the factors have a clear interpretation as structural shocks. This provides a natural framework for incorporating narrative restrictions in the form of prior distributions on the structural shocks, which allows to develop a simple yet efficient algorithm.

Our approach deviates from existing methods such as those in antolin2018narrative and giacomini2021identification, which typically incorporate narrative restrictions through a pseudo-likelihood. Instead, imposing shock sign restrictions via prior distributions allows us to introduce a new sampling algorithm that offers significant computational efficiency by ensuring that each draw is accepted with certainty. This feature enables us to handle a large number of restrictions on both impact responses and structural shocks, overcoming computational challenges often encountered with accept-and-reject algorithms. Our approach therefore represents a major advancement in the scalability and feasibility of SVAR models, particularly in high-dimensional settings where existing algorithms may be very slow or even infeasible. We compare our proposed algorithm with the approach of antolin2018narrative in a synthetic exercise in which we estimate several ten-variable VARs with different sets of impact and shock sign restrictions. The results indicate that our algorithm achieves substantially faster computation times than the method proposed by antolin2018narrative across the different identification schemes.

In an empirical application, we demonstrate the practical utility of our approach by identifying five structural shocks in a large VAR. In this respect, we follow stock2012disentangling and estimate 30 variable VAR and disentangle structurally interpretable shocks by means of impact and shock sign restrictions.\footnote{Specifically, we use the same set of 30 variables and identify the same structural shocks as in stock2012disentangling. h24 estimates the same 30-variables VAR and identifies the same 5 structural shocks.} Importantly, we show that shock sign restrictions can help to separate structural shocks and to overall sharpen identification. In our empirical illustration, we provide evidence suggesting that financial risk shocks are a key driver of business cycle dynamics. This example underscores the flexibility and scalability of our methodology, which is capable of handling high-dimensional data and a large number of inequality restrictions on the impact responses as well as on the structural shocks.

The remainder of this paper is organized as follows. Section 2 lays out and discusses the econometric framework. Section 3 studies the computational efficiency. Section 4 applies the model to the study of stock2012disentangling to identify and investigate the effects of five structural shocks on the economy. Section 4 concludes.

A large structural VAR with factor structure

Let $\boldsymbol{y}_t = (y_{1,t}, \dots, y_{n,t})'$ represent an $n \times 1$ vector of endogenous variables at time $t$. The model can be expressed as:

equation[equation omitted — 160 chars of source]

where the error term $\boldsymbol{u}_t$ is decomposed as

equation[equation omitted — 86 chars of source]

with $\boldsymbol{v}_t \sim N(\boldsymbol{0}, \boldsymbol{\Sigma})$, where $\boldsymbol{\Sigma} = \text{diag}(\sigma_1^2, \dots, \sigma_n^2)$, and $\boldsymbol{f}_t$ is an $r \times 1$ vector of factors such that $\boldsymbol{f}_t \sim N(\boldsymbol{0}, \boldsymbol{I})$. Concisely, the system can be reformulated as:

equation[equation omitted — 152 chars of source]

where $\boldsymbol{I}_n$ is the identity matrix of dimension $n$, $\otimes$ represents the Kronecker product, and $\boldsymbol{\beta} = \text{vec}([\boldsymbol{b}_0, \boldsymbol{B}_1, \dots, \boldsymbol{B}_p]')$. The vector $\boldsymbol{x}_t = (1, \boldsymbol{y}'_{t-1}, \dots, \boldsymbol{y}'_{t-p})'$ of dimension $k = 1 + np$ contains an intercept and lagged values. The idiosyncratic component $\boldsymbol{v}_t$ accounts for measurement error or other idiosyncratic noise. In contrast, the $r$ latent factors, $\boldsymbol{f}_t$ impact multiple variables in the system and hence have the interpretation as structural shocks. More formally we multiply the model with the generalized inverse of $\boldsymbol{L}$ to obtain:

equation[equation omitted — 135 chars of source]

where $\boldsymbol{A} = (\boldsymbol{L}' \boldsymbol{L})^{-1} \boldsymbol{L}'$ and $\boldsymbol{B} = (\boldsymbol{A} \boldsymbol{b}_0, \boldsymbol{A} \boldsymbol{B}_1, \dots, \boldsymbol{A} \boldsymbol{B}_p)$. Given that $\boldsymbol{v}_t$ is uncorrelated noise, the Central Limit Theorem (see bai2003) implies that $\boldsymbol{A} \boldsymbol{v}_t \to 0$ as $n \to \infty$. Hence, we can write

equation[equation omitted — 107 chars of source]

Thus, $\boldsymbol{v}_t$ is treated as noise shocks without structural meaning, while $\boldsymbol{f}_t$ provides a projection of structural shocks in $\mathbb{R}^r$. Consequently, this formulation supports structural analysis using standard methods, including impulse response functions (e.g., Froni2019, Korobilis2022, chan2022large). Under the assumption of uncorrelated $\boldsymbol{f}_t$ and $\boldsymbol{v}_t$, the conditional variance of $\boldsymbol{u}_t$ is given by:

equation[equation omitted — 138 chars of source]

To separately identify the common and idiosyncratic components, we adopt the condition $r \leq (n-1)/2$ from anderson1956statistical. Under this restriction, for any observationally equivalent model $(\boldsymbol{L},\boldsymbol{\Sigma})$, it holds that $\boldsymbol{L} \boldsymbol{L}' = \boldsymbol{L}^* \boldsymbol{L}^{*'}$ and $\boldsymbol{\Sigma} = \boldsymbol{\Sigma}^*$. However, without additional restrictions, $\boldsymbol{L}$ is not identified. Any orthogonal matrix $\boldsymbol{Q} \in \mathcal{O}(r)$, where $\mathcal{O}(r) = \{ \boldsymbol{Q} \in \mathbb{R}^{r \times r} : \boldsymbol{Q} \boldsymbol{Q}' = \boldsymbol{I}_m\}$, generates an equivalent model via the transformation $\tilde{\boldsymbol{L}} \tilde{\boldsymbol{f}}_t = \boldsymbol{L} \boldsymbol{Q} \boldsymbol{Q}' \boldsymbol{f}_t$. In this paper, we achieve set identification by placing inequality restrictions on $\boldsymbol{L}$ and $\boldsymbol{f}_{t}$.\footnote{As shown below, our model also allows for inequality restrictions on the product of $\boldsymbol{L}$ and $\boldsymbol{f}$.}

Linear inequality restrictions via prior distributions

We consider inequality restriction on the $\boldsymbol{L}$, $\boldsymbol{f}_{t}$, and on their product.\footnote{In many cases we have a strong consensus in economic theory about the signs of impulse responses at impact but not at longer horizons (see, e.g. canova2011business).} We can express these inequality restrictions as follows:

align[align omitted — 340 chars of source]

where $\boldsymbol{l}_i$ denote the elements of $\boldsymbol{L}$ in the $i-$th equation, $\boldsymbol{R}_{i}^{L}$ is $r_i^L\times r$, $\boldsymbol{R}_{\tilde{t}}^{f}$ is $r_{\tilde{t}}^f\times r$, $\boldsymbol{R}_{i\tilde{t}}^{Lf} $ is $r_{i\tilde{t}}^{Lf}\times r$ and $\tilde{t}\in t=1,\dots,T$. We assume that $\ell_i^{L} <\upsilon_i^{L}$, $\ell_{\tilde{t}}^{f}< \upsilon_{\tilde{t}}^{f}$ and $\ell_{i\tilde{t}}^{Lf}< \upsilon_{i\tilde{t}}^{Lf}$. Furthermore, we allow elements in the lower and upper bound to be $-\infty$ or $\infty$ to indicate that some inequity restrictions are single-sided. In many applications $\boldsymbol{R}_i^{L}$, $\boldsymbol{R}_{\tilde{t}}^{f}$ and $\boldsymbol{R}_{i\tilde{t}}^{Lf}$ will be equal to $\boldsymbol{I}_r$. The first set of inequality restrictions can be used to implement sign restrictions on $\boldsymbol{L}$ as in Korobilis2022. The second set of inequality restriction can be used to implement sign restrictions on the structural shocks $\boldsymbol{f}_{t}$ at specific time periods as in antolin2018narrative. Finally, the last set of inequality restrictions can be used to implement a magnitude restrictions on the product of both factors loadings and structural shocks as in Badinger2023. In particular, Badinger2023 impose that one particular shock explains more then half of the unexplained movement of a certain variable at a specific period. This provides a flexible framework for the identification of structural shocks. Instead of only relaying on impact sign restrictions, researchers can also use shock sign restrictions to distinguish between different structural shocks. Or even use a single magnitude restriction as in Badinger2023 to separate one shock from the others.\footnote{It should be noted that non-linear restrictions can be incorporated via an accept-rejection step. However, this may significantly increase the computational burden or even render sampling infeasible. If such non-linear restrictions can be well approximated by linear restrictions that fit within our framework, the accept-rejection step may become feasible and the additional computational cost can be kept minimal.}

We implement our inequality restrictions using truncated Gaussian prior distributions. In particular we assume

align[align omitted — 642 chars of source]

where $\tilde{T}$ denotes the set of periods in which a restrictions is specified.

We complete our model specification by describing the further prior distributions. For the variance terms of the noise we assume $\sigma_{j}^2 \sim \mathcal{IG}(\alpha_0,\beta_0)$. In our empirical application, we set $\boldsymbol{l}_{0,i}= \boldsymbol{0}$, $ \boldsymbol{V}_{\boldsymbol{l}_i}=10 \times \boldsymbol{I}_r$ and $\alpha_0=\beta_0=0$. In high-dimensional settings such as large VARs, it is important to impose shrinkage prior to mitigate the risk of overfitting. We follow Korobilis2022 and use the horseshoe prior proposed by carvalho2010horseshoe. Let $\boldsymbol{\beta}_i$ the VAR coefficients in the $i-$th equation, $i=1,\dots,n$ and $\beta_{i,j}$, the $j-$th coefficient in the $i-$th equation. Consider the prior for $\beta_{i,j}$, $i=1,\dots,n$ and $j=2,\dots,k$:

align[align omitted — 151 chars of source]

where $C^+(0,1)$ denotes the standard half-Cauchy distribution. The hyperparameter $\lambda_i$ is the global variance components that are common to all elements in $\boldsymbol{\beta}_i$ , whereas each $\psi_{i,j}$ is a local variance component specific to the coefficients $\beta_{i,j}$. To facilitate sampling, we follow makalic2015simple and use the following latent variables representations of the half-Cauchy distributions:

align[align omitted — 241 chars of source]

for $i=1,\dots,n$ and $j=2,\dots,k$.

Gibbs Sampler

In this section, we develop an efficient posterior sampler to estimate the model. To impose the inequality restrictions on the factors or factor loadings we use the sampler from botev2017normal. In contrast to existing rejecting sampling approaches (see, e.g. antolin2018narrative or giacomini2021identification) each draw is accepted with certainty. Posterior draws can be obtained by sampling sequentially from the conditional distributions:

enumerate$p(\boldsymbol{f}|\boldsymbol{y},\boldsymbol{\beta}, \boldsymbol{L},\boldsymbol{\Sigma}, \boldsymbol{\lambda}, \boldsymbol{\psi}, \boldsymbol{z}_{\lambda},\boldsymbol{z}_{\boldsymbol{\psi}})=p(\boldsymbol{f}|\boldsymbol{y}, \boldsymbol{\beta}, \boldsymbol{L}, \boldsymbol{W}, \boldsymbol{\Sigma})$; • $p( \boldsymbol{L}|\boldsymbol{y}\boldsymbol{\beta},\boldsymbol{f},\boldsymbol{\Sigma}, \boldsymbol{\lambda}, \boldsymbol{\psi}, \boldsymbol{z}_{\lambda},\boldsymbol{z}_{\boldsymbol{\psi}})=\prod_{i=1}^n=p(\boldsymbol{l}_i|\boldsymbol{y}_i, \boldsymbol{\beta}_i,\boldsymbol{f}, \boldsymbol{\sigma}^2_i)$$p(\boldsymbol{\beta}|\boldsymbol{y}, \boldsymbol{L}, \boldsymbol{f},\boldsymbol{\Sigma}, \boldsymbol{\lambda}, \boldsymbol{\psi}, \boldsymbol{z}_{\lambda},\boldsymbol{z}_{\boldsymbol{\psi}})=\prod_{i=1}^n=p(\boldsymbol{\beta}_i|\boldsymbol{y}_i,\boldsymbol{l}_i,\boldsymbol{f}, \boldsymbol{\sigma}^2_i)$$p(\boldsymbol{\Sigma}|\boldsymbol{y},\boldsymbol{\beta}, \boldsymbol{L},\boldsymbol{f}, \boldsymbol{\lambda}, \boldsymbol{\psi}, \boldsymbol{z}_{\lambda},\boldsymbol{z}_{\boldsymbol{\psi}})=\prod_i^n p(\sigma_i^2|\boldsymbol{y}_i,\boldsymbol{f}_i,\boldsymbol{l}_i,\boldsymbol{\beta}_i) $; • $p(\boldsymbol{\lambda}|\boldsymbol{y},\boldsymbol{\beta}, \boldsymbol{L},\boldsymbol{f},\boldsymbol{v}, \boldsymbol{\psi}, \boldsymbol{z}_{\lambda},\boldsymbol{z}_{\boldsymbol{\psi}})=\prod_{l=1}^n p(\lambda_i|\boldsymbol{\beta}, \boldsymbol{\psi}, z_{\lambda_i})$; • $p(\boldsymbol{\psi}|\boldsymbol{y},\boldsymbol{\beta}, \boldsymbol{L},\boldsymbol{f},\boldsymbol{\Sigma}, \boldsymbol{\lambda}, \boldsymbol{z}_{\lambda},\boldsymbol{z}_{\boldsymbol{\psi}})=\prod_{i=1}^n\prod_{j=2}^k p(\psi_{i,j}|\beta_{i,j}, \lambda_i,z_{\psi_{i,j}})$; • $p( \boldsymbol{z}_{\lambda}|\boldsymbol{y},\boldsymbol{\beta}, \boldsymbol{L},\boldsymbol{f},\boldsymbol{\Sigma}, \boldsymbol{\lambda}, \boldsymbol{\psi},\boldsymbol{z}_{\boldsymbol{\psi}})=\prod_{i=1}^n p(z_{\lambda_i}|\lambda_i)$; • $p(\boldsymbol{z}_{\boldsymbol{\psi}}|\boldsymbol{y},\boldsymbol{\beta}, \boldsymbol{L},\boldsymbol{f},\boldsymbol{\Sigma}, \boldsymbol{\lambda}, \boldsymbol{\psi}, \boldsymbol{z}_{\lambda})=\prod_{i=1}^n\prod_{j=2}^k p(z_{ \psi_{i,j}}| \psi_{i,j})$,

with $\boldsymbol{y}_i=(y_{i,1},\dots,y_{i,T})'$ be a $T\times 1$ vector of observations of the $i-$th variable.

Step 1 First, we sample $\boldsymbol{f}_t$ for $t=1,\dots,T$. We can use standard regression results (see, e.g., chan2019bayesian) to obtain

equation[equation omitted — 152 chars of source]

where

equation[equation omitted — 315 chars of source]

We use the efficient truncated Normal generator provided by botev2017normal in order to sample the factors.

Step 2 Second, we sample $\boldsymbol{L}$. Given the latent factors $\boldsymbol{f}$, the VAR becomes $n$ unrelated regressions and we can sample $\boldsymbol{L}$ equation by equation. Remember that $\boldsymbol{\beta}_i$ and $\boldsymbol{l}_i$ denote, respectively, the VAR coefficients and the factor loadings in the $i-$th equation. Then, the $i-$th equation of the VAR can be expressed as

equation[equation omitted — 133 chars of source]

where $\boldsymbol{F}=(\boldsymbol{f}_1,\dots,\boldsymbol{f}_r)$ the $ T\times r$ matrix of factors with $ \boldsymbol{f}_i=(f_{i,1},\dots,f_{i,T})'$. The vector of noise $\boldsymbol{v}=(v_{i,1},\dots,v_{i,T})'$ is distributed as $N(\boldsymbol{0},\boldsymbol{I}_T \sigma_i^2)$.

Then using standard linear regression results, we get

equation[equation omitted — 384 chars of source]

where

align*[align* omitted — 360 chars of source]

We use the efficient truncated Normal generator provided by botev2017normal in order to sample $\boldsymbol{l}_i$.

Step 3 Third, we sample $\boldsymbol{\beta}$ equation by equation based on ((ref)). Equation by equation estimation simplifies the estimation and allows for the estimation with a large number of variables. Again, using standard linear regression results, we get

equation[equation omitted — 175 chars of source]

where

align*[align* omitted — 379 chars of source]

Step 4 Next, we sample $\sigma^2_{i}$ for $i=1,\dots,n$. Given $\boldsymbol{f}$ the model reduces to $n$ independent linear regressions. Therefore, we can use standard regression results (see, e.g., chan2019bayesian) to obtain

equation[equation omitted — 236 chars of source]

Step 5 Lastly, we sample the hyperparameter $\lambda_{i}$ and $\psi_{i,j}$ from our shrinkage prior for the VAR coefficients as well as the mixing variables $z_{\lambda_{i}}$ and $z_{\psi_{i,j}}$. Using the latent variable representation of the half Cauchy distribution, we obtain

align*[align* omitted — 362 chars of source]

which is the kernel of the following inverse-gamma distribution:

equation[equation omitted — 164 chars of source]

Furthermore,

align*[align* omitted — 434 chars of source]

which is the kernel of the following inverse-gamma distribution:

equation[equation omitted — 205 chars of source]

Finally, we sample the latent variables $\boldsymbol{z}_{\boldsymbol{\psi}}$ and $\boldsymbol{z}_{\boldsymbol{\lambda}}$. In particular, $z_{\psi_{i,j}}\sim \mathcal{IG}(1,1+\psi_{i,j}^{-1})$ for $i=1,\dots,n$ and $j=2,\dots,k$. Similarly, we have $z_{\lambda_{l}}\sim \mathcal{IG}(1,1+\lambda_i^{-1})$ for $i=1,\dots,n$.

Computational Efficiency

This section evaluates the computational performance of our proposed algorithm in comparison to the approach developed by antolin2018narrative (henceforth DDR18). To assess the relative efficiency, we consider three different structural VAR specifications with $n=10$ variables, lag length $p=4$, $k=5$ structural shocks, and a sample size of $T=148$ observations.\footnote{In our empirical application in Section (ref), we use quarterly data from 1983Q1 to 2019Q4, corresponding to $T = 148$. We identify $k=5$ structural shocks and set the lag length to $p=4$, consistent with the simulation design here.}

Each VAR specification differs in the number of shock sign restrictions imposed: (i) 15 impact sign restrictions and no shock sign restrictions, (ii) 15 impact sign restrictions and 3 shock sign restrictions, and (iii) 15 impact sign restrictions and 6 shock sign restrictions. For each specification, we simulate 10 synthetic data sets and report the average runtime required by both algorithms to generate 100 admissible draws.\footnote{For our algorithm, we discard the first 1,000 draws as burn-in and retain every \nth{10} draw thereafter, yielding an effective sample of 1,000 draws. For the DDR18 algorithm, we adhere to the specifications outlined by antolin2018narrative, employing 1,000 importance-weighted draws and capping the number of attempted draws of the reduced-form parameters and the $\mathbf{Q}$ matrix at 100 each. We use flat priors as antolin2018narrative. It is important to note that antolin2018narrative estimate a three-variable VAR, which naturally requires less prior shrinkage than the ten-variable VAR used in our simulations. This difference affects computational burden and must be considered when interpreting runtime comparisons. Additionally, antolin2018narrative generate independent posterior draws for the reduced-form coefficients.} Details of the data-generating process are provided in Appendix (ref).

table[table omitted — 906 chars of source]

Table (ref) summarizes the computational comparison between the proposed algorithm and DDR18 across the three VAR configurations. It is evident that the proposed algorithm achieves substantially faster runtimes in all cases. As highlighted in Footnote (ref), the DDR18 algorithm was originally developed for small-scale VARs, whereas our approach is explicitly designed for high-dimensional VARs with factor error structures, which requires more sophisticated shrinkage methods. Importantly, in our simulation design, we construct the data-generating process such that the imposed shock sign restrictions align with the true signs of the underlying structural shocks. In the absence of this feature, the DDR18 algorithm fails to find a sufficient number of admissible draws.\footnote{See Appendix (ref) for a detailed discussion of the data-generating process. Our computations were carried out using MATLAB R2024b on an Intel Core Ultra 7 with a 1.70 GHz base speed and 8 GB of RAM.}

Table (ref) clearly shows that the proposed algorithm achieves substantially faster runtimes than the DDR18 algorithm. Notably, the runtime of the proposed method remains stable even as the number of shock sign restrictions increases. In contrast, as the number of variables $n$ or the number of identifying restrictions grows - e.g., to $n = 20$ or $n = 30$ with more impact and shock sign restrictions - the DDR18 algorithm struggles or even fails to find admissible draws within a feasible timeframe. These results underscore the computational advantages of our approach and highlight its suitability for estimating large-scale VARs with extensive identifying restrictions. In the empirical application that follows, we apply the proposed algorithm to a thirty-variable VAR identified with 60 impact sign restrictions and 26 shock sign restrictions.

Application to stock2012disentangling

This section presents an empirical application based on the framework of stock2012disentangling, who estimate a 30-variable vector autoregression (VAR) to assess the relevance of five macroeconomic shocks during the Great Financial Crisis. This empirical application serves the purpose to illustrate the effectiveness of our proposed algorithm to combine impact and shock sign restrictions in large a structural VAR.\footnote{h24 also estimates the same 30-variable VAR and identifies the structural shocks using multiple instruments in a proxy SVAR.} In line with their approach, we estimate a large-scale VAR model comprising the same set of variables, simultaneously identifying five structural shocks: an oil supply (news) shock, a monetary policy shock, a technology shock, a financial risk shock, and a government spending shock. Differing from stock2012disentangling, we employ the above-outlined structural VAR with factor error structure. Further, we extend the sample period to span from 1983Q1 to 2019Q4 and employ both impact sign and shock sign restrictions for the identification of structural shocks.\footnote{bgk24 note that this sample period captures a relatively stable macroeconomic environment, excluding the heightened oil price volatility of the 1970s as well as the COVID-19 pandemic and the geopolitical shock resulting from the Russian invasion of Ukraine.} A detailed overview of the 30 variables included in the VAR is provided in Section (ref). Following the transformations employed in chan2025rankingrestrictions, most variables are expressed in logarithms. Given the quarterly frequency of the data, the model is estimated using a lag length of $p=4$. Posterior inference is conducted using the described Gibbs sampler. Specifically, we perform 6,000 iterations, discarding the first 1,000 as burn-in and retaining every 10th of the remaining draws for inference.

Impact Sign and Shock Sign Restrictions

In this section, we outline and discuss the implementation of impact sign and shock sign restrictions. Particular emphasis is placed on the shock sign restrictions for 2001Q3 and 2008Q4, given their role in disentangling structural shocks. The reasoning of the remaining shock sign restrictions are deferred to Appendix (ref).

Regarding the impact sign restrictions imposed on the $L$-matrix via $\ell_i^{L} < \boldsymbol{R}_i^{L} \boldsymbol{l}_i' < \upsilon_i^{L}$ in Equation (7), we draw on a well-established literature on shock identification using sign restrictions. For the identification of a positive oil supply (news) shock, as presented in Table (ref), we follow banbura2023drives and assume that such a shock leads to an increase in oil prices, as well as in both consumer and producer prices.\footnote{banbura2023drives elaborate on the identification of oil supply and demand shocks, referring to peersman2009oil, kilian2008economic, kilian2009not, kilian2014role, and baumeister2019structural. Notably, our oil supply news is closely related to the shock identified by kaenzig2021, where a positive shock implies the expectation of lower oil supply.} In addition, we assume that a positive oil supply shock induces a monetary tightening, depresses stock prices, and worsens labor market conditions. Output and its components are left unrestricted. Concerning monetary policy shocks, we adopt the identification strategy of chan2025rankingrestrictions, under which a tightening of monetary policy is associated with increases in interest rates and bond yields, declines in stock prices, and a contraction in economic activity.\footnote{Extensive discussions on the identification of monetary policy shocks in VAR frameworks can be found in sims1972money, sims1980macroeconomics, sims1986forecasting, faust1998robustness, canova2005var, canova2002monetary, uhlig2005effects, and rubio2010structural.} The technology shock is identified following chan2025rankingrestrictions, with the assumption that it increases economic activity, raises labor demand, and exerts downward pressure on prices. In response to such a shock, stock markets are expected to appreciate, while the policy rate declines in reaction to the deflationary environment.\footnote{For additional discussion on the sign identification of technology shocks, see dedola2007does and peersman2009technology.} To identify a financial risk shock, we assume a widening of the yield spread between corporate and ten year government bonds, a decline in stock market valuations, and a decrease in consumer prices, an approach consistent with chan2025rankingrestrictions.\footnote{See also Korobilis2022 and furlanetto2019identification for further discussion on the identification of financial shocks.} Finally, a government spending shock is identified by assuming an increase in government expenditures, a reduction in unemployment, and a rise in (stock) prices. This identification strategy also aligns with chan2025rankingrestrictions.\footnote{See laumer2020government and forni2010fiscal for detailed treatments of fiscal shock identification.}

As previously noted, we impose shock sign restrictions on the oil supply, financial risk, and government spending shocks in 2001Q3 and 2008Q4 through $\ell_{\tilde{t}}^{f} < \boldsymbol{R}_{\tilde{t}}^{f} \boldsymbol{f}_{\tilde{t}} < \upsilon_{\tilde{t}}^{f}$ in Equation (8), as these restrictions are essential for the separation of the respective shocks. In 2001Q3, the terrorist attacks on September 11th resulted in a sudden and pronounced increase in financial risk, reflecting heightened uncertainty across financial markets. The temporary closure of the New York Stock Exchange and the flight of investors toward safe-haven assets such as Treasury bonds substantiate this assessment.\footnote{This interpretation is further supported by the Financial Stress Index; see stlfsi4.} In terms of the oil shock, we assume it to be positive, as uncertainty surrounding oil supply rose significantly following the attacks, with market participants anticipating potential production cuts and hence, higher oil prices. Additionally, we assume a positive government spending shock in response to the enactment of the Emergency Supplemental Appropriations Act on September 18th, 2001.\footnote{For further details, see publiclaw107-38, ramey2018government, and blanchard2002empirical.} For the fourth quarter of 2008, we impose a negative oil supply news shock, a positive financial risk shock, and a negative government spending shock. The bankruptcy of Lehman Brothers in late September 2008 precipitated widespread turmoil in global financial markets, leading to a significant increase in financial uncertainty during the subsequent quarter. Consequently, we posit a positive financial risk shock. \footnote{See again stlfsi4 for empirical evidence.} Regarding the oil supply shock, the financial crisis contributed to a downward revision of oil price expectations, thereby alleviating supply constraints. \footnote{In fact, the Organization of the Petroleum Exporting Countries (OPEC) held two meetings in 2008Q4 to address the decline in oil prices opecmeetings2008.} In terms of government spending, the election of President Obama resulted in a reduced expectation of government expenditure due to announcements of cuts in military spending. \footnote{Additionally, the American Recovery and Reinvestment Act, aimed at stimulating the economy, was announced by President Obama in January 2009 and signed into law in February 2009 ARRA2009.} As illustrated in Table (ref), the narrative shock sign restrictions for both the financial risk and government spending shocks are critical for disentangling the two shocks. The additional shock sign restrictions further refine the identification process. It is noted that, for the monetary policy and financial risk shocks, in principle, it would be possible for the two shocks to exhibit the same sign patterns. However, given our set of restrictions, the data do not support the same sign patterns for both shocks, and thus our restrictions are sufficient to disentangle the shocks. In particular, we verify the uniqueness of the sign patterns for every posterior draw of the impact matrix $L$ and the structural shocks $f_{t}$.

table[table omitted — 2,870 chars of source]

Empirical Results

In this section, we present selected results from our empirical analysis. Figure (ref) displays the impulse response functions (IRFs) of key macroeconomic and financial variables, including real GDP, industrial production, unemployment, the Consumer Price Index (CPI), and the S&P 500 index, in response to identified structural shocks. The oil supply news shock, which manifests as an increase in oil prices, leads to a contraction in economic activity, a decline in labor demand, and an increase in consumer prices. As expectations about future economic conditions deteriorate, equity markets also decline. These dynamics are consistent with the empirical findings in the literature (e.g., kaenzig2021, kilian2009impact, kilian2009not, hamilton2003oil, ramey2011oil). A contractionary monetary policy shock is found to reduce output, raise unemployment, and lower both prices and equity valuations (christiano2005nominal, gertler2015monetary, jarocinski2020deconstructing). This response pattern suggests that the identified monetary policy shock is purged from any information effects or central bank signaling, thereby capturing a pure policy shock (see jarocinski22, bs23a, bs23b). A positive technology shock increases both output and labor demand, while exerting downward pressure on prices due to productivity improvements. The stock market responds favorably. This is largely in line with existing evidence on the expansionary effects of technological innovations (peersman2009technology). A shock to financial risk has strongly contractionary effects: output and prices decline persistently, unemployment rises, and stock markets fall. These results corroborate prior findings on the macroeconomic consequences of financial stress (gilchrist2012credit, caldara2016macroeconomic, christiano2014risk, jurado2015measuring, Korobilis2022). Finally, an increase in government spending yields an expansion in output, a reduction in unemployment, upward pressure on prices, and a short-run increase in equity prices. These responses are consistent with a substantial body of empirical work (blanchard2002empirical, ramey1998costly, ramey2011identifying, auerbach2012measuring, mountford2009effects, laumer2020government). The complete set of impulse responses is provided in Section (ref) of the Online Appendix.

figure[figure omitted — 616 chars of source]

Table (ref) presents the forecast error variance decomposition (FEVD) of GDP across various forecast horizons. The FEVD illustrates the proportion of the variance in GDP forecast errors that can be attributed to each of the identified structural shocks. Across all horizons, financial risk shocks emerge as the dominant driver of GDP fluctuations. In contrast, government spending shocks exhibit only a moderate contribution to the forecast error variance of GDP, while monetary policy and technology shocks account for a relatively smaller share of GDP fluctuations across all horizons. Our empirical results are in line with h24 who estimates the same 30-variable VAR and proxy-identifies the 5 structural shocks. In the Online Appendix, Section (ref) shows the historical decomposition of GDP corroborating the finding that financial risk shocks are the largest driver of the business cycle. Further, in Section (ref) we show that if the sample is shortened to 2012 financial risk shocks take on an even more important role in driving the business cycle.\footnote{Further results on the above-explained shortened sample can be obtained on request. It is noted that our results from the shortened sample are even more similar to h24 as their baseline sample ends in 2012.}

table[table omitted — 821 chars of source]

Conclusion

This paper presents a novel high-dimensional SVAR framework that accommodates a large number of linear inequality restrictions on both impact impulse responses and structural shocks, including their interactions. By incorporating a factor structure in the error terms and introducing a computationally efficient sampling algorithm that avoids rejection-based procedures, our approach addresses key limitations of existing methods - most notably, their limited scalability in high-dimensional settings. The flexibility to combine sign and shock restrictions enhances the precision of structural shock identification and enables economically meaningful inference even in large systems.

In a synthetic data exercise, we demonstrate the advantages of our proposed algorithm in large-scale VAR settings with multiple impact and shock sign restrictions, comparing our results to those of antolin2018narrative. In our empirical application, we identify five structural shocks in a VAR with 30 variables, highlighting the practical utility of the method, particularly in revealing the prominent role of financial shocks in driving business cycle fluctuations. Regarding narrative shock sign restrictions, we show their effectiveness in sharpening inference by shrinking the identified set. More importantly, we demonstrate that shock sign restrictions can facilitate the disentanglement of structurally distinct shocks, as exemplified by the separation of financial risk and government risk shocks.

Overall, the proposed framework substantially broadens the scope for credible and computationally tractable structural analysis in high-dimensional VAR environments.

Disclosure statement

The authors have no conflicts of interest to declare. DeepL (version 1.47.0) was used to assist with grammar and language refinement in selected sections of the manuscript. All final edits and interpretations remain the sole responsibility of the authors. The data used in this study are available from the Federal Reserve Economic Data (FRED) database and from LSEG Data & Analytics.

{\setstretch{0.85} \addcontentsline{toc}{section}{References}

\emptythanks

\newcounter{footnotecounter} \setcounter{footnotecounter}{\value{footnote}} \setcounter{footnote}{0}

titlepage{-3cm} \begin{abstract} \begin{singlespace} This is the online appendix for “Large structural VARs with multiple linear shock and impact inequality restrictions”. The reader is referred to the paper for more detailed information. \end{singlespace} \end{abstract} Keywords: Structural Identification, Sign Restrictions, Narrative Restrictions, Large BVARs JEL classification: C11, C32, C55, E50

\pagenumbering{arabic}