EconBase
← Back to paper

Transmission of Macroeconomic Shocks to Risk Parameters: Their uses in Stress Testing

Extracted main text — title through conclusion, appendix excluded. This is what our citation measures are computed over, published so the extraction can be checked by eye.

76,664 characters · 33 sections · 34 citation commands

Rendered from LaTeX for readability, not typeset faithfully. Citation keys are highlighted; maths is left as source; figures, tables and equation environments are summarised rather than reproduced; unrecognised commands are greyed out so nothing is silently dropped. Email addresses are removed.

Transmission of Macroeconomic Shocks to Risk Parameters: Their uses in Stress Testing

\baselineskip = 5.7mm

\affil{{Institute of Mathematics and Statistics, University of S\ ao Paulo -- IME-USP, Brazil}}

\affil{{Institute of Mathematical and Computer Sciences, University of S\ ao Paulo -- ICMC-USP, Brazil}} \affil[1,2]{{Santander Bank, Brazil}}

abstractIn this paper, we are interested in evaluating the resilience of financial portfolios under extreme economic conditions. Therefore, we use empirical measures to characterize the transmission process of macroeconomic shocks to risk parameters. We propose the use of an extensive family of models, called General Transfer Function Models, which condense well the characteristics of the transmission described by the impact measures. The procedure for estimating the parameters of these models is described employing the Bayesian approach and using the prior information provided by the impact measures. In addition, we illustrate the use of the estimated models from the credit risk data of a portfolio. \\ \\ Keywords: Transmission of Shocks, Stress Testing, Risk Parameters, General Transfer Function Models, Bayesian Approach.

Introduction

Stress testing resurfaced as a key tool for financial supervision after the 2007-2009 crisis; before the crisis these tests were rarely exhaustive and rigorous and financial institutions often considered them only as a regulatory exercise with no impact at capital. The lack of scope and rigor in these tests played a decisive role in the fact that financial institutions were not prepared for the financial crisis. Currently, the regulatory demands on stress tests are strict and the academic interest for related issues has increased considerably. Stress tests were designed to assess the resilience of financial institutions to possible future risks and, in some cases, to help establish policies to promote resilience. Stress tests generally begin with the specification of stress scenarios in macroeconomic terms, such as severe recessions or financial crisis, which are expected to have an adverse impact on banks. A variety of different models are used to estimate the impact of scenarios on banks’ profits and balance sheets. Such impacts are measured through the risk parameters corresponding to the portfolios evaluated. For a more detailed discussion of the use, future and limitations of stress tests, see dent2016stress, SiddiqueHasan2012, ScheuleRoesch2008 and aymanns2018models. For an optimal evaluation of the resilience, it is necessary to implement models that condense the relevant empirical characteristics of the transmission process of the macroeconomic shocks to the risk parameters.

In this work, we study empirical characteristics of the transmission process of macroeconomic shocks to financial risk parameters, and suggest measures that describe the propagation and persistence of impacts in the transmission. We propose the use of an extensive family of models, called General Transfer Function Models, that condense well the characteristics of the transmission described by the impact measurements. In addition, we incorporated to the model the possibility of including a learning or deterioration process of resilience, through a flexible structure which implies a stochastic transfer of shocks. We describe the procedure for estimating model parameters, where the prior information provided by the impact measures is incorporated into the procedure of estimation through Bayesian approach. We also illustrate the use of the impact measures and model, estimated from the credit risk data of a portfolio.

Outline

The paper is organized as follows. In Section (ref) we state the objective of the models in stress testing and present a case study. In Section (ref) we describe the characteristics of the shock transmission process and propose impact measures. Section (ref) describes the General Transfer Function Models. In section (ref) we describe how to incorporate a flexible behavior to the resilience of the risk parameter. Section (ref) describes the procedure for estimating model parameters. In the section (ref) we present the simulation study to evaluate the Response Function dacay and the parameters estimation. Finally in section (ref) we present results of the models applied to a case study.

Stress Test Modeling

Stress testing models are equations that express quantitatively how macroeconomic shocks impact the different risk dimensions of financial institutions. Each risk dimension (for example, credit risk, interest rate risk, market risk, among others) is monitored by parameters, called risk parameters, which quantify the exposure of the portfolio to possible losses. As we mentioned before, the objective in stress tests is to assess the resilience of the portfolios and this is carried out through the risk parameters corresponding to each portfolio. Consequently, a specific objective is to model the behavior of risk parameters in terms of macroeconomic variables and to use these dependency relationships to extrapolate the behavior to the risk parameters in hypothetical downturn scenarios. Stress tests are also done in parameters that monitor the profitability and performance of banks, which is the case of stress tests for Pre-Provision Net Revenue known as PPNR models; all the concepts and tools discussed in this paper are also applicable to this type of models. Portfolios used in stress tests are usually organized with each line of business or following some criterion of homogeneity with commercial utility.

The complex relationship between series, the restricted availability of observations (usually between 20 and 30 quarters of observation), the diffuse behavior of all series (Random Walk) and the need to maintain a simple economic narrative are relevant issues to take into account when proposing models for stress testing.

Case Study: Credit Risk Data

Credit risk has great potential to generate losses on its assets and, therefore, has significant effects on capital adequacy. In addition, the credit risk is, possibly, the dimension of risk with the biggest bank regulation regarding stress tests. The most relevant credit risk parameters to assess resilience are: Probability of Default (PD) and Loss Given Default (LGD). Other risk parameters can be considered, but these have a definition superimposed with the parameters already mentioned. Furthermore, PD and LGD are used explicitly in the calculation of capital for a financial institution.

The definition of default and of the parameters mentioned above is given below:

definitionIn Stress Testing, it is considered to be the default the borrower who does not fulfill the obligations for determined period of time.
definitionProbability of default (PD) is a financial term describing the likelihood of a default over a particular time horizon (this time horizon depends of the financial institution). It provides an estimate of the likelihood that a borrower will be unable to meet its debt obligations. The most intuitive way to estimate PD is through the Frequency of Observed Default (ODF), which is given by number of bad borrower (or borrower in default) in time t and under total number borrower (number of bad and good borrower) in time t.
definitionLoss Given Default or LGD is the share of an asset that is lost if a borrower defaults. Theoretically, LGD is calculated in different ways, but the most popular is 'Gross' LGD, where total losses are divided by Exposure at Default (EAD). Thus, the LGD is the total debt% minus recuperate of the debt%, i.e., the percentage of debt that was recovered by the financial institution with the borrower's payments.

For a better understanding of the different parameters of credit risk and their use to calculate the capital, see HenryKok2013macro.

figure[figure omitted — 551 chars of source]

Without loss of generality, since they can be applied to other risk dimensions, we analyzed the credit risk data of a portfolio from the first quarter of 2008 to the fourth quarter of 2014, observed quarterly. . We considered three macroeconomic variables that represent the macroeconomic environment for this specific portfolio: Gross Domestic Product (GDP), Interbank Deposit Rate (IDR) and Unemployment (Unemp). The macroeconomic variables were observed in the same period as the risk parameters, LGD\footnote{For information security reasons, we used monotonic transformations in the original series.} in this case (see Figure (ref)). We considered four macroeconomic scenarios Optimistic, Baseline, Global and Local, where the Local is more severe than Global, Global is more severe than Baseline and Baseline is more severe than Optimistic (see Appendix (ref)). The severity of the scenarios depends of the regulatory exercise of supervisory authorities and central banks. In our study, the scenarios were generated from the first quarter of 2016 to the fourth quarter of 2021.

Modeling on Efficient Environment

Due to its simplicity, possibly, a first stress test model that could be considered is linear regression for time series. A linear regression will link the macroeconomic variables with the evaluated risk parameter, where the coefficients of the regression will represent the impact of the macroeconomic shocks on the risk parameter. A linear regression establishes a static relationship between the variables, which would imply an immediate transfer of the shocks to the risk parameters. An immediate transfer of shocks would only be possible in a fully efficient economic environment, that is, perfectly rational and well informed. The assumption of an efficient environment, where the agents that determine the behavior of the risk parameter are fully informed and rational agents, is generally far from what is observed in the stress tests data. On the contrary, agents are very focused in the short term and are blind in the long term, which generates collective irrationality and a slow process of adaptation to shocks.

Using a static relationship, a linear regression in this case, and without taking into account the inefficiency of the environment; it would generate non-coherent forecasts of scenarios (overlapping scenarios forecasts) (see Appendix (ref) for more results). Overlapping scenarios would constitute evidences that the model may not be adequate, given that the model would be making predictions inconsistent with stress levels, which would generate a relevant source of risk. It is worth mentioning that, in some cases, in environments that do not deviate much from the ideal conditions of an efficient environment, some transformations of the variables can be applied to comply with the assumptions of a linear regression model, but in general, these transformations eliminate relevant information at transmission process of shocks that a good model should contain.

Transmission of Shocks

Propagation and Persistence

To understand the transmission process of macroeconomic shocks to risk parameters, it is necessary to consider the characteristics and limitations of the agents that determine the dynamic behavior of the risk parameters. The series of risk parameters observed, derive from the relationship established between imperfect agents (in terms of rationality), which interact within an imperfect economic environment (in terms of information). For example, the agents that determine the dynamic behavior of the credit risk parameters are: the financial institution that grants the credits, the clients whom the credit is granted to and the financial regulator. All these agents have relevant limitations of rationality and information.

The imperfection of the economic environment conditions determines a slow processing of information by agents (behavioral or institutional inertia) and at the same time causes an unequal distribution of information among agents (asymmetry of information). The behavioral inertia and information asymmetry cause that macroeconomic shocks dynamically impact the risk parameters, that is, the impacts are propagated in time. The decay in the propagation of the impacts can be fast or slow; when the decay is slow we say that the propagation is of high persistence; and when the decay is fast, then the propagation is low. Thus, in inefficient environments the transmission process is characterized by persistence in the propagation. The nature and intensity of the persistence in the data of a specific portfolio need to be evaluated empirically and corroborated with some coherent economic criteria.

Impact Measures

To empirically contrast the persistent nature of the impacts, we propose the use of two simple functions (Response and Diffusion):

i) Response Function

A simple way to assess the impact of macroeconomic shocks on the average level of risk parameters is through what we call Response Function $\mathfrak{R}(j)$ that is defined as:

equation[equation omitted — 128 chars of source]

where $Y_{t}$ is the stochastic process generator of risk parameter with mean function $\mu_{Y}=\mathbb{E}[Y_{t}]$ and $X_{t}$ is the stochastic process generator of a univariate macroeconomic series with mean function $\mu_{X}=\mathbb{E}[X_{t}]$. To each $j$, the quantity $\mathfrak{R}(j)$ measures the mean impact on $Y$, a time $j$ later, of a shock in $X$ at time $t$. The time window used to obtain the estimation of $\mathfrak{R}(j)$ generally comprises the entire length of the series, in this case, in statistics and probability, $\mathfrak{R}(j)$ it is referred to as Cross-Covariance Function (CCF). In some specific cases, it may happen that the macroeconomic series and the risk parameter present divergent relationships from those expected economically. These divergences occur only temporarily due to political changes in the portfolio. In the presence of these divergences, the temporal window to estimate $\mathfrak{R}(j)$ has to be chosen trying to maintain the economic relationships that we know priori that make sense. To be more recurrent, in this work we will use the whole length of the series.

Given the bivariate series $\{Y_{t},X_{t}\}_{t=1}^{T}$, the consistent estimator of $\mathfrak{R}(j)$ (denoted by $\widehat{\mathfrak{R}}$(j)), is given by

equation[equation omitted — 104 chars of source]

where $\bar{Y}$ and $\bar{X}$ are the sample mean values of both series observed respectively (see box2008time).

It is known from the literature that the CCF is not recommended to be used in the model specification when the series are not stationary (which is the case of the series used in Stress Testing). On the other hand, it is worth mentioning that $\widehat{\mathfrak{R}}(j)$, in this work, is not used to specify the model, but to remove qualitative information from the decay of shocks, that is, the functional form of the decays that representing the endogenous behavior of the parameter to macroeconomic shocks.

ii) Diffusion Function

The Diffusion Function $\mathfrak{D}(j)$ is defined as:

equation[equation omitted — 124 chars of source]

where $Y_{t}$ is the stochastic process generator of risk parameter with mean function $\mu_{Y}=\mathbb{E}[Y_{t}]$ and $X_{t}$ is the stochastic process generator of macroeconomic series with mean function $\mu_{X}=\mathbb{E}[X_{t}]$. To each $j$, the quantity $\mathfrak{D}(j)$ measures the average fluctuation of $Y$ between time $t$ and $t+j$, associated to shocks on $X$ at time $t$. Note that $\mathfrak{D}(j)$ is used to analyze the shocks impact on the variance of the risk parameter. If there is a persistent pattern in this measure, it should be considered in construction from candidates models.

Given the series $\{Y_{t},X_{t}\}_{t=1}^{T}$, the consistent estimator of $\mathfrak{D}(j)$ (denoted by $\widehat{\mathfrak{D}}$(j)), is given by

equation[equation omitted — 106 chars of source]

where $\bar{Y}$ and $\bar{X}$ are sample mean values.

Interpretation of Impact Measures

The impact measurements presented in ((ref)) and ((ref)) describe the dynamic response of the risk parameter $Y$ to shocks in the macroeconomic variable $X$. In other words, they describe the way risk parameter absorbs the shocks, which is determined by the reaction of the agents to return to equilibrium. A slow decrease of the impact measures implies a high persistence at the absorption of shocks. Through these impact measures, two types of impact can be identified: Impacts with permanent effects and impacts with temporary effects. Shocks with permanent effects impact the stochastic trend of the risk parameter and their effects are measured by $\mathfrak{R}(j)$. Shocks with temporary effects impact the stationary fluctuations around the stochastic trend and their effects are measured by $\widehat{\mathfrak{D}}(j)$. For example, we can empirically identify these two types of impacts with the impact measurements referring to the credit risk data, see Figure (ref).

figure[figure omitted — 807 chars of source]

The persistent nature of the shocks is observed in the decay rate of the impacts measured by $\widehat{\mathfrak{R}}(j)$. It is important to mention that due to the aggregate nature of the macroeconomic variables, the type of persistence needs to be contrasted with economic criteria that justify its nature. For example in the case of GDP and IDR, observing their corresponding Response Functions, see Figure (ref), and considering that both macroeconomic variables are directly involved in the credit recovery policies and LGD calculation respectively, we can consider that the impact of the shocks of these two macroeconomic variables on the LGD of this portfolio are of low persistence. In the case of the macroeconomic variable Unemployment, its response function shows slow decay in the first lags, besides, it is a variable associated with labor laws and institutional policies that justify the high persistence of shocks in this macroeconomic variable. In this case, $ \widehat{\mathfrak{D}}(j) $ not present a persistent pattern, that is, the shocks do not affect the variance of the risk parameter of persistence way, the only effect that the macroeconomic shocks do is to balance around their averages.

Transfer Function Models

We are interested in the transmission of shocks in macroeconomic variables, for example $ X$, to risk parameters, for example $Y$. In the previous section, we showed that the $ X $ effect on $ Y $ persists for a period and decays to zero as time passes. A simplified and didactic way to represent the cumulative impact of $X_{t}$ on $Y_{t}$ is

equation[equation omitted — 79 chars of source]

where any shock in $X_{t}$ will impact $\mathbb{E}[Y_{t}]$ $(E_{t})$ in all the later periods. The term $\beta(j)$ in ($\ref{TFM}$) is a function of $j$ that represents the $j$-th impact coefficient. A condition generally assumed for dynamic stability of system (all finite shock has a finite cumulative impact), is that $\sum_{j=0}^{\infty}\beta(j)<\infty$, which implies $\lim_{j\to\infty}\beta(j)=0$. It is important to mention that in the last condition nor $X_{t}$ neither $Y_{t}$ must to be stationary. But the condition establish is the stability of the responses of the parameter risk to the macroeconomic shocks, which is a recurring characteristic in ergodic systems.

It is common to establish specific restrictions on the impact coefficients functionally related, often known as Distributed Lag Models Zellner1971introduction. The principal Distributed Lag Models are described in Koyck1954distributed, Solow1960family and Almon1965distributed.

In this work we adopted a more general perspective described by BoxJenkins2015time. Alternatively, the impact of $X_{t}$ can be written as a linear filter:

equation[equation omitted — 141 chars of source]

The polynomial on the lag operator $\mathcal{T}(L)=\beta(0)+\beta(1)L+\beta(2)L^{2}+\ldots$ is called Transfer Function and represents the cumulative impact of $X$ on $Y$. The sequence $\{\beta(j)\}_{j=0}^{\infty}$ called Dynamic Multiplier is the mechanism of propagation of impacts underlying the Transfer Function and expresses the instantaneous impact of $X$ on $Y$ at present and further times.

General Transfer Function Model

Let $X_{t}$ be the value of an macroeconomic, scalar variable $X$ and $Y_{t}$ the value of risk parameter at time $t$. A General Transfer Function Model (GTFM) for the dynamic effect of $X$ on the risk parameter $Y$, consistent with the empirical facts discussed in section ((ref)), is defined as

subequations\begin{align} Y_{t} &=\mathbf{F}_{t}^{\top}\mathbf{\theta}_{t}+v_{t}\\ \theta_{t} &=\mathbf{G}_{t}\theta_{t-1}+\psi_{t}X_{t}+\partial\theta_{t}\\ \psi_{t} &=\psi_{t-1}+\partial\psi_{t} \end{align}

with terms described as follows: $\theta_{t}$ is an $\mathit{n}$-dimensional states vector (unobserved), $\mathbf{F}_{t}$ is a known $\mathit{n}$-vector (design vector), $\mathbf{G}_{t}$ a known transition matrix (evolution matrix) which determines the evolution of the states vector and, $v_{t}$ and $\partial\theta_{t}$ are observation and evolution noise terms. All these terms are precisely as in the standard Dynamic Linear Models (DLM), with the usual independence assumptions for the noise terms hold here. The term $\psi_{t}$ is an $\mathit{n}$-vector of parameters, evolving via addition of a noise term $\partial\psi_{t}$, assumed to be zero-mean normally distributed independently of $v_{t}$ (though not necessarily of $\partial\theta_{t}$).

In accordance with the empirical facts discussed in section ((ref)), the GTFM defined in ((ref)) establishes that $X$ impacts the stochastic trend of $Y$ ((ref)) and, at the same time, generates stationary fluctuations around that stochastic trend ((ref)). The parameters that determine the transfer can vary over time, allowing greater flexibility to the transmission process of shocks ((ref)). This model constitutes a small variation of the first model presented by HarissonWest1999.

The states vector $\theta_{t}$ carries the effect of current and past values of the $X$ series to $Y_{t}$ in equation ((ref)); this is formed in ((ref)) as the sum of a linear function of past effects; $\theta_{t-1}$, and the current effect $\psi_{t}X_{t}$, plus a noise term. Notice that the parameterization ((ref)) is more flexible than that of equation ((ref)), allowing different stochastic interpretations for the effect of $X$ on $Y$.

The general model ((ref)) can be rewritten in the standard DLM as follows. Define a 2$\mathit{n}$-dimensional state parameters vector $\tilde{\theta}_{t}$ by concatenating $\theta_{t}$ and $\psi_{t}$, giving $\tilde{\theta}_{t}=(\theta_{t}^{\top}, \psi_{t}^{\top})$. Similarly, extend the $\mathbf{F}_{t}$ vector by concatenating an $\mathit{n}$-vector of zeros, giving a new $\tilde{\mathbf{F}}_{t}$ such that $\tilde{\mathbf{F}}_{t}=(\mathbf{F}_{t}, 0, \cdots, 0)$. For the evolution matrix, define by

align*[align* omitted — 154 chars of source]

where $\mathbf{I}_{\mathit{n}}$ is the $\mathit{n}$ x $\mathit{n}$ identity matrix. Finally, let $\mathbf{\omega}_{t}$ be the noise vector defined by $\omega_{t}^{\top}=(\partial\theta_{t}^{\top}+X_{t}\partial\psi_{t}^{\top}, \partial\psi_{t}^{\top})$. Then the model ((ref)) can be written as

equation[equation omitted — 197 chars of source]

Thus the GTFM ((ref)) is written in the standard DLM form ((ref)) and the usual inferential analysis corresponding to State Space Models can be applied (see DurbinKoopman2012time and WestHarrison2006bayesian). Note that the generalization to include several macroeconomic variables in the model ((ref)) is trivial.

Some Transfer Functions

In this section, we present some Transfer Functions underlying models that are specific cases of the GTFM.

Geometric Propagation

Assuming a Dynamic Multiplier with geometric decay in ((ref)), this is:

equation[equation omitted — 158 chars of source]

the mechanism of propagation ((ref)) was first associated to Koyck1954distributed, and postulates the gradual and rapid decay of the impacts and generates the following Transfer Function

equation[equation omitted — 69 chars of source]

replacing ((ref)) in ((ref)), the Transfer Function model would be

subequations\begin{align} Y_{t}&=E_{t}+v_{t}\\ E_{t}&=\phi E_{t-1}+\beta X_{t}+\partial E_{t}, \end{align}

where

align*[align* omitted — 103 chars of source]

with $v_{t}$ ($\partial E_{t}$) normally distributed with zero mean and variance $\sigma_{v}^{2}$ ($\sigma_{E}^{2}$), $\eta=\frac{\sigma_{v}}{\sigma_{E}}$ is signal-to-noise ratio. In this model, the dynamic response of $Y_{t}$ is $\beta \phi^j X_{t-j}$. Based on the representation in ((ref)), we have $\mathit{m}=1$ (model with only one latent state), $\theta_{t}=E_{t}$, the cumulative impact until $t$, $\psi_{t}=\beta$, the current impact for all $t$, $\mathbf{F}_{t}=1$, $\mathbf{G}_{t}=\phi$ and $\partial \theta_{t}=\partial E_{t}$. Two simplifications of the equation ((ref)) can be obtained assuming $\eta=0$ or $\sigma_{E}=0$, these simplifications are often known as Autoregressive Distributed Lag (ADL) model. For a detailed review of ADL models from a classical approach, see PesaranShin1998ARDL, and from a Bayesian approach, see BauwensLubrano1999bayesian, Zellner1971introduction.

figure[figure omitted — 906 chars of source]

Note that the condition $|\phi|<1$ in ((ref)), corresponds to the stability condition mentioned in ((ref)), in which case the Dynamic Multiplier may exhibit the following patterns: (i) If $0<\phi<1$, shows monotone geometric decay and (ii) If $-1<\phi<0$, shows geometric decay with alternating signs. The parameter $\phi$ (resilience parameter) controls the velocity of the decay; absolute values next to 1 imply high persistent of impacts. Such patterns are exhibited in the Figure (ref), with $\beta >0$ (positive relationship) and different values for $\phi$, for $\beta<0$ (negative relationship) the patterns are similar but the decay has inverted sign.

An extension of Koyck Distributed Lag ((ref)) was proposed by Solow1960family, generally known as Solow Distributed Lag. Solow's assumption about the $\beta(j)$ is that they are generated by a Pascal distribution, which generates a more flexible transfer. An even more flexible proposal is known as Almon Distributed Lag, proposed by Almon1965distributed. Almon's assumption is that the coefficients are well approximated by a low order polynomial on the lags, this polynomial approximation provides a wide variety of shapes for the lag function $\beta(j)$. With an adequate reparameterization and identification of the corresponding Transfer Function, both the Koyck, Solow and the Almon proposals can be expressed as a particular case of the GTFM (see, RavinesSchmidtMingon2006revisiting).

Propagation of High Persistence

The simplification proposed by Koyck implies a rapid decay of the impacts. However, many times the shock impact of some macroeconomic variables may present a high persistence, for example, the Unemployment variable referring to the credit risk data in section ((ref)). An alternative to model highly persistent shocks is to choose Transfer Functions corresponding to either the Solow's or Almon's simplifications presented in the previous section. A disadvantage of these alternatives is the need to observe high persistence in all the shocks of the macroeconomic variables considered. But in general, as seen in credit risk data, high persistence is empirically observed only in some of the total macroeconomic variables considered, and in many cases only in one of them. Then, it is necessary to generate a flexible propagation mechanism only for the variables identified with high persistence. A simple way to approach this problem is to propose a general geometric decay for all the variables and generate a superposition effect of lag decays for the variables identified with high persistence.

For example, suppose that $X$ is a variable identified with high persistence; then we assume the following superposition of decays as Dynamic Multiplier

equation[equation omitted — 181 chars of source]

where $|\phi|<1$ and $\beta_{k} \in \mathbb{R}$ for all $k \in \{0,1,\ldots,s\}$ and $\mathbb{I}_{\{.\}}$ is the indicator function. Note that $s$ is an arbitrary number that determines the number of lags used in the superposition (superposition-order), which is calibrated in function of intensity of the persistence in the data. Every lag has geometric decay but generates interference in the decay of their predecessor lags, see Figure (ref).

figure[figure omitted — 566 chars of source]

The propagation defined in ((ref)) generates the following Transfer Function

equation[equation omitted — 76 chars of source]

then, the Transfer Function model is

subequations\begin{align} Y_{t}&=E_{t}+v_{t}\\ E_{t}&=\phi E_{t-1}+\beta_{0} X_{t}+\beta_{1} X_{t-1}+\ldots + \beta_{s} X_{t-s}+\partial E_{t} \end{align}

with $v_{t}$ ($\partial E_{t}$) normally distributed with zero mean and variance $\sigma_{v}^{2}$ ($\sigma_{E}^{2}$), $\eta=\frac{\sigma_{v}}{\sigma_{E}}$ is signal-to-noise ratio. Same as in the model ((ref)), the model ((ref)) can be easily written using the general representation ((ref)).

Choice of Transfer Function

Impact measures discussed in section (ref) represent the empirical propagation of shocks. We will to use this qualitative information embedded in the impact measures to choose the Transfer Function. The functional form of the theoretical decay of the impacts (Dynamic Multiplier) must be consistent with the empirical propagation of the permanent impacts. Then, a criterion that we employed to choose the functional form of the Dynamic Multiplier of a macroeconomic variable, was that the functional form be consistent with the rapidity of decay of its Response Function $\widehat{\mathfrak{R}}(j)$. For example, if the Response Function $\widehat{\mathfrak{R}}(j)$ of a macroeconomic variable shows a high persistence or a behavior with wave decay, then the functional form proposed for its Dynamic Multiplier has to imitate that pattern, because it condenses the empirical shocks transmission.

It is important, as seen above, to point out that these measures are not used to specify the model with precision, as in the case of the autocorrelation function in the modeling of stationary series. Rather, due to the aggregate nature and the diffuse behavior (random walk) of the macroeconomic variables and the risk parameter, these measures are used to have a general approximation of the functional form of the impact decay that characterizes the transmission process.

Time-varying Resilience and Stochastic Transfer

In the transfer functions discussed in the previous section the Resilience Parameter is constant over time, this implies a state of equilibrium between the shocks and the response of the risk parameter to these shocks, which remains constant. A constant resilience does not consider the learning process or deterioration of the agents' response, and therefore nor of the response of the risk parameter to recurrent macroeconomic shocks. It may be important to identify and incorporate this learning process or deterioration of the resilience of the parameter into the modeling, in order not to overestimate or underestimate the projections in the stress scenarios.

To incorporate a dynamic behavior to resilience, we assume a mechanism of stochastic propagation:

subequations\begin{align} \beta_{t}(j)&=\beta\prod_{i=0}^{j-1}\phi_{t-i}\\ \phi_{t}&=\phi_{t-1}+\partial\phi_{t} \end{align}

where $\beta \in \mathbb{R}$ and $\partial \phi_{t}$ are normally distributed with zero mean and variance $\sigma_{\phi}^{2}$. Note that the resilience changes slowly over time according to the simple random walk ((ref)).

The Stochastic Dynamic Multiplier in ((ref)) generates the following Stochastic Transfer Function

subequations\begin{align} \mathcal{T}_{t}(L)&=\beta\sum_{j=0}^{\infty}\prod_{i=0}^{j-1}\phi_{t-i}L^{j} \\ \phi_{t}&=\phi_{t-1}+\partial\phi_{t} \end{align}
figure[figure omitted — 564 chars of source]

Then, the Transfer Function model is

subequations\begin{align} Y_{t}&=E_{t}+v_{t}\\ E_{t}&=\phi_{t} E_{t-1}+\beta X_{t}+\partial E_{t}\\ \phi_{t}&=\phi_{t-1}+\partial\phi_{t} \end{align}

with $v_{t}$ ($\partial E_{t}$) normally distributed with zero mean and variance $\sigma_{v}^{2}$ ($\sigma_{E}^{2}$), $\eta=\frac{\sigma_{v}}{\sigma_{E}}$ signal-to-noise ratio. Like all models previously presented, the model ((ref)) can be easily written using the general representation ((ref)).

Equation ((ref)) also establishes a relationship of equilibrium between shocks and the response of the risk parameter, but that equilibrium is sensitive and changes before variations occur in the macroeconomic environment, recurrent characteristic of metastable systems. The flexible structure for the risk parameter makes it possible to identify a learning effect or deterioration of resilience over time. In Figure (ref) is shown an example of deterioration of resilience.

Inference Procedure

Macroeconomic variables are usually correlated, this could significantly compromise the estimation of the parameters of the models described in this paper. Another problem emerges when it is necessary to model shocks with high persistence through the method of superposition of lags proposed in the section ((ref)), implying moderate to large values of $s$ (superposition-order), in which case the estimation of the parameters of the model, when we use the classic inference, is compromised by the autocorrelation in $X_{t}, \ldots ,X_{t-s}$. For these reasons we recommend the use of Bayesian inference whenever possible. Note that both problems could cause that the Response Functions and the Dynamic Multipliers do not have the same decay patterns. We propose to estimate the parameters of these models using the qualitative information contained in the Impact Measures (decay pattern of Response Functions), incorporating them into the prior using estimation procedure through Bayesian approach.

Prior Specification

Following the Bayesian paradigm, the specification of a model is complete after specifying the prior distribution of all parameters of interest. To complete the specification of the models described, we defined the prior distribution according to the declination pattern of the Response Function.

For the Resilience Parameter $\phi \in (0,1)$, that corresponds to Response functions with monotone decay pattern, we specified a $\mathcal{B}eta(a,b)$ distribution; and for $\phi \in (-1,0)$, that corresponds to Response functions with wave decay pattern, it is specified by setting $\phi=-\phi^{*}$ where $\phi^{*}\sim\mathcal{B}eta(c,d)$. We chose to work with the normal distribution for the regressors parameters, however, we decided to restrict the support according to the Response Function observed at the data, that is, considering the dynamic relationship between the risk parameter and the macroeconomic series.

itemize• If the relation present by the Response Function of the $j$-th macroeconomic variable is positive, \begin{align*} \beta_{j}\simHalf-Normal(\mu_{0}, \sigma_{0}^{2}) with \beta_{j} \in [0,\infty). \end{align*} • If the relation of the Response Function of the $j$-th macroeconomic variable is negative, \begin{align*} \beta_{j}\simHalf-Normal(\mu_{0}, \sigma_{0}^{2}) with \beta_{j} \in (-\infty,0]. \end{align*}

Finally, the prior distribution for $\sigma_{E}$, $\eta$ and $\sigma_{\phi}$ are specified as,

align*[align* omitted — 179 chars of source]

MCMC-based Computation

There are several approximation methods based on simulations to obtain samples of the posterior distribution. The most widely used is the Markov Chain Monte Carlo Methods (MCMC), described in RobertCasella2010MCMC. Within the MCMC family stand out the algorithms: Metropolis metropolis1953equation, Metropolis-Hastings Hastings1970monte, Gibbs Sampling GelfandSmith1990Gibbs and Hamiltonian Monte Carlo Neal2011hmc. Great advances have been made recently with Hamiltonian Monte Carlo algorithms (HMC, see for example Girolami2011AdvanceHMC) for the estimation Bayesian models. The HMC has an advantage over the other algorithms mentioned, since it avoids the random walk behavior, presents smaller correlations structure and converges with less simulations. In this paper we used the HMC to obtain samples of the posterior distribution.

Model Comparison

One of the challenges in the modeling process is choosing the model that best represents the data structure and that is parsimonious. A widely used metric, in the case of Bayesian approach, is the Deviance Information Criterion (DIC) proposed by spiegelhalter2002dic, which is defined as,

align*[align* omitted — 92 chars of source]

where $\mathbf{Y}=(Y_{1},\dots ,Y_{T})$ are the data, $\theta$ is the vector of parameters of the model, $\hat{\theta}_{Bayes} = E[\theta|\mathbf{Y}]$ and $\textrm{p}_{DIC}$ is a penalizing term given by,

align*[align* omitted — 147 chars of source]

Then, given M simulations from the posterior distribution of $\theta$, $p_{DIC}$ can be approximated as,

align*[align* omitted — 168 chars of source]

Additionally, in this paper, we used the Watanabe Information Criterion (WAIC) described by watanabe2010waic. The WAIC does not depend on Fisher's asymptotic theory; therefore, it is not responsible for later displacement to a single point, as well as it is an alternative to more advanced models (see watanabe2013widely, vehtari2015efficient and vehtari2017practical). It is defined as,

equation[equation omitted — 89 chars of source]

where $\textrm{p}_{WAIC} = \sum_{i=1}^{T}V_{\theta|\mathbf{Y}}\big(\log p(Y_{i}|\theta)\big)$ penalizes for the effective number of parameters and $\textrm{lpd} = \sum_{i=1}^{T}\log p(Y_{i}|\mathbf{Y})$ is the log pointwise predictive density, as defined in watanabe2010waic, where,

align*[align* omitted — 105 chars of source]

Similar to DIC, given M simulated from posterior distribution, the two terms in ((ref)) are estimated as,

align*[align* omitted — 107 chars of source]

and

align*[align* omitted — 93 chars of source]

where $V_{m=1}^{M}$ denotes the sample variances of log $p(Y_{i}|\theta^{(1)}), \ldots, p(Y_{i}|\theta^{(M)})$. Thus, WAIC is defined as,

align*[align* omitted — 46 chars of source]

The log pointwise predictive density can also be estimated using approximate leave-one-out cross-validation (LOOIC) as,

align*[align* omitted — 148 chars of source]

where $Y_{-i}$ denotes the data vector with the $i$th observation deleted. vehtari2015efficient introduced an efficient approach to compute LOOIC using Pareto-Smoothed Importance Sampling (PSIS). The lower the value of the selection criteria, the better the model is considered. It is worth mentioning that the estimates for WAIC and LOOIC are obtained as the sum of independent components, so it is possible to calculate the approximate standard errors for the estimated predictive errors.

Forecasting with Posterior Distribution

Without loss of generality, we will assume a univariate case (only a macroeconomic variable). We want to apply our model analysis to a new set of data, where we observed the new covariate $X_{T+1}$ corresponding to one of the stress scenarios, and we wish to predict the corresponding outcome $Y_{T+1}$. From a Bayesian perspective, this process can be performed considering the posterior predictive distribution.

equation[equation omitted — 93 chars of source]

where $D_{T}$ is all the information until the instant $T$ and $\gamma = (E_{T+1}, \sigma_{E})^{\top}$. However, since there is no analytical determination of this distribution, we adopted the following procedure.

align*[align* omitted — 283 chars of source]

where $m = 1, \ldots, M$ are MCMC samples of the posterior distribution. Then, $\Big(Y_{t+1}^{(1)}, \ldots, Y_{t+1}^{(M)}\Big)$ are an i.i.d. sample from $p(Y_{T+1}|D_{T})$. The process can be repeated until $h$-steps-ahead, considering $X_{T+2},\ldots,X_{T+h}$, to obtain the forecast to $Y_{T+2},\ldots,Y_{T+h}$.

Simulations Study

In this section , the theoretical results of the previous sections are illustrated through the simulations study. First we will provide the details of the process of simulating of the data series. We will explain how evaluating the simulations and to comment about the results obtained.

Data Generating Process. The simulation process of data series was done follow the steps below:

itemize• Set the number of observations ($T$) in the series you will simulated; • Generate two different random walk processes $X_{1}$ and $X_{2}$ of size $T$, which in this case represent the macroeconomic variables; • Set the values of the parameters that will generate the response variable $Y$, which in this case represents the risk parameter; • Finally $Y_{t}\sim\mathcal{N}(\alpha+\phi y_{t-1}+\beta_{1}x_{1t}+\beta_{2}x_{2t}, \sigma_{v}^{2})$ of size $T$.

As in general the series used in the Stress Test exercises are short, we consider using $T = 40$ and we generated $M = 500$ replicates of series of this size. We divided the simulation study into two steps. In the first stage we check the functional form of the decays presented by the simulated series and calculate the hit rate of the response functions. For the first step we used several configurations, as shown in Table (ref). In the second step, we used two of the configurations presented in step 1, performed the estimation of the N series and evaluated some metrics in relation to the estimates, they are: General mean, general deviation, Mean Squared Error (MSE) and Median Absolute Error (MAE).

To calculate the hit rate of the response function, we consider the following process. Be $\Delta\widehat{\mathfrak{R}}(j) = \widehat{\mathfrak{R}}(j) - \widehat{\mathfrak{R}}(j-1)$, where $j=0,\ldots, L$, consider the mean dacay below:

align*[align* omitted — 98 chars of source]

where $\widehat{\mathfrak{R}}(j)^{(m)}$ is the estimative of the response function of $m$-th iteration of the $M$ simulations. The Hit Rate of the Response Function is given by:

align*[align* omitted — 101 chars of source]

We used numbers of lags $L=10$.

Results

In this section we present the results of the two steps of the simulation study.

table[table omitted — 1,092 chars of source]

According the Table (ref), the hit rate of the response function is satisfactory. With this we can be use the response function it to evaluate the decay of the shocks of the macroeconomic variables. Note that when the value of one of the betas is double the other, its decay hit rate is higher than the which one with of the lowest weight, but yet is satisfactory to evaluate the decay.

To the step two we evaluate the estimations considering two configurations and compare the Response Function Mean of the estimates with the Expected decay from of each variables.

The configuration 1 we considered is: $\alpha=3$, $\phi=0.4$, $\beta_{1}=-0.4$, $\beta_{2}=0.4$ and $\sigma_{v}=0.1$.

table[table omitted — 507 chars of source]

As present in the Table (ref), the estimates obtained are close to the true values of the parameters. Also we observated that the standard deviation of $\alpha$ and the difference in scale of the $\sigma$ parameter is slightly higher than the others presented.

figure[figure omitted — 847 chars of source]

In the Figure (ref) we presented the Reponse Function Mean of the $M=500$ simulations and present the Expected decay of the variables used to construct the data series. As shown the average empirical functional form of decay is in accordance with the expected form of theoretical decay of the variables.

The configuration 1 we considered is: $\alpha=3$, $\phi=0.7$, $\beta_{1}=-0.8$, $\beta_{2}=0.4$ and $\sigma_{v}=0.01$.

table[table omitted — 503 chars of source]

Like in configuration 1, the the estimates obtained in configuration 2, showing in the Table (ref), are close to the true values of the parameters. We also note that the standard deviation of all estimates has decreased considerably. We believe it is due to the resilience parameter.

figure[figure omitted — 847 chars of source]

As observed in both configuration the functional form of average decay that is the qualitative information that will be used to build the priori, present in Figures (ref) and (ref) , are in agreement with the form of theoretical decay already defined by construction of the data series. It is worth mentioning that, since Bayesian inference was used, the estimates presented are based on the mean and standard deviation of a posteriori means.

Application: Case study

In this section, we illustrate the use of the proposed model in credit risk data that was presented in section (ref). We consider the same macroeconomic series: GDP, IDR and Unemployment. It is noteworthy that the Unemployment series shocks on risk parameter LGD are of persistent nature, since the empirical measures of impact show a slow decay, as presented in section (ref) and there are economic arguments to corroborate this fact; therefore, we considered a superposition with three lags (order 3) for the Unemployment variable and geometric decay for GDP and IDR. We also considered the use of the four scenarios mentioned before, being: Baseline, Optimistic, Local and Global.

In order to demonstrate the flexibility of general model, we tested some variations of the model; these were denominated models I, II, III and IV. For each model we drew 10,000 MCMC samples for the parameters using HMC with Stan and NUTS where 5,000 were discarded as warmup, as detailed below.

Model I

The Model I can be represented as,

align[align omitted — 335 chars of source]

where, $\eta=0$, $X_{t}^{\top}=(1, GDP_{t}, IDR_{t}, Unemp_{t}, Unemp_{t-1}, Unemp_{t-2}, Unemp_{t-3} )$ and $\beta=(\alpha, \beta_{1}, \ldots, \beta_{6})^{\top}$. We defined the following prior distributions,

align*[align* omitted — 651 chars of source]

The results of this model are presented below,

table[table omitted — 977 chars of source]
figure[figure omitted — 450 chars of source]
figure[figure omitted — 157 chars of source]

Model II

The Model II can be represented as,

align[align omitted — 324 chars of source]

where $X_{t}^{\top}=(1, GDP_{t}, IDR_{t}, Unemp_{t}, Unemp_{t-1}, Unemp_{t-2}, Unemp_{t-3} )$ and $\beta=(\alpha, \beta_{1}, \ldots, \beta_{6})^{\top}$. We defined the following prior distributions,

align*[align* omitted — 752 chars of source]

The results of this model are presented below,

table[table omitted — 1,054 chars of source]
figure[figure omitted — 451 chars of source]
figure[figure omitted — 154 chars of source]

Model III

The Model III can be represented as,

align[align omitted — 511 chars of source]

where, $\eta=0$, $X_{t}^{\top}=(1, GDP_{t}, IDR_{t}, Unemp_{t}, Unemp_{t-1}, Unemp_{t-2}, Unemp_{t-3} )$ and $\beta=(\alpha, \beta_{1}, \ldots, \beta_{6})^{\top}$. Note that the resilience parameter varies over time according to random walk process; generating stochastic transmissions. We defined the following prior distributions,

align*[align* omitted — 682 chars of source]

The results of this model are presented below,

table[table omitted — 1,068 chars of source]
figure[figure omitted — 905 chars of source]
figure[figure omitted — 170 chars of source]
figure[figure omitted — 159 chars of source]

Model IV

The Model IV can be represented as,

align[align omitted — 509 chars of source]

where $X_{t}^{\top}=(1, GDP_{t}, IDR_{t}, Unemp_{t}, Unemp_{t-1}, Unemp_{t-2}, Unemp_{t-3} )$ and $\beta=(\alpha, \beta_{1}, \ldots, \beta_{6})^{\top}$. Note that the resilience parameter varies over time according to random walk process; generating stochastic transmissions. We defined the following prior distributions,

align*[align* omitted — 783 chars of source]

The results of this model are presented below,

table[table omitted — 1,143 chars of source]
figure[figure omitted — 904 chars of source]
figure[figure omitted — 170 chars of source]
figure[figure omitted — 162 chars of source]

Model Comparison Results

The four models proposed for the credit risk data condense the empirical propagation of the shocks, described by the impact measurements in section ((ref)). It is important to highlight that they also reproduce satisfactorily the persistence nature of the shocks of the variable Unemployment. That is, all the proposed models summarize well the transmission process of macroeconomic shocks underlying the data. As an effect of condensing well the transmissions of empirical shocks, the models present good adjustment capacity in the response variable (see the Table (ref)). In addition, stationarity and all other assumptions of the models are attended. For a correct interpretation of adjustment capacity measures see Appendix (ref).

table[table omitted — 346 chars of source]
table[table omitted — 421 chars of source]

Although the four models are empirically consistent and with good performance, models III and IV are more complete because they consider a flexible structure for the resilience parameter. In both cases, the resilience of the risk parameter improves over time (see Figures (ref) and (ref)). That is, during the observation period, the risk parameter improved its response and adaptation to shocks. Given that in the credit risk data there is a significant learning process of adaptation to shocks, the information criteria tend to choose models III and IV (see Table (ref)).

The four models generated coherent forecasts (non-overlapping scenarios forecast). An advantage of the last two models is that by incorporating the learning process by means of a flexible structure for the resilience parameter, they generated more optimistic projections for all scenarios as an effect of the improvement in the underlying resilience in the data. Note that if there was a deterioration on resilience, the effect on the projections would be the opposite. And, of course, an advantage of models I and II is that with a simpler structure they manage to reproduce the transmission of shocks and generate coherent projections in the scenarios.

Conclusions and Future work

We showed that in inefficient economic environments, the transmission of macroeconomic shocks to risk parameters, in general, is characterized by a persistent propagation of impacts. To identify the intensity of this persistence, we proposed the use of impact measures that describe the dynamic relationship between the risk parameter and the macroeconomic series. We presented a family of models called General Transfer Function Models, and considered the use of impact measures to propose specific models, that satisfactorily synthesize the relevant characteristics about the transmission process.

We present a simulation study to evaluate the parameter estimation process and the functional form of the Response Function decays. Moreover, from the empirical data of a credit risk portfolio, we show that these models, with an adequate identification of the transfer functions through the impact measures, satisfactorily describe the underlying data transmission process and at the same time generate coherent projections in the scenarios of stress.

These models can be extended to portfolios where there is a correlation between their risk parameters, eg LGD and PD, including this correlation and allowing simultaneous modeling. It is also possible to extend these models to include a dependence between the resilience and the severity of the stress scenario, that is, the response of the parameter would depend on the intensity of the shocks corresponding to the scenario. We intend to explore these extensions in future work.

Acknowledgments

The authors thank Luciana Guardia, Tatiana Yamanouchi and Rafael Veiga Pocai for the fruitful discussions and contributions to this research. This research was supported by Santander Brazil.

The vision presented in this article is entirely attributed to the authors, not being the responsibility of Santander S.A.