Extracted main text — title through conclusion, appendix excluded. This is what our citation measures are computed over, published so the extraction can be checked by eye.
19,117 characters · 2 sections · 34 citation commands
Testing Partial Instrument Monotonicity
Keywords: Partial monotonicity, instrument validity, nonparametric test
Exclusion, random assignment, and monotonicity are three fundamental conditions for instrument variables (IVs) to be valid in the identification and estimation of causal effects. When the instrument variable is multi-dimensional, mogstad2021causal show that the monotonicity condition only holds if choice behavior is effectively homogeneous. However, such homogeneity may break down in many empirical applications. mogstad2021causal then consider a weaker version of monotonicity, partial monotonicity, which permits heterogeneous choice behavior. Under partial monotonicity, multi-dimensional instruments (multiple instruments) can be used for causal effects estimation even if the full monotonicity condition fails.
In the literature, the IV validity conditions can be tested by the methods of huber2015testing, kitagawa2015test, mourifie2016testing, and sun2018ivvalidity. In this paper, we extend the frameworks of kitagawa2015test and sun2018ivvalidity to the partial monotonicity of mogstad2021causal. We show that the proposed test based on kitagawa2015test and sun2018ivvalidity performs well in practice. In the appendix, Monte Carlo studies demonstrate the finite sample properties of the test. We then apply the test to an empirical example from thornton2008demand discussed in mogstad2021causal.
mogstad2021causal introduce the concept of partial instrument monotonicity for multi-dimensional IVs. We follow mogstad2021causal and mainly focus on the multivalued ordered treatment case.\footnote{It would be easy to extend the test to unordered treatments.} Let $( \Omega, \mathcal{A}, \mathbb{P} )$ be a probability space on which all the random elements are well defined. Suppose that the outcome variable $Y\in\mathbb{R}$, the treatment $D\in\mathcal{D}=\left\{ d_{1},\ldots,d_J\right\}$ for some $J\ge 2$, and the instrument $\boldsymbol{Z}\in\mathcal{Z}=\{\boldsymbol{z}_1,\ldots,\boldsymbol{z}_K\}$ for some $K\ge 2$, where every value $\boldsymbol{z}\in\mathcal{Z}$ is a vector $\boldsymbol{z}=(z_1,\ldots,z_L)$ for some $L\ge 2$. That is, $\boldsymbol{Z}$ is an $L$-dimensional instrument with $\boldsymbol{Z}=(Z_1,\ldots,Z_L)$, where $Z_l$ is a scalar variable for every $l$. Suppose the $l$th dimension of $\boldsymbol{Z}$ has $k_l$ possible values, i.e., $Z_l\in\{z_l^1,\ldots,z_l^{k_l}\}$. Following the rectangular support assumption of mogstad2021causal, we suppose that $\mathcal{Z}=\mathrm{supp}(Z_1)\times\cdots\times\mathrm{supp}(Z_L)$, which implies $K=\prod_{l=1}^L k_l$. For every $\boldsymbol{z}=(z_1,\ldots,z_L)$, we define the vector $z_{-l}=(z_1,\ldots,z_{l-1},z_{l+1},\ldots,z_L)$, and $\boldsymbol{z}$ may be written as $\boldsymbol{z}=(z_{l},z_{-l})$. Suppose that $Y_{d\boldsymbol{z}}$ for $d\in\mathcal{D}$ and $\boldsymbol{z}\in\mathcal{Z}$, and $D_{\boldsymbol{z}}$ for $\boldsymbol{z}\in \mathcal{Z}$ are the potential random variables. The following assumption formalizes the IV validity assumption for multi-dimensional instrument $\boldsymbol{Z}$ proposed by mogstad2021causal.
Suppose that $D$ has maximum value $d_{\max}$ and minimum value $d_{\min}$. We provide a testable implication for Assumption (ref) in the following lemma.
Lemma (ref) can be proved analogously to Lemma 2.1 of sun2018ivvalidity. In the following, we extend the tests of kitagawa2015test and sun2018ivvalidity to partial IV validity. Without loss of generality, we assume that $d_{\min}=d_1\le\cdots\le d_J=d_{\max}$ with $d_{\min}=0$ and $d_{\max}=1$. Then the inequalities in (ref) and (ref) are equivalent to
for all $1 \leq l \leq L$, all $1\le k\le k_{l}-1$, all possible values $z_{-l}$, all closed intervals $B$ in $\mathbb{R}$, each $d\in\{0,1\}$, and all $C=(-\infty,c]$ with $c\in\mathbb{R}$. Following {Lemma B.7} of kitagawa2015test, here we only consider all closed intervals $B$ instead of all Borel sets $B$ when constructing the test. By definition, for all $B,C\in\mathcal{B}_{\mathbb{R}}$ and all possible values $\boldsymbol{z}$, $ \mathbb{P}\left( Y\in B,D\in C|\boldsymbol{Z}=\boldsymbol{z}\right)={\mathbb{P}\left( Y\in B,D\in C,\boldsymbol{Z}=\boldsymbol{z}\right) }/{\mathbb{P}\left( \boldsymbol{Z}=\boldsymbol{z}\right) }. $ We then define function spaces
Let $\{\left( Y_i,D_i,\boldsymbol{Z}_i \right)\}_{i=1}^{n} $ be an i.i.d.\ sample, which is distributed according to some probability distribution $P$, that is, the measure $P(G)=\mathbb{P}((Y_i,D_i,\boldsymbol{Z}_i)\in G)$ for all $G\in\mathcal{B}_{\mathbb{R}^{2+L}}$. For every measurable function $v$, by an abuse of notation, we define
For every $\left( h,g\right) \in {\bar{\mathcal{H}}\times\mathcal{G}}$ with $g=(g_{1},g_{2})$, define
The null hypothesis equivalent to (ref) is
and the alternative hypothesis is
We then define the sample analogue of $\phi$ by
where $\hat{P}$ denotes the empirical probability measure of $P$ such that for every measurable function $v$,
and $\{( Y_{i},D_{i},\boldsymbol{Z}_{i})\}_{i=1}^n$ is the i.i.d.\ sample distributed according to $P$. Define
for all $(h,g)\in\bar{\mathcal{H}}\times\mathcal{G}$ with $g=(g_1,g_2)$, where $T_n=n\cdot\prod_{k=1}^{K}\hat{P}\left(1_{\mathbb{R}\times\mathbb{R}\times\{\boldsymbol{z}_k\}}\right)$.
We then specify a closed set $\Xi\subset(0,1]$ such that $\Xi$ contains all the values of $\xi$ used for constructing the test statistic in the following. We also specify a positive measure $\nu$ on $\Xi$ that satisfies the following assumption.
We follow kitagawa2015test and sun2018ivvalidity and construct the test statistic as
We may set the measure $\nu$ to be a Dirac measure centered at some fixed $\xi\in\Xi$, and this is equivalent to using a particular value of $\xi$ to construct the statistic. We construct a random set $\widehat{\Psi_{{\mathcal{H}}\times\mathcal{G}}}$ by
with $\tau_{n}\rightarrow\infty$ and $\tau_{n}/\sqrt{n}\rightarrow0$ as $n\rightarrow\infty$, where $\xi_0$ is a fixed small positive number. In practice, we suggest setting $\xi_0=10^{-10}$ which is used in the simulations and the application in the paper. This random set is an estimator for some contact set similar to those in Beare2015improved and Beare2017improved in different contexts.\footnote{See linton2010improved and lee2013testing for more discussions on estimation of contact sets.} Algorithm (ref) illustrates the test procedure.
Proposition (ref) may be proved analogously to Theorem 3.2 in sun2018ivvalidity, so we omit the proof. The simulation study in Section (ref) in the supplementary appendix demonstrates the good finite sample properties of the test. The numerical results show that the test is asymptotically size controlled and consistent. Section (ref) in the appendix provides an empirical application thornton2008demand for the proposed test. We examine the partial validity of monetary incentives and distance from results centers as instruments for the knowledge of HIV status, and find that these instruments passed our test.