Wenlong Ji, Lihua Lei, Asher Spector
arXiv 12 Oct 2023 · Econometrics · 1 citations (OpenAlex)
arXiv:2310.08115 · PDF · DOI · OpenAlex · Extracted main text
Many causal estimands are only partially identifiable since they depend on the unobservable joint distribution between potential outcomes. Stratification on pretreatment covariates can yield sharper bounds; however, unless the covariates are discrete with relatively small support, this approach typically requires binning covariates or estimating the conditional distributions of the potential outcomes given the covariates. Binning can result in substantial efficiency loss and become challenging to implement, even with a moderate number of covariates. Estimating conditional distributions, on the other hand, may yield invalid inference if the distributions are inaccurately estimated, such as when a misspecified model is used or when the covariates are high-dimensional. In this paper, we propose a unified and model-agnostic inferential approach for a wide class of partially identified estimands. Our method, based on duality theory for optimal transport problems, has four key properties. First, in randomized experiments, our approach can wrap around any estimates of the conditional distributions and provide uniformly valid inference, even if the initial estimates are arbitrarily inaccurate. A simple extension of our method to observational studies is doubly robust in the usual sense. Second, if nuisance parameters are estimated at semiparametric rates, our estimator is asymptotically unbiased for the sharp partial identification bound. Third, we can apply the multiplier bootstrap to select covariates and models without sacrificing validity, even if the true model is not selected. Finally, our method is computationally efficient. Overall, in three empirical applications, our method consistently reduces the width of estimated identified sets and confidence intervals without making additional structural assumptions.
appendix boundary found by appendix_command · 38% of the source is main text. Read the extracted text to check this.
The works this paper leans on most, across its whole bibliography — not restricted to papers in our corpus. Ranked by composite intensity, which combines how often a work is mentioned, how many sections mention it, and how much of that falls in the main text rather than the appendix.
| Reference | Intensity | Mentions | Sections | Main text | |
|---|---|---|---|---|---|
| 1 | Lee, D. S (2009) Training, wages, and sample selection: Estimating sharp bounds on treatment effects | 1.000 | 6 | 3 | 100% |
| 2 | Imbens, G. and Manski, C (2004) Confidence intervals for partially identified parameters | 0.928 | 4 | 3 | 100% |
| 3 | Stoye, J (2009) More on confidence intervals for partially identified parameters | 0.928 | 4 | 3 | 100% |
| 4 | Semenova, V (2023) Adaptive estimation of intersection bounds: a classification approach | 0.874 | 8 | 2 | 100% |
| 5 | Semenova, V (2021) Generalized lee bounds | 0.811 | 4 | 2 | 100% |
| 6 | Chernozhukov, V., Chetverikov, D., and Kato, K (2018) Inference on Causal and Structural Parameters using Many Moment Inequalities | 0.737 | 4 | 3 | 50% |
| 7 | Jun, S. J. and Lee, S (2023) Identifying the effect of persuasion | 0.737 | 3 | 2 | 100% |
| 8 | Gerber, A. S., Karlan, D., and Bergan, D (2009) Does the media matter? a field experiment measuring the effect of newspapers on voting behavior and political opinions | 0.644 | 4 | 1 | 100% |
| 9 | Chernozhukov, V., Chetverikov, D., and Kato, K (2013) Gaussian approximations and multiplier bootstrap for maxima of sums of high-dimensional random vectors | 0.644 | 2 | 2 | 100% |
| 10 | Fang, Z., Santos, A., Shaikh, A. M., and Torgovitsky, A (2023) Inference for large-scale linear systems with known coefficients | 0.644 | 2 | 2 | 100% |
Showing the top 10 of 106 scored citations.
arXiv econ.EM papers that cite this one, ranked by how heavily they lean on it.