EconBase
← All papers

Inferring Treatment Effects in Large Panels by Uncovering Latent Similarities

Ben Deaner, Chen-Wei Hsiang, Andrei Zeleneev

arXiv 26 Mar 2025 · Econometrics

arXiv:2503.20769 · PDF · DOI · OpenAlex · Extracted main text

Abstract

The presence of unobserved confounders is one of the main challenges in identifying treatment effects. In this paper, we propose a new approach to causal inference using panel data with large large $N$ and $T$. Our approach imputes the untreated potential outcomes for treated units using the outcomes for untreated individuals with similar values of the latent confounders. In order to find units with similar latent characteristics, we utilize long pre-treatment histories of the outcomes. Our analysis is based on a nonparametric, nonlinear, and nonseparable factor model for untreated potential outcomes and treatments. The model satisfies minimal smoothness requirements. We impute both missing counterfactual outcomes and propensity scores using kernel smoothing based on the constructed measure of latent similarity between units, and demonstrate that our estimates can achieve the optimal nonparametric rate of convergence up to log terms. Using these estimates, we construct a doubly robust estimator of the period-specifc average treatment effect on the treated (ATT), and provide conditions, under which this estimator is $\sqrt{N}$-consistent, and asymptotically normal and unbiased. Our simulation study demonstrates that our method provides accurate inference for a wide range of data generating processes.

Citation extraction

51
references
170
in-text mentions
56
distinct cited
1
self-citations
14,638
main-text words

appendix boundary found by appendix_command · 47% of the source is main text. Read the extracted text to check this.

Most heavily cited references

The works this paper leans on most, across its whole bibliography — not restricted to papers in our corpus. Ranked by composite intensity, which combines how often a work is mentioned, how many sections mention it, and how much of that falls in the main text rather than the appendix.

ReferenceIntensityMentionsSectionsMain text
1Abadie, A., Agarwal, A., Dwivedi, R., and Shah, A (2024) Doubly robust inference in causal latent factor models1.000213100%
2Chernozhukov, V., Chetverikov, D., Demirer, M., Duflo, E., Hansen, C… (2018) Double/debiased machine learning for treatment and structural parameters: Double/debiased machine learning1.00093100%
3Fernández-Val, I., Freeman, H., and Weidner, M (2021) Low-rank approximations of nonseparable panel models1.00053100%
4Feng, Y (2024) Causal inference in possibly nonlinear factor models0.874202100%
5Zhang, Y., Levina, E., and Zhu, J (2017) Estimating network edge probabilities by neighbourhood smoothing0.874102100%
6Stone, C. J (1980) Optimal rates of convergence for nonparametric estimators0.87472100%
7Athey, S., Bayati, M., Doudchenko, N., Imbens, G., and Khosravi, K (2021) Matrix completion methods for causal panel data models0.81142100%
8Robins, J., Li, L., Tchetgen, E., van der Vaart, A., et al (2008) Higher order influence functions and minimax estimation of nonlinear functionals0.81142100%
9Sun, L. and Abraham, S (2021) Estimating dynamic treatment effects in event studies with heterogeneous treatment effects0.81142100%
10Agarwal, A., Dahleh, M., Shah, D., and Shen, D (2021) Causal matrix completion0.73732100%

Showing the top 10 of 56 scored citations.

Cited by, within the corpus

arXiv econ.EM papers that cite this one, ranked by how heavily they lean on it.

Citing paperIntensityMentionsSections
1Inference on Linear Regressions with Two-Way Unobserved Heterogeneity0.92853
2Flexible Imputation of Incomplete Network Data0.87462
3Inference after discretizing time-varying unobserved heterogeneity0.40511