arXiv 22 Feb 2024 · Statistics — Machine Learning · 1 citations (OpenAlex)
arXiv:2402.14264 · PDF · DOI · OpenAlex · Extracted main text
Average treatment effect estimation is the most central problem in causal inference with application to numerous disciplines. While many estimation strategies have been proposed in the literature, the statistical optimality of these methods has still remained an open area of investigation, especially in regimes where these methods do not achieve parametric rates. In this paper, we adopt the recently introduced structure-agnostic framework of statistical lower bounds, which poses no structural properties on the nuisance functions other than access to black-box estimators that achieve some statistical estimation rate. This framework is particularly appealing when one is only willing to consider estimation strategies that use non-parametric regression and classification oracles as black-box sub-processes. Within this framework, we prove the statistical optimality of the celebrated and widely used doubly robust estimators for both the Average Treatment Effect (ATE) and the Average Treatment Effect on the Treated (ATT), as well as weighted variants of the former, which arise in policy evaluation.
appendix boundary found by appendix_command · 47% of the source is main text. Read the extracted text to check this.
The works this paper leans on most, across its whole bibliography — not restricted to papers in our corpus. Ranked by composite intensity, which combines how often a work is mentioned, how many sections mention it, and how much of that falls in the main text rather than the appendix.
| Reference | Intensity | Mentions | Sections | Main text | |
|---|---|---|---|---|---|
| 1 | Sivaraman Balakrishnan, Edward H Kennedy, and Larry Wasserman (2023) The fundamental limits of structure-agnostic functional estimation | 1.000 | 12 | 4 | 100% |
| 2 | James Robins, Eric Tchetgen Tchetgen, Lingling Li, and Aad van der V… (2009) Semiparametric minimax rates | 1.000 | 10 | 3 | 100% |
| 3 | S Balakrishnan and L Wasserman (2019) Hypothesis testing for densities and high-dimensional multinomials: Sharp local minimax rates | 0.811 | 4 | 2 | 100% |
| 4 | Edward H Kennedy, Sivaraman Balakrishnan, James M Robins, and Larry… (2022) Minimax rates for heterogeneous causal effect estimation | 0.644 | 4 | 1 | 100% |
| 5 | Ery Arias-Castro, Bruno Pelletier, and Venkatesh Saligrama (2018) Remember the curse of dimensionality: The case of goodness-of-fit testing in arbitrary dimension | 0.644 | 2 | 2 | 100% |
| 6 | Max H Farrell, Tengyuan Liang, and Sanjog Misra (2021) Deep neural networks for estimation and inference | 0.644 | 2 | 2 | 100% |
| 7 | Yu I Ingster (1994) Minimax detection of a signal in $_p$ metrics | 0.644 | 2 | 2 | 100% |
| 8 | Anselm Johannes Schmidt-Hieber (2020) Nonparametric regression using deep neural networks with relu activation function | 0.644 | 2 | 2 | 100% |
| 9 | Vasilis Syrgkanis and Manolis Zampetakis (2020) Estimation and inference with trees and forests in high dimensions self | 0.644 | 2 | 2 | 100% |
| 10 | Alexandre B Tsybakov (2008) Introduction to nonparametric estimation | 0.644 | 2 | 2 | 100% |
Showing the top 10 of 79 scored citations.
arXiv econ.EM papers that cite this one, ranked by how heavily they lean on it.