Extracted main text — title through conclusion, appendix excluded. This is what our citation measures are computed over, published so the extraction can be checked by eye.
62,508 characters · 44 sections · 61 citation commands
Nonparametric Identification and Estimation of Spatial Treatment Effect Boundaries: Evidence from 42 Million Pollution Observations
Spatial spillovers are ubiquitous in economic applications. Environmental policies affect pollution in neighboring regions deryugina2019mortality, knittel2016caution, transportation infrastructure impacts economic activity far from construction sites donaldson2018railroads, duranton2014roads, and place-based policies generate effects that propagate across space busso2013assessing, kline2019place. A fundamental challenge in such settings is determining where treatment effects end—that is, identifying the spatial boundary beyond which spillover effects become negligible and units can serve as valid controls.
Recent methodological advances have made progress on this question. butts2023difference provides a comprehensive treatment of spatial difference-in-differences estimators, showing that neglecting spillovers can severely bias treatment effect estimates. kikuchi2024unified and kikuchi2024navier derive spatial boundaries from first principles using atmospheric dispersion models, demonstrating that under simplified transport assumptions, treatment effects should decay exponentially with distance. kikuchi2024stochastic develops stochastic approaches for settings with pervasive spillovers and general equilibrium effects. However, existing approaches face a fundamental tension: parametric methods assume strong functional forms (exponential, power law) that may be misspecified, while nonparametric distance bin estimators are inefficient and cannot extrapolate beyond observed distances.
This paper develops a nonparametric framework that resolves this tension. I establish conditions under which spatial boundaries can be identified without parametric assumptions on decay functions, propose kernel-based estimators with optimal bandwidth selection, and derive asymptotic theory for inference. The approach maintains the interpretability and extrapolation capability of parametric methods while providing robustness to functional form misspecification.
Using 42 million satellite observations of nitrogen dioxide (NO$_2$) from the TROPOMI instrument on board the Sentinel-5P satellite, combined with locations of 318 coal-fired power plants, I estimate spatial decay functions nonparametrically and compare results to parametric baselines. The data cover 2019-2021, providing both cross-sectional variation in distance from plants and temporal variation through the COVID-19 pandemic.
Finding 1: Nonparametric methods reduce prediction errors. Across 42 million grid-cell observations, nonparametric kernel regression reduces mean absolute prediction errors by 1.0 percentage point compared to exponential decay assumptions. The improvement is largest at policy-relevant distances: 2.8 percentage points at 10 km (where near-source damages are highest) and 3.7 percentage points at 100 km (where long-range transport matters for aggregate damages).
Finding 2: Parametric methods exhibit systematic bias patterns. Exponential decay underestimates pollution near sources (12.9% error at 10 km) and overestimates decay at distance (16.9% error at 100 km), while nonparametric methods reduce these errors to 10.1% and 13.2%, respectively. This pattern is consistent with violations of the idealized assumptions underlying exponential decay—steady winds, homogeneous atmospheres, flat terrain—that theory predicts should be most severe at extreme distances.
Finding 3: The COVID-19 pandemic validates temporal sensitivity. NO$_2$ concentrations dropped 4.6% in 2020 relative to 2019, then recovered 5.7% in 2021. Both parametric and nonparametric methods detect this temporal variation, but nonparametric approaches maintain superior prediction accuracy across all years (0.9-1.1 percentage point improvement), demonstrating robustness to temporary shocks.
Finding 4: Results are robust to spatial correlation. Using spatial HAC standard errors following muller2022spatial, I find that accounting for residual spatial correlation increases confidence interval widths by 10-15% but does not alter the main conclusions. This demonstrates the value of combining nonparametric boundary estimation with spatial correlation robust inference.
This paper makes four main contributions to the literature on spatial treatment effects and econometric methodology:
1. Nonparametric Identification Framework. I establish that spatial boundaries are nonparametrically identified under weak regularity conditions—smoothness and monotonicity—without requiring knowledge of the decay function's parametric form (Theorems (ref) and (ref)). When monotonicity fails, I provide partial identification results yielding bounds on the boundary (Theorem (ref)). This extends recent work on spatial econometrics muller2022spatial, muller2024spatial by showing how their insights about flexible spatial structures apply to treatment effect propagation.
2. Kernel-Based Estimation with Optimal Bandwidth Selection. I propose a local polynomial regression estimator for the decay function $m(d)$ and a plug-in estimator for the boundary $d^*$. Crucially, I develop a data-driven cross-validation procedure for bandwidth selection that optimizes boundary estimation directly, rather than only minimizing mean squared error of the decay function. This addresses a key practical challenge in nonparametric spatial analysis.
3. Asymptotic Theory and Inference. I derive the asymptotic distribution of the boundary estimator (Theorem (ref)), showing that it achieves the optimal nonparametric rate of $n^{-2/5}$ under standard kernel regression conditions. I provide consistent variance estimators and develop both analytical and bootstrap-based confidence intervals. Importantly, I show how to integrate spatial correlation robust inference muller2022spatial with nonparametric boundary estimation when residual spatial dependence is present.
4. Large-Scale Empirical Application. The application to coal plant pollution demonstrates the framework's practical value using an unprecedented scale of data (42 million observations) and provides the first comprehensive test of whether atmospheric dispersion theory's exponential decay prediction holds empirically. The systematic rejection of exponential functional form across multiple pollutants and years validates the need for flexible estimation methods in environmental applications.
This work connects three literatures in econometrics and environmental economics.
The spatial econometrics literature has long recognized that treatments can have geographic spillovers anselin1988spatial, conley1999gmm, lesage2009introduction. Recent work develops spatial difference-in-differences estimators butts2023difference and spatial HAC inference colella2019inference. This literature typically specifies spatial weights matrices based on ad hoc assumptions. My contribution is providing nonparametric methods for data-driven weights specification based on estimated decay functions.
Most recently, muller2022spatial develop a comprehensive framework for spatial correlation robust inference, showing how to construct valid confidence intervals when the spatial correlation structure is unknown. muller2024spatial demonstrate that spatial data can exhibit behavior analogous to time series unit roots, leading to spurious spatial regressions when persistent spatial trends are present.
My work complements these advances in three ways. First, while muller2022spatial address nuisance correlation in errors, I study treatment-induced spatial patterns that decay with distance from sources. Both sources of correlation may be present simultaneously, and I show how to combine the methods (Section (ref)). Second, I address the spurious regression concerns of muller2024spatial through regional heterogeneity analysis: decay patterns change sign at 100 km (positive within, negative beyond), which is inconsistent with spurious trends but consistent with heterogeneous treatment effects (Section (ref)). Third, I extend their theoretical framework from spatial correlation to treatment effect boundaries, generalizing their "half-life" concept to arbitrary policy-relevant thresholds.
kikuchi2024navier derives spatial boundaries from atmospheric dispersion models based on the Navier-Stokes equations, showing that under idealized assumptions (steady-state flow, homogeneous atmospheres, flat terrain), pollution should decay exponentially with distance. kikuchi2024unified extends this framework to provide a unified treatment of both spatial and temporal boundaries, while kikuchi2024stochastic develops stochastic approaches for settings where general equilibrium effects and pervasive spillovers complicate boundary detection. However, these frameworks assume researchers know the correct parametric form. Environmental economics studies of air pollution deryugina2019mortality, knittel2016caution, muller2011damages typically impose exponential or power law decay without testing these functional form assumptions.
I bridge atmospheric science and econometrics by treating dispersion theory as providing testable predictions rather than maintained assumptions. The mathematical framework underlying atmospheric transport—the advection-diffusion equation derived from Navier-Stokes principles—corresponds to a specific parametric form for the spatial kernel in muller2022spatial's representation. But when the underlying assumptions fail (as they inevitably do with real-world coal plants), the true decay function may deviate substantially from the exponential benchmark. This paper provides the tools to test these deviations and estimate boundaries nonparametrically.
My estimator builds on local polynomial regression fan1996local, ruppert1995effective and optimal bandwidth selection fan1992variable, jones1996brief. The novelty is adapting these tools to boundary estimation rather than function estimation, requiring modified cross-validation criteria. The asymptotic theory extends fan1996local's results to functionals of nonparametric estimators, similar in spirit to muller2009inference's work on changepoint detection but for smooth decay rather than discrete jumps.
My estimator builds on local polynomial regression fan1996local, ruppert1995effective and optimal bandwidth selection fan1992variable, jones1996brief. The novelty is adapting these tools to boundary estimation rather than function estimation, requiring modified cross-validation criteria. The asymptotic theory extends fan1996local's results to functionals of nonparametric estimators, similar in spirit to muller2009inference's work on changepoint detection but for smooth decay rather than discrete jumps.
This paper builds on and extends my earlier work on spatial boundaries. kikuchi2024navier derives boundaries from physical principles (Navier-Stokes equations), demonstrating that atmospheric dispersion implies exponential decay under idealized assumptions. kikuchi2024unified provides a unified framework for both spatial and temporal boundaries, showing how to identify treatment effect propagation in both dimensions simultaneously. kikuchi2024stochastic extends to settings with stochastic diffusion and general equilibrium effects, where boundaries must be characterized probabilistically rather than deterministically.
This paper complements these earlier contributions by:
The progression across papers is: kikuchi2024navier establishes the physical foundation, kikuchi2024unified extends to multiple dimensions, kikuchi2024stochastic handles stochastic settings, and this paper provides robust nonparametric methods when functional forms are uncertain.
Section 2 develops the theoretical framework, connecting atmospheric dispersion models to spatial econometric representations and defining spatial boundaries. Section 3 establishes nonparametric identification results. Section 4 develops the kernel-based estimator and derives asymptotic theory. Section 5 describes the TROPOMI NO$_2$ data and coal plant locations. Section 6 presents the empirical results, comparing parametric and nonparametric estimates and testing functional form assumptions. Section 7 discusses extensions, spatial correlation robust inference, and diagnostics for spurious regression. Section 8 concludes with implications for research design and policy analysis.
The spatial propagation of airborne pollutants from point sources follows well-established principles of fluid dynamics and mass transport. Building on kikuchi2024navier, who derives these relationships from the Navier-Stokes equations, under idealized conditions—steady-state flow, homogeneous atmospheric conditions, and flat terrain—the concentration of a pollutant at distance $d$ from a source can be characterized by the advection-diffusion equation seinfeld2016atmospheric:
where $C$ is pollutant concentration, $u$ is wind velocity, $D$ is the diffusion coefficient, and $S$ represents source emissions. The solution under Gaussian plume assumptions yields pasquill1976atmospheric:
where $\kappa$ is the spatial decay parameter governed by meteorological conditions and pollutant chemistry, and $A$ is source strength. This exponential form is the cornerstone of the parametric approach in kikuchi2024unified and kikuchi2024navier.
where $\kappa$ is the spatial decay parameter governed by meteorological conditions and pollutant chemistry, and $A$ is source strength.
The exponential decay form in equation ((ref)) provides a useful theoretical benchmark but relies on restrictive assumptions that may be violated in practice. Recent developments in spatial econometrics offer a complementary perspective. muller2022spatial develop a general framework for spatial processes with arbitrary correlation structures:
where $Y(s)$ represents the outcome at location $s$, $K(\cdot)$ is a spatial kernel function, and $\varepsilon(r)$ captures innovations at location $r$. Critically, they show that imposing incorrect parametric forms for $K(\cdot)$ can lead to substantial bias in both estimation and inference.
In our context, the exponential decay corresponds to a specific parametric assumption about the kernel: $K(d) = A\exp(-\kappa d)$. However, this form is justified only when the underlying theoretical assumptions hold:
Real-world emissions violate these assumptions in systematic ways. Wind patterns vary across time and space, terrain channels pollutant flows, multiple emission sources create complex concentration fields, and pollutants like NO$_2$ undergo photochemical reactions (NO$_2$ $\leftrightarrow$ NO + O$_3$). These violations suggest that the actual spatial decay function $m(d) = E[\text{Pollution} | \text{Distance} = d]$ may deviate substantially from the exponential form.
Following muller2022spatial, I characterize spatial decay through the concept of a spatial boundary. Define $d^*(\varepsilon)$ as the distance where pollution decays to $\varepsilon$ times its source level:
where $\varepsilon \in (0,1)$ is a policy-relevant threshold. This generalizes muller2022spatial's "half-life" measure (corresponding to $\varepsilon = 0.5$) to thresholds more relevant for environmental policy. The concept extends the deterministic boundaries of kikuchi2024unified and kikuchi2024navier by allowing for nonparametric estimation, and complements the stochastic boundaries of kikuchi2024stochastic by focusing on point-source treatments where deterministic decay dominates.
Under the exponential decay assumption, the boundary has a closed form:
However, if the true decay function deviates from exponential form, this parametric boundary will be misspecified. My nonparametric approach estimates $d^*$ directly from data without imposing functional form restrictions.
My framework uses theoretical predictions as a benchmark while remaining agnostic about functional form. Rather than assuming exponential decay holds, I test whether it provides an adequate approximation to actual spatial patterns. This approach offers several advantages:
The key insight from muller2022spatial—that spatial correlation structures should be estimated flexibly rather than imposed parametrically—applies directly to our setting. Just as time series methods have moved from parametric ARMA models to flexible nonparametric alternatives when confronted with misspecification concerns, spatial econometrics increasingly favors data-driven approaches when underlying theoretical restrictions may not hold.
We observe a random sample $(Y_i, D_i, \mathbf{X}_i)$ for $i = 1, \ldots, n$, where:
The conditional mean function is:
In a spatial treatment effects setting, $m(d)$ represents the expected outcome as a function of distance from treatment. Under standard identifying assumptions (conditional independence, overlap), $m(d)$ has a causal interpretation as the spatially-varying treatment effect.
This section establishes identification of the decay function $m(d)$ and boundary $d^*$ under weak regularity conditions, without parametric restrictions.
We first consider identification of the conditional mean function $m(d)$.
We now turn to identification of the spatial boundary $d^*$. This requires additional structure beyond smoothness.
When strict monotonicity fails, we can still obtain bounds on the boundary.
This section develops a kernel-based estimator for the decay function and boundary, derives its asymptotic distribution, and provides methods for inference.
I estimate $m(d)$ using local polynomial regression of order $p$ fan1996local.
Common choices include the Epanechnikov kernel:
Given the estimated decay function $\hat{m}(d)$, I estimate the boundary via a plug-in approach.
In practice, I evaluate $\hat{m}(d)$ on a fine grid $\{d_1, \ldots, d_G\}$ and find:
Bandwidth selection is critical for nonparametric estimation. Standard mean squared error (MSE) minimization for $\hat{m}(d)$ may not be optimal for boundary estimation. I use cross-validation with Silverman's rule of thumb as a starting point:
where $\sigma_D$ is the standard deviation of distance. This provides the optimal rate for local linear regression while remaining computationally tractable for the large-scale TROPOMI dataset.
I now derive the asymptotic distribution of the boundary estimator.
Given the large sample size and computational constraints, I use a subsample bootstrap approach:
The subsample bootstrap provides valid inference while remaining computationally feasible for the 42 million observation dataset.
I use nitrogen dioxide (NO$_2$) concentration data from the TROPOspheric Monitoring Instrument (TROPOMI) on board the European Space Agency's Sentinel-5P satellite. TROPOMI provides daily global coverage at 3.5 × 5.5 km spatial resolution, offering unprecedented detail for studying air pollution dispersion.
Sample construction:
Data quality:
Coal plant locations come from the EPA Emissions and Generation Resource Integrated Database (eGRID), which provides comprehensive data on all power plants in the United States:
Table (ref) presents summary statistics for the analysis sample.
Table (ref) presents the main comparison of parametric and nonparametric boundary estimation.
Key finding: Nonparametric methods reduce prediction errors by 1.0 percentage point on average, with consistent improvements across all three years (0.9-1.1 pp). The improvement is economically meaningful given the scale of NO$_2$ concentrations and the policy applications of these estimates.
Figure (ref) illustrates the spatial decay patterns for each year, comparing actual observations (black dots) with parametric exponential predictions (dashed red lines) and nonparametric kernel estimates (solid blue lines).
Table (ref) decomposes prediction errors by distance from coal plants, revealing where parametric misspecification is most severe.
Key finding: The pattern of improvement is precisely what theory predicts. Exponential decay assumptions—derived under idealized conditions—perform worst where those conditions are most violated:
This U-shaped pattern of parametric bias validates both the theoretical framework (exponential decay is approximately correct in mid-range) and the need for flexibility at extremes.
Figure (ref) visualizes this U-shaped error pattern, showing how nonparametric methods maintain more consistent accuracy across all distances.
Figure (ref) presents a comprehensive view of nonparametric improvements across all year-distance combinations, revealing consistent patterns of superior performance at extremes.
The COVID-19 pandemic provides a natural experiment for validating the framework's ability to detect temporal changes in pollution patterns.
Key findings:
Figure (ref) illustrates the COVID-19 natural experiment, showing both the temporal dynamics of NO$_2$ concentrations and the consistent superiority of nonparametric methods across all three years.
I formally test whether exponential decay provides adequate functional form using specification tests based on residuals.
Procedure:
Test statistic: Integrated squared deviation:
where $\hat{g}(d)$ is the nonparametric regression of residuals on distance.
Results:
Conclusion: Strong evidence against exponential functional form in all years, validating the need for nonparametric estimation.
Figure (ref) visualizes the residual patterns from both estimation approaches, demonstrating the systematic nature of parametric bias at extreme distances.
To validate that decay patterns reflect causal effects rather than spurious trends, I estimate spatial decay separately by distance from coal plants.
Key finding: Decay parameters **change sign** at the 100 km threshold. Within 100 km, pollution decreases with distance (positive spatial decay), consistent with coal plant effects. Beyond 100 km, pollution increases with distance (negative decay), reflecting dominance of urban sources. This sign reversal is inconsistent with spurious spatial trends but consistent with treatment effect heterogeneity—precisely what muller2024spatial recommend as a diagnostic for ruling out spurious regression.
While my baseline asymptotic theory assumes independence across observations, outcomes may exhibit spatial correlation beyond treatment-induced patterns. Following muller2022spatial, I compute spatial HAC standard errors that remain valid under arbitrary spatial correlation.
The spatial HAC standard errors are 15-20% larger than iid standard errors, indicating modest residual spatial correlation in pollution beyond treatment effects. However, all main results remain statistically significant, and the relative performance of parametric vs nonparametric methods is unchanged.
kikuchi2024stochastic develops a complementary framework for settings where treatment effects propagate through stochastic diffusion processes and general equilibrium channels. While that framework is designed for pervasive spillovers (e.g., technology adoption networks, trade shocks), my nonparametric approach is most suitable for point-source treatments (e.g., power plants, infrastructure) where:
The two frameworks are complementary:
Future work could integrate these frameworks, developing nonparametric methods for stochastic boundary detection in general equilibrium settings. This would combine the flexibility of my kernel-based estimator with the generality of stochastic diffusion models.
A potential concern is that estimated decay patterns reflect spurious spatial trends rather than causal treatment effects. I address this through state fixed effects.
Results are robust to state fixed effects, further confirming that spatial trends do not drive findings. Combined with the regional heterogeneity analysis (sign reversal at 100 km), these diagnostics provide strong evidence against spurious regression concerns raised by muller2024spatial.
The difference between parametric and nonparametric estimates has important implications for environmental policy and research design:
1. Damage function estimation: Parametric methods that underestimate near-source concentrations (10 km: 12.9% error) while overestimating long-range transport (100 km: 16.9% error) lead to systematic bias in damage calculations. Given that health damages are nonlinear in pollution exposure, underestimating peak concentrations near sources may substantially understate total damages.
2. Research design for spatial DiD: When designing studies of coal plant effects, parametric boundaries would suggest placing controls farther from plants than necessary, reducing the available control region and potentially introducing bias if distant controls are affected by other pollution sources (as our regional heterogeneity analysis suggests).
3. Regulatory buffer zones: Environmental regulations often specify geographic buffer zones around pollution sources. Nonparametric estimates suggest these zones should be designed with attention to near-source complexities rather than simple exponential extrapolation.
This paper develops a nonparametric framework for identifying and estimating spatial boundaries of treatment effects when decay functions may deviate from parametric assumptions. Using 42 million satellite observations of NO$_2$ concentrations near coal plants, I demonstrate that flexible, data-driven methods substantially outperform parametric exponential decay assumptions commonly used in environmental economics.
Main findings:
Practical implications: For applied researchers conducting spatial difference-in-differences studies:
Broader implications: This work demonstrates the value of combining economic theory with statistical flexibility. Atmospheric dispersion models provide interpretable parameters and causal mechanisms, while nonparametric estimation ensures robustness when real-world phenomena deviate from idealized models. This combination offers a template for empirical research in settings where treatment effects propagate across space—from infrastructure projects to technology adoption to disease transmission—where theory provides guidance but data should discipline functional form assumptions.
Future directions: Natural extensions include developing methods for time-varying spatial boundaries, incorporating multiple treatment sources through additive models, and addressing endogenous treatment placement through nonparametric instrumental variables. The framework could also be extended to other pollutants (PM$_{2.5}$, SO$_{2}$, CO) and other applications beyond environmental economics where spatial treatment propagation is theoretically motivated but parametric forms uncertain.
This research was supported by a grant-in-aid from Zengin Foundation for Studies on Economics and Finance.