EconBase
← All papers

Scale-Insensitive Neural Network Significance Tests

Hasan Fallahgoul

arXiv 27 Jan 2025 · Statistics — Machine Learning

arXiv:2501.15753 · PDF · DOI · OpenAlex · Extracted main text

Abstract

This paper develops a scale-insensitive framework for neural network significance testing, substantially generalizing existing approaches through three key innovations. First, we replace metric entropy calculations with Rademacher complexity bounds, enabling the analysis of neural networks without requiring bounded weights or specific architectural constraints. Second, we weaken the regularity conditions on the target function to require only Sobolev space membership $H^s([-1,1]^d)$ with $s > d/2$, significantly relaxing previous smoothness assumptions while maintaining optimal approximation rates. Third, we introduce a modified sieve space construction based on moment bounds rather than weight constraints, providing a more natural theoretical framework for modern deep learning practices. Our approach achieves these generalizations while preserving optimal convergence rates and establishing valid asymptotic distributions for test statistics. The technical foundation combines localization theory, sharp concentration inequalities, and scale-insensitive complexity measures to handle unbounded weights and general Lipschitz activation functions. This framework better aligns theoretical guarantees with contemporary deep learning practice while maintaining mathematical rigor.

Citation extraction

18
references
86
in-text mentions
18
distinct cited
0
self-citations
7,195
main-text words

appendix boundary found by appendix_command · 45% of the source is main text. Read the extracted text to check this.

Most heavily cited references

The works this paper leans on most, across its whole bibliography — not restricted to papers in our corpus. Ranked by composite intensity, which combines how often a work is mentioned, how many sections mention it, and how much of that falls in the main text rather than the appendix.

ReferenceIntensityMentionsSectionsMain text
1Fallahgoul, Franstianto \ Lin (2024) `Asset pricing with neural networks: Significance tests', Journal of Econometrics 238(1), 1055741.000163100%
2Horel \ Giesecke (2020) `Significance tests for neural networks', Journal of Machine Learning Research 21(227), 1–291.000113100%
3Farrell, Liang \ Misra (2021) `Deep neural networks for estimation and inference', Econometrica 89(1), 181–2130.95229586%
4Bartlett, Bousquet \ Mendelson (2005) `Local rademacher complexities', The Annals of Statistics 33, 1497–15370.73732100%
5Koltchinskii \ Panchenko (2000) Rademacher processes and bounding the risk of function learning, in `High Dimensional Probability II', Springer, pp. 443–4570.73732100%
6Yarotsky (2017) `Error bounds for approximations with deep relu networks', Neural networks 94, 103–1140.73732100%
7Yarotsky (2018) Optimal approximation of continuous functions by very deep relu networks, in `in 31st Conference on learning theory', PMLR, pp.…0.73732100%
8Anthony \ Bartlett (1999) Neural Network Learning: Theoretical Foundations, Cambridge University Press, Cambridge0.64422100%
9Goodfellow, Bengio \ Courville (2016) Deep Learning, MIT Press, Cambridge0.64422100%
10van der Vaart \ Wellner (1996) Weak Convergence and Empirical Processes: With Applications to Statistics, Springer Series in Statistics, Springer, New York0.5855420%

Showing the top 10 of 18 scored citations.