Paper: 2604.19738 Authors: Simmaco Di Lillo, Leonardo Maini, Domenico Marinucci (University of Rome Tor Vergata) Categories: cs.LG, stat.ML
Abstract
Establishes central and non-central limit theorems for sequences of functionals of the Gaussian output of an infinitely-wide random neural network on the d-dimensional sphere.
Three Distinct Limiting Regimes
The asymptotic behavior depends crucially on the fixed points of the covariance function:
-
Low-disorder regime (κ’(1)
< 1):- Convergence to same functional of a limiting Gaussian field
- Typically non-Gaussian limiting distribution
-
High-disorder regime (κ’(1)
> 1):- Convergence to Gaussian or non-Gaussian with phase transition
- Depends on activation function and input space dimension
-
Sparse regime (κ’(1) = 1):
- Critical boundary behavior
- Includes ReLU activation with appropriate normalization
Mathematical Tools
Proofs exploit:
- Hermite expansions
- Diagram Formula
- Stein-Malliavin techniques
- Fixed-point analysis of iterative operators
Key Insight
Asymptotic behavior is determined by the fixed-point structure of the iterative operator associated with the covariance, whose nature and stability governs the different limiting regimes.
Takeaways
- Rich phase transition phenomena in deep random networks
- Fixed-point analysis provides unifying framework
- Higher dimensions make Gaussian behavior harder to achieve
论文: 2604.19738 作者: Simmaco Di Lillo, Leonardo Maini, Domenico Marinucci(罗马大学) 分类: cs.LG, stat.ML
摘要
为d维球面上无限宽随机神经网络高斯输出的泛函序列建立了中心和非中心极限定理。
三个不同极限状态
渐近行为关键取决于协方差函数的不动点:
-
低无序状态(κ’(1)
< 1):- 收敛到极限高斯场的相同泛函
- 通常为非高斯极限分布
-
高无序状态(κ’(1)
> 1):- 收敛到高斯或非高斯,带相变
- 取决于激活函数和输入空间维度
-
稀疏状态(κ’(1) = 1):
- 临界边界行为
- 包括适当归一化的ReLU激活
数学工具
证明利用了:
- Hermite展开
- 图公式
- Stein-Malliavin技术
- 迭代算子的不动点分析
关键洞察
渐近行为由与协方差相关的迭代算子的不动点结构决定,其性质和稳定性控制不同的极限状态。
要点总结
- 深度随机网络中丰富的相变现象
- 不动点分析提供统一框架
- 更高维度使高斯行为更难实现