Technology & AIJul 12, 2026
Singular perturbations and hierarchical learning in two-layer neural networks
We study the population gradient flow of an infinitely wide two-layer neural network learning a misspecified single-index model in high dimension.
We study the population gradient flow of an infinitely wide two-layer neural network learning a misspecified single-index model in high dimension. The two layers are optimized jointly, with a perturbative parameter tuning the relative training speed between the first and second…
Sign in to learn & save →
The frontier is open to all. Sign in to learn this from first principles and save it to your knowledge base.