Machine learningOther
Deep Linear Network Training Dynamics from Random Initialization: Data, Width, Depth, and Hyperparameter Transfer
The paper explores the dynamics of gradient descent in deep linear networks, discussing the impact of network width and depth, and comparing various training dynamics.
Featured in No. 107 on 25 Jul 2025 · · 17 citations today · published in International Conference on Machine Learning
- Released
- 4 Feb 2025
- First featured
- No. 107 · 25 Jul 2025
- Citations (Semantic Scholar)
- 17
- Influential citations
- 0
- Published in
- International Conference on Machine Learning
- Shares when featured
- 26
- Identifier
- arXiv:2502.02531
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).