ML-QuantSubscribe

Machine learningOther

Deep Linear Network Training Dynamics from Random Initialization: Data, Width, Depth, and Hyperparameter Transfer

The paper explores the dynamics of gradient descent in deep linear networks, discussing the impact of network width and depth, and comparing various training dynamics.

Featured in No. 107 on 25 Jul 2025 · · 17 citations today · published in International Conference on Machine Learning

Released
4 Feb 2025
First featured
No. 107 · 25 Jul 2025
Citations (Semantic Scholar)
17
Influential citations
0
Published in
International Conference on Machine Learning
Shares when featured
26
Identifier
arXiv:2502.02531

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page