Feature Learning and Signal Propagation in Deep Neural Networks
Yizhang Lou, Chris Mingard, Yoonsoo Nam, Soufiane Hayou
阅读操作
确认中在文库中上传 PDF 后可生成中文音频讲解。
摘要与影响
Recent work by Baratin et al. (2021) sheds light on an intriguing pattern that occurs during the training of deep neural networks: some layers align much more with data compared to other layers (where the alignment is defined as the euclidean product of the tangent features matrix and the data labels matrix). The curve of the alignment as a function of layer index (generally) exhibits an ascent-descent pattern where the maximum is reached for some hidden layer. In this work, we provide the first explanation for this phenomenon. We introduce the Equilibrium Hypothesis which connects this alignment pattern to signal propagation in deep neural networks. Our experiments demonstrate an excellent match with the theoretical predictions.
逐年被引趋势
关键指标
同类平均 = 1
同领域 · 同年份 · 同类型
Google Scholar 与 OpenAlex 的被引统计范围不同,数值存在差异属正常。
AI 辅助阅读
依据:摘要
可就本文提问;依据不足时会说明。
学术脉络
学科主题
计算机 / AINeural Networks and Applications
Stochastic Gradient Optimization Techniques · Machine Learning and Data Classification
参考文献 0
引用本文 3
按被引量排序,此处列出前 3 条