PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 4, 20241 citationsOpen Access

Neural Redshift: Random Networks are not Random Functions

View Full Paper
DTDamien TeneyIdiap Research InstituteANArmand Mihai NicolicioiuETH ZurichVHValentin N. HartmannUniversity of Stuttgart

Key Points

Key points are not available for this paper at this time.

Abstract

Our understanding of the generalization capabilities of neural networks (NNs) is still incomplete. Prevailing explanations are based on implicit biases of gradient descent (GD) but they cannot account for the capabilities of models from gradient-free methods nor the simplicity bias recently observed in untrained networks. This paper seeks other sources of generalization in NNs. Findings. To understand the inductive biases provided by architectures independently from GD, we examine untrained, random-weight networks. Even simple MLPs show strong inductive biases: uniform sampling in weight space yields a very biased distribution of functions in terms of complexity. But unlike common wisdom, NNs do not have an inherent "simplicity bias". This property depends on components such as ReLUs, residual connections, and layer normalizations. Alternative architectures can be built with a bias for any level of complexity. Transformers also inherit all these properties from their building blocks. Implications. We provide a fresh explanation for the success of deep learning independent from gradient-based training. It points at promising avenues for controlling the solutions implemented by trained models.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Teney et al. (2024) studied this question.

synapsesocial.com/papers/68e75ca9b6db6435876d4112https://doi.org/10.48550/arxiv.2403.02241
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1How Uniform Random Weights Induce Non-uniform Bias: Typical Interpolating Neural Networks Generalize with Narrow Teachers2024
  2. 2Simplicity Bias in Overparameterized Machine Learning2024 · 2 citations
  3. 3Bias of Stochastic Gradient Descent or the Architecture: Disentangling the Effects of Overparameterization of Neural Networks2024
  4. 4Can Biases in ImageNet Models Explain Generalization?2024
  5. 5Simplicity Bias of Two-Layer Networks beyond Linearly Separable Data2024