Deep neural networks have an inbuilt Occam's razor
Fuente:
arXiv
Salvato in:
| Autori principali: | Mingard, Chris, Rees, Henry, Valle-Pérez, Guillermo, Louis, Ard A. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
In-context learning and Occam's razor
di: Elmoznino, Eric, et al.
Pubblicazione: (2024)
di: Elmoznino, Eric, et al.
Pubblicazione: (2024)
Exploiting the equivalence between quantum neural networks and perceptrons
di: Mingard, Chris, et al.
Pubblicazione: (2024)
di: Mingard, Chris, et al.
Pubblicazione: (2024)
Feature learning is decoupled from generalization in high capacity neural networks
di: Göring, Niclas Alexander, et al.
Pubblicazione: (2025)
di: Göring, Niclas Alexander, et al.
Pubblicazione: (2025)
A simple mean field model of feature learning
di: Göring, Niclas, et al.
Pubblicazione: (2025)
di: Göring, Niclas, et al.
Pubblicazione: (2025)
Statistical learning theory and Occam's razor: The core argument
di: Sterkenburg, Tom F.
Pubblicazione: (2023)
di: Sterkenburg, Tom F.
Pubblicazione: (2023)
Sufficient Conditions for Stability of Minimum-Norm Interpolating Deep ReLU Networks
di: Harzli, Ouns El, et al.
Pubblicazione: (2026)
di: Harzli, Ouns El, et al.
Pubblicazione: (2026)
Occam's model: Selecting simpler representations for better transferability estimation
di: Singh, Prabhant, et al.
Pubblicazione: (2025)
di: Singh, Prabhant, et al.
Pubblicazione: (2025)
Occam's Razor for Self Supervised Learning: What is Sufficient to Learn Good Representations?
di: Ibrahim, Mark, et al.
Pubblicazione: (2024)
di: Ibrahim, Mark, et al.
Pubblicazione: (2024)
Planning in a recurrent neural network that plays Sokoban
di: Taufeeque, Mohammad, et al.
Pubblicazione: (2024)
di: Taufeeque, Mohammad, et al.
Pubblicazione: (2024)
Characterising the Inductive Biases of Neural Networks on Boolean Data
di: Mingard, Chris, et al.
Pubblicazione: (2025)
di: Mingard, Chris, et al.
Pubblicazione: (2025)
An exactly solvable model for emergence and scaling laws in the multitask sparse parity problem
di: Nam, Yoonsoo, et al.
Pubblicazione: (2024)
di: Nam, Yoonsoo, et al.
Pubblicazione: (2024)
In-Context Occam's Razor: How Transformers Prefer Simpler Hypotheses on the Fly
di: Deora, Puneesh, et al.
Pubblicazione: (2025)
di: Deora, Puneesh, et al.
Pubblicazione: (2025)
Deep-layered machines have a built-in Occam's razor
di: Fink, Thomas M. A.
Pubblicazione: (2026)
di: Fink, Thomas M. A.
Pubblicazione: (2026)
OccamLLM: Fast and Exact Language Model Arithmetic in a Single Step
di: Dugan, Owen, et al.
Pubblicazione: (2024)
di: Dugan, Owen, et al.
Pubblicazione: (2024)
On permutation-invariant neural networks
di: Kimura, Masanari, et al.
Pubblicazione: (2024)
di: Kimura, Masanari, et al.
Pubblicazione: (2024)
Sobolev acceleration for neural networks
di: Oh, Jong Kwon, et al.
Pubblicazione: (2025)
di: Oh, Jong Kwon, et al.
Pubblicazione: (2025)
Attention mechanisms in neural networks
di: Hays, Hasi
Pubblicazione: (2026)
di: Hays, Hasi
Pubblicazione: (2026)
Principles of Lipschitz continuity in neural networks
di: Luo, Róisín
Pubblicazione: (2026)
di: Luo, Róisín
Pubblicazione: (2026)
Linearity-based neural network compression
di: Dobler, Silas, et al.
Pubblicazione: (2025)
di: Dobler, Silas, et al.
Pubblicazione: (2025)
Applying graph neural network to SupplyGraph for supply chain network
di: Han, Kihwan
Pubblicazione: (2024)
di: Han, Kihwan
Pubblicazione: (2024)
LayerCollapse: Adaptive compression of neural networks
di: Shabgahi, Soheil Zibakhsh, et al.
Pubblicazione: (2023)
di: Shabgahi, Soheil Zibakhsh, et al.
Pubblicazione: (2023)
Understanding the dynamics of the frequency bias in neural networks
di: Molina, Juan, et al.
Pubblicazione: (2024)
di: Molina, Juan, et al.
Pubblicazione: (2024)
Graph neural networks informed locally by thermodynamics
di: Tierz, Alicia, et al.
Pubblicazione: (2024)
di: Tierz, Alicia, et al.
Pubblicazione: (2024)
Graph neural networks and non-commuting operators
di: Velasco, Mauricio, et al.
Pubblicazione: (2024)
di: Velasco, Mauricio, et al.
Pubblicazione: (2024)
Randomness and signal propagation in physics-informed neural networks (PINNs): A neural PDE perspective
di: Tucny, Jean-Michel, et al.
Pubblicazione: (2025)
di: Tucny, Jean-Michel, et al.
Pubblicazione: (2025)
Elimination-compensation pruning for fully-connected neural networks
di: Ballini, Enrico, et al.
Pubblicazione: (2026)
di: Ballini, Enrico, et al.
Pubblicazione: (2026)
Concealed Adversarial attacks on neural networks for sequential data
di: Sokerin, Petr, et al.
Pubblicazione: (2025)
di: Sokerin, Petr, et al.
Pubblicazione: (2025)
Dopamine-driven synaptic credit assignment in neural networks
di: Nambusubramaniyan, Saranraj, et al.
Pubblicazione: (2025)
di: Nambusubramaniyan, Saranraj, et al.
Pubblicazione: (2025)
Understanding polysemanticity in neural networks through coding theory
di: Marshall, Simon C., et al.
Pubblicazione: (2024)
di: Marshall, Simon C., et al.
Pubblicazione: (2024)
Variational autoencoder-based neural network model compression
di: Cheng, Liang, et al.
Pubblicazione: (2024)
di: Cheng, Liang, et al.
Pubblicazione: (2024)
Conditional computation in neural networks: principles and research trends
di: Scardapane, Simone, et al.
Pubblicazione: (2024)
di: Scardapane, Simone, et al.
Pubblicazione: (2024)
Utilizing Lyapunov Exponents in designing deep neural networks
di: Mittra, Tirthankar
Pubblicazione: (2024)
di: Mittra, Tirthankar
Pubblicazione: (2024)
Grey-informed neural network for time-series forecasting
di: Xie, Wanli, et al.
Pubblicazione: (2024)
di: Xie, Wanli, et al.
Pubblicazione: (2024)
A comparative analysis of a neural network with calculated weights and a neural network with random generation of weights based on the training dataset size
di: Geidarov, Polad
Pubblicazione: (2025)
di: Geidarov, Polad
Pubblicazione: (2025)
Self-adaptive weights based on balanced residual decay rate for physics-informed neural networks and deep operator networks
di: Chen, Wenqian, et al.
Pubblicazione: (2024)
di: Chen, Wenqian, et al.
Pubblicazione: (2024)
EEG-Bench: A Benchmark for EEG Foundation Models in Clinical Applications
di: Kastrati, Ard, et al.
Pubblicazione: (2025)
di: Kastrati, Ard, et al.
Pubblicazione: (2025)
Understanding the learned look-ahead behavior of chess neural networks
di: Cruz, Diogo
Pubblicazione: (2025)
di: Cruz, Diogo
Pubblicazione: (2025)
Confidence-gated training for efficient early-exit neural networks
di: Mokssit, Saad, et al.
Pubblicazione: (2025)
di: Mokssit, Saad, et al.
Pubblicazione: (2025)
Adaptive multiple optimal learning factors for neural network training
di: Challagundla, Jeshwanth
Pubblicazione: (2024)
di: Challagundla, Jeshwanth
Pubblicazione: (2024)
Stochastic stem bucking using mixture density neural networks
di: Schmiedel, Simon
Pubblicazione: (2024)
di: Schmiedel, Simon
Pubblicazione: (2024)
Documenti analoghi
-
In-context learning and Occam's razor
di: Elmoznino, Eric, et al.
Pubblicazione: (2024) -
Exploiting the equivalence between quantum neural networks and perceptrons
di: Mingard, Chris, et al.
Pubblicazione: (2024) -
Feature learning is decoupled from generalization in high capacity neural networks
di: Göring, Niclas Alexander, et al.
Pubblicazione: (2025) -
A simple mean field model of feature learning
di: Göring, Niclas, et al.
Pubblicazione: (2025) -
Statistical learning theory and Occam's razor: The core argument
di: Sterkenburg, Tom F.
Pubblicazione: (2023)