Representation Learning on a Random Lattice
Fuente:
arXiv
Guardado en:
| Autor principal: | Brill, Aryeh |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Neural Scaling Laws Rooted in the Data Distribution
por: Brill, Ari
Publicado: (2024)
por: Brill, Ari
Publicado: (2024)
Lattice Protein Folding with Variational Annealing
por: Khandoker, Shoummo Ahsan, et al.
Publicado: (2025)
por: Khandoker, Shoummo Ahsan, et al.
Publicado: (2025)
Towards Worst-Case Guarantees with Scale-Aware Interpretability
por: Greenspan, Lauren, et al.
Publicado: (2026)
por: Greenspan, Lauren, et al.
Publicado: (2026)
Spectral Dynamics in Deep Networks: Feature Learning, Outlier Escape, and Learning Rate Transfer
por: Lauditi, Clarissa, et al.
Publicado: (2026)
por: Lauditi, Clarissa, et al.
Publicado: (2026)
Learning Shrinks the Hard Tail: Training-Dependent Inference Scaling in a Solvable Linear Model
por: Levi, Noam
Publicado: (2026)
por: Levi, Noam
Publicado: (2026)
Applications of Statistical Field Theory in Deep Learning
por: Ringel, Zohar, et al.
Publicado: (2025)
por: Ringel, Zohar, et al.
Publicado: (2025)
Grokking vs. Learning: Same Features, Different Encodings
por: Manning-Coe, Dmitry, et al.
Publicado: (2025)
por: Manning-Coe, Dmitry, et al.
Publicado: (2025)
Parameter Symmetry Potentially Unifies Deep Learning Theory
por: Ziyin, Liu, et al.
Publicado: (2025)
por: Ziyin, Liu, et al.
Publicado: (2025)
Quantifying Hyperparameter Transfer and the Importance of Embedding Layer Learning Rate
por: Kalra, Dayal Singh, et al.
Publicado: (2026)
por: Kalra, Dayal Singh, et al.
Publicado: (2026)
Scaling Laws and Spectra of Shallow Neural Networks in the Feature Learning Regime
por: Defilippis, Leonardo, et al.
Publicado: (2025)
por: Defilippis, Leonardo, et al.
Publicado: (2025)
A Geometric Perspective on the Difficulties of Learning GNN-based SAT Solvers
por: Skenderi, Geri
Publicado: (2025)
por: Skenderi, Geri
Publicado: (2025)
Enhancing Noise-Robust Losses for Large-Scale Noisy Data Learning
por: Staats, Max, et al.
Publicado: (2023)
por: Staats, Max, et al.
Publicado: (2023)
Connecting NTK and NNGP: A Unified Theoretical Framework for Wide Neural Network Learning Dynamics
por: Avidan, Yehonatan, et al.
Publicado: (2023)
por: Avidan, Yehonatan, et al.
Publicado: (2023)
More Bang for the Buck: Improving the Inference of Large Language Models at a Fixed Budget using Reset and Discard (ReD)
por: Meir, Sagi, et al.
Publicado: (2026)
por: Meir, Sagi, et al.
Publicado: (2026)
Predictive Coding Networks and Inference Learning: Tutorial and Survey
por: van Zwol, Björn, et al.
Publicado: (2024)
por: van Zwol, Björn, et al.
Publicado: (2024)
Generalization through variance: how noise shapes inductive biases in diffusion models
por: Vastola, John J.
Publicado: (2025)
por: Vastola, John J.
Publicado: (2025)
Identifying internal patterns in (1+1)-dimensional directed percolation using neural networks
por: Parkhomenko, Danil, et al.
Publicado: (2025)
por: Parkhomenko, Danil, et al.
Publicado: (2025)
A Spin Glass Characterization of Neural Networks
por: Li, Jun
Publicado: (2025)
por: Li, Jun
Publicado: (2025)
Towards Distributed Neural Architectures
por: Cowsik, Aditya, et al.
Publicado: (2025)
por: Cowsik, Aditya, et al.
Publicado: (2025)
A Scalable Measure of Loss Landscape Curvature for Analyzing the Training Dynamics of LLMs
por: Kalra, Dayal Singh, et al.
Publicado: (2026)
por: Kalra, Dayal Singh, et al.
Publicado: (2026)
A method for quantifying the generalization capabilities of generative models for solving Ising models
por: Ma, Qunlong, et al.
Publicado: (2024)
por: Ma, Qunlong, et al.
Publicado: (2024)
KAN: Kolmogorov-Arnold Networks
por: Liu, Ziming, et al.
Publicado: (2024)
por: Liu, Ziming, et al.
Publicado: (2024)
How Do Transformers "Do" Physics? Investigating the Simple Harmonic Oscillator
por: Kantamneni, Subhash, et al.
Publicado: (2024)
por: Kantamneni, Subhash, et al.
Publicado: (2024)
On the origin of neural scaling laws: from random graphs to natural language
por: Barkeshli, Maissam, et al.
Publicado: (2026)
por: Barkeshli, Maissam, et al.
Publicado: (2026)
The Persian Rug: solving toy models of superposition using large-scale symmetries
por: Cowsik, Aditya, et al.
Publicado: (2024)
por: Cowsik, Aditya, et al.
Publicado: (2024)
Smooth Kolmogorov Arnold networks enabling structural knowledge representation
por: Samadi, Moein E., et al.
Publicado: (2024)
por: Samadi, Moein E., et al.
Publicado: (2024)
Estimating Global Input Relevance and Enforcing Sparse Representations with a Scalable Spectral Neural Network Approach
por: Chicchi, Lorenzo, et al.
Publicado: (2024)
por: Chicchi, Lorenzo, et al.
Publicado: (2024)
A Minimal Model of Representation Collapse: Frustration, Stop-Gradient, and Dynamics
por: Yao, Louie Hong, et al.
Publicado: (2026)
por: Yao, Louie Hong, et al.
Publicado: (2026)
Predictive Coding Graphs are a Superset of Feedforward Neural Networks
por: van Zwol, Björn
Publicado: (2026)
por: van Zwol, Björn
Publicado: (2026)
Nature-Inspired Local Propagation
por: Betti, Alessandro, et al.
Publicado: (2024)
por: Betti, Alessandro, et al.
Publicado: (2024)
Approximation Theory for Neural Networks: Old and New
por: Mukherjee, Soumendu Sundar, et al.
Publicado: (2026)
por: Mukherjee, Soumendu Sundar, et al.
Publicado: (2026)
Preisach Attention: A Hysteretic Model of Sequential Memory
por: Frydrych, Piotr
Publicado: (2026)
por: Frydrych, Piotr
Publicado: (2026)
Theory of Optimal Learning Rate Schedules and Scaling Laws for a Random Feature Model
por: Bordelon, Blake, et al.
Publicado: (2026)
por: Bordelon, Blake, et al.
Publicado: (2026)
Graph Learning Metallic Glass Discovery from Wikipedia
por: Ouyang, K. -C., et al.
Publicado: (2025)
por: Ouyang, K. -C., et al.
Publicado: (2025)
Disordered Dynamics in High Dimensions: Connections to Random Matrices and Machine Learning
por: Bordelon, Blake, et al.
Publicado: (2026)
por: Bordelon, Blake, et al.
Publicado: (2026)
High-Dimensional Learning Dynamics of Quantized Models with Straight-Through Estimator
por: Ichikawa, Yuma, et al.
Publicado: (2025)
por: Ichikawa, Yuma, et al.
Publicado: (2025)
How Deep Networks Learn Sparse and Hierarchical Data: the Sparse Random Hierarchy Model
por: Tomasini, Umberto, et al.
Publicado: (2024)
por: Tomasini, Umberto, et al.
Publicado: (2024)
Scaling Laws and Representation Learning in Simple Hierarchical Languages: Transformers vs. Convolutional Architectures
por: Cagnetta, Francesco, et al.
Publicado: (2025)
por: Cagnetta, Francesco, et al.
Publicado: (2025)
Random features and polynomial rules
por: Aguirre-López, Fabián, et al.
Publicado: (2024)
por: Aguirre-López, Fabián, et al.
Publicado: (2024)
Formation of Representations in Neural Networks
por: Ziyin, Liu, et al.
Publicado: (2024)
por: Ziyin, Liu, et al.
Publicado: (2024)
Ejemplares similares
-
Neural Scaling Laws Rooted in the Data Distribution
por: Brill, Ari
Publicado: (2024) -
Lattice Protein Folding with Variational Annealing
por: Khandoker, Shoummo Ahsan, et al.
Publicado: (2025) -
Towards Worst-Case Guarantees with Scale-Aware Interpretability
por: Greenspan, Lauren, et al.
Publicado: (2026) -
Spectral Dynamics in Deep Networks: Feature Learning, Outlier Escape, and Learning Rate Transfer
por: Lauditi, Clarissa, et al.
Publicado: (2026) -
Learning Shrinks the Hard Tail: Training-Dependent Inference Scaling in a Solvable Linear Model
por: Levi, Noam
Publicado: (2026)