The Graphon Limit Hypothesis: Understanding Neural Network Pruning via Infinite Width Analysis
Fuente:
arXiv
Saved in:
| Main Authors: | Pham, Hoang, Ta, The-Anh, Jacobs, Tom, Burkholz, Rebekka, Tran-Thanh, Long |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Pruning at Initialisation through the lens of Graphon Limit: Convergence, Expressivity, and Generalisation
by: Pham, Hoang, et al.
Published: (2026)
by: Pham, Hoang, et al.
Published: (2026)
Pay Attention to Small Weights
by: Zhou, Chao, et al.
Published: (2025)
by: Zhou, Chao, et al.
Published: (2025)
Flatness-aware Sequential Learning Generates Resilient Backdoors
by: Pham, Hoang, et al.
Published: (2024)
by: Pham, Hoang, et al.
Published: (2024)
Mask in the Mirror: Implicit Sparsification
by: Jacobs, Tom, et al.
Published: (2024)
by: Jacobs, Tom, et al.
Published: (2024)
Revisiting LARS for Large Batch Training Generalization of Neural Networks
by: Do, Khoi, et al.
Published: (2023)
by: Do, Khoi, et al.
Published: (2023)
The Neural Pruning Law Hypothesis
by: Barbulescu, Eugen, et al.
Published: (2025)
by: Barbulescu, Eugen, et al.
Published: (2025)
HORST: Composing Optimizer Geometries for Sparse Transformer Training
by: Jacobs, Tom, et al.
Published: (2026)
by: Jacobs, Tom, et al.
Published: (2026)
Mirror, Mirror of the Flow: How Does Regularization Shape Implicit Bias?
by: Jacobs, Tom, et al.
Published: (2025)
by: Jacobs, Tom, et al.
Published: (2025)
Never Saddle for Reparameterized Steepest Descent as Mirror Flow
by: Jacobs, Tom, et al.
Published: (2026)
by: Jacobs, Tom, et al.
Published: (2026)
On the Infinite Width and Depth Limits of Predictive Coding Networks
by: Innocenti, Francesco, et al.
Published: (2026)
by: Innocenti, Francesco, et al.
Published: (2026)
Efficient Bilevel Optimization for Meta Label Correction in Noisy Label Learning
by: Nguyen, Ba Hoang Anh, et al.
Published: (2026)
by: Nguyen, Ba Hoang Anh, et al.
Published: (2026)
Adaptive Width Neural Networks
by: Errica, Federico, et al.
Published: (2025)
by: Errica, Federico, et al.
Published: (2025)
Improving Graph Convolutional Networks with Transformer Layer in social-based items recommendation
by: Hoang, Thi Linh, et al.
Published: (2024)
by: Hoang, Thi Linh, et al.
Published: (2024)
Tree-Sliced Wasserstein Distance with Nonlinear Projection
by: Tran, Thanh, et al.
Published: (2025)
by: Tran, Thanh, et al.
Published: (2025)
Constructing Decision Trees from Data Streams
by: Pham, Huy, et al.
Published: (2024)
by: Pham, Huy, et al.
Published: (2024)
Bridging Domains through Subspace-Aware Model Merging
by: Chaves, Levy, et al.
Published: (2026)
by: Chaves, Levy, et al.
Published: (2026)
Understanding the Countably Infinite: Neural Network Models of the Successor Function and its Acquisition
by: Gupta, Vima, et al.
Published: (2023)
by: Gupta, Vima, et al.
Published: (2023)
Global Convergence and Rich Feature Learning in $L$-Layer Infinite-Width Neural Networks under $μ$P Parametrization
by: Chen, Zixiang, et al.
Published: (2025)
by: Chen, Zixiang, et al.
Published: (2025)
Sign-In to the Lottery: Reparameterizing Sparse Training From Scratch
by: Gadhikar, Advait, et al.
Published: (2025)
by: Gadhikar, Advait, et al.
Published: (2025)
Hyperbolic Aware Minimization: Implicit Bias for Sparsity
by: Jacobs, Tom, et al.
Published: (2025)
by: Jacobs, Tom, et al.
Published: (2025)
Virtual Width Networks
by: Seed, et al.
Published: (2025)
by: Seed, et al.
Published: (2025)
MP-PINN: A Multi-Phase Physics-Informed Neural Network for Epidemic Forecasting
by: Nguyen, Thang, et al.
Published: (2024)
by: Nguyen, Thang, et al.
Published: (2024)
The WidthWall: A Strict Expressivity Hierarchy for Hypergraph Neural Networks
by: Jiang, Fengqing, et al.
Published: (2026)
by: Jiang, Fengqing, et al.
Published: (2026)
Tree-Sliced Wasserstein Distance: A Geometric Perspective
by: Tran, Viet-Hoang, et al.
Published: (2024)
by: Tran, Viet-Hoang, et al.
Published: (2024)
Don't Trust Stubborn Neighbors: A Security Framework for Agentic Networks
by: Abedini, Samira, et al.
Published: (2026)
by: Abedini, Samira, et al.
Published: (2026)
Geometric Limits of Knowledge Distillation: A Minimum-Width Theorem via Superposition Theory
by: Sarkar, Nilesh, et al.
Published: (2026)
by: Sarkar, Nilesh, et al.
Published: (2026)
Infinite Width Models That Work: Why Feature Learning Doesn't Matter as Much as You Think
by: Sernau, Luke
Published: (2024)
by: Sernau, Luke
Published: (2024)
NNiT: Width-Agnostic Neural Network Generation with Structurally Aligned Weight Spaces
by: Kim, Jiwoo, et al.
Published: (2026)
by: Kim, Jiwoo, et al.
Published: (2026)
Spectral Graph Pruning Against Over-Squashing and Over-Smoothing
by: Jamadandi, Adarsh, et al.
Published: (2024)
by: Jamadandi, Adarsh, et al.
Published: (2024)
Spherical Tree-Sliced Wasserstein Distance
by: Tran, Viet-Hoang, et al.
Published: (2025)
by: Tran, Viet-Hoang, et al.
Published: (2025)
Distance-Based Tree-Sliced Wasserstein Distance
by: Tran, Hoang V., et al.
Published: (2025)
by: Tran, Hoang V., et al.
Published: (2025)
A Gap Between the Gaussian RKHS and Neural Networks: An Infinite-Center Asymptotic Analysis
by: Kumar, Akash, et al.
Published: (2025)
by: Kumar, Akash, et al.
Published: (2025)
Identifying the Best Arm in the Presence of Global Environment Shifts
by: Srisawad, Phurinut, et al.
Published: (2024)
by: Srisawad, Phurinut, et al.
Published: (2024)
Learning to Stop Overthinking at Test Time
by: Bao, Hieu Tran, et al.
Published: (2025)
by: Bao, Hieu Tran, et al.
Published: (2025)
PruneSymNet: A Symbolic Neural Network and Pruning Algorithm for Symbolic Regression
by: Wu, Min, et al.
Published: (2024)
by: Wu, Min, et al.
Published: (2024)
Using Synthetic Data to estimate the True Error is theoretically and practically doable
by: Thanh, Hai Hoang, et al.
Published: (2025)
by: Thanh, Hai Hoang, et al.
Published: (2025)
A Dual Certificate Approach to Sparsity in Infinite-Width Shallow Neural Networks
by: Del Grande, Leonardo, et al.
Published: (2026)
by: Del Grande, Leonardo, et al.
Published: (2026)
Beyond the Frontier: Stochastic Backtracking for Efficient Test-Time Scaling
by: Tran, Dao, et al.
Published: (2026)
by: Tran, Dao, et al.
Published: (2026)
Neural ODE Transformers: Analyzing Internal Dynamics and Adaptive Fine-tuning
by: Tong, Anh, et al.
Published: (2025)
by: Tong, Anh, et al.
Published: (2025)
Infinite Width Limits of Self Supervised Neural Networks
by: Fleissner, Maximilian, et al.
Published: (2024)
by: Fleissner, Maximilian, et al.
Published: (2024)
Similar Items
-
Pruning at Initialisation through the lens of Graphon Limit: Convergence, Expressivity, and Generalisation
by: Pham, Hoang, et al.
Published: (2026) -
Pay Attention to Small Weights
by: Zhou, Chao, et al.
Published: (2025) -
Flatness-aware Sequential Learning Generates Resilient Backdoors
by: Pham, Hoang, et al.
Published: (2024) -
Mask in the Mirror: Implicit Sparsification
by: Jacobs, Tom, et al.
Published: (2024) -
Revisiting LARS for Large Batch Training Generalization of Neural Networks
by: Do, Khoi, et al.
Published: (2023)