Theoretical Compression Bounds for Wide Multilayer Perceptrons
Fuente:
arXiv
Saved in:
| Main Authors: | Cheairi, Houssam El, Gamarnik, David, Mazumder, Rahul |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Classification by Separating Hypersurfaces: An Entropic Approach
by: Arratia, Argimiro, et al.
Published: (2025)
by: Arratia, Argimiro, et al.
Published: (2025)
Randomized Matrix Sketching for Neural Network Training and Gradient Monitoring
by: Antil, Harbir, et al.
Published: (2025)
by: Antil, Harbir, et al.
Published: (2025)
Random feature-based double Vovk-Azoury-Warmuth algorithm for online multi-kernel learning
by: Rokhlin, Dmitry B., et al.
Published: (2025)
by: Rokhlin, Dmitry B., et al.
Published: (2025)
A hierarchical Vovk-Azoury-Warmuth forecaster with discounting for online regression in RKHS
by: Rokhlin, Dmitry B.
Published: (2025)
by: Rokhlin, Dmitry B.
Published: (2025)
Distributed Sparse Linear Regression under Communication Constraints
by: Fonseca, Rodney, et al.
Published: (2023)
by: Fonseca, Rodney, et al.
Published: (2023)
Ballistic Convergence in Hit-and-Run Monte Carlo and a Coordinate-free Randomized Kaczmarz Algorithm
by: Bou-Rabee, Nawaf, et al.
Published: (2024)
by: Bou-Rabee, Nawaf, et al.
Published: (2024)
Stealth edits to large language models
by: Sutton, Oliver J., et al.
Published: (2024)
by: Sutton, Oliver J., et al.
Published: (2024)
A Generalization Bound for a Family of Implicit Networks
by: Fung, Samy Wu, et al.
Published: (2024)
by: Fung, Samy Wu, et al.
Published: (2024)
Provable Post-Training Quantization: Theoretical Analysis of OPTQ and Qronos
by: Zhang, Haoyu, et al.
Published: (2025)
by: Zhang, Haoyu, et al.
Published: (2025)
ParaRNN: Unlocking Parallel Training of Nonlinear RNNs for Large Language Models
by: Danieli, Federico, et al.
Published: (2025)
by: Danieli, Federico, et al.
Published: (2025)
Super-fast Rates of Convergence for Neural Network Classifiers under the Hard Margin Condition
by: Tepakbong, Nathanael, et al.
Published: (2025)
by: Tepakbong, Nathanael, et al.
Published: (2025)
On the Sample Complexity of One Hidden Layer Networks with Equivariance, Locality and Weight Sharing
by: Behboodi, Arash, et al.
Published: (2024)
by: Behboodi, Arash, et al.
Published: (2024)
Credal and Interval Deep Evidential Classifications
by: Caprio, Michele, et al.
Published: (2025)
by: Caprio, Michele, et al.
Published: (2025)
The General Theory of Localization Methods
by: Song, Congwei
Published: (2026)
by: Song, Congwei
Published: (2026)
Ranking Perspective for Tree-based Methods with Applications to Symbolic Feature Selection
by: Luo, Hengrui, et al.
Published: (2024)
by: Luo, Hengrui, et al.
Published: (2024)
Exascale Multi-Task Graph Foundation Models for Imbalanced, Multi-Fidelity Atomistic Data
by: Pasini, Massimiliano Lupo, et al.
Published: (2026)
by: Pasini, Massimiliano Lupo, et al.
Published: (2026)
Ergodic Generative Flows
by: Brunswic, Leo Maxime, et al.
Published: (2025)
by: Brunswic, Leo Maxime, et al.
Published: (2025)
Multipole Semantic Attention: A Fast Approximation of Softmax Attention for Pretraining
by: Mitchell, Rupert, et al.
Published: (2025)
by: Mitchell, Rupert, et al.
Published: (2025)
Robust Tensor CUR Decompositions: Rapid Low-Tucker-Rank Tensor Recovery with Sparse Corruption
by: Cai, HanQin, et al.
Published: (2023)
by: Cai, HanQin, et al.
Published: (2023)
Towards a Rigorous Understanding of the Population Dynamics of the NSGA-III: Tight Runtime Bounds
by: Opris, Andre
Published: (2025)
by: Opris, Andre
Published: (2025)
Directional Convergence, Benign Overfitting of Gradient Descent in leaky ReLU two-layer Neural Networks
by: Hashimoto, Ichiro
Published: (2025)
by: Hashimoto, Ichiro
Published: (2025)
Asymptotics of Stochastic Gradient Descent with Dropout Regularization in Linear Models
by: Li, Jiaqi, et al.
Published: (2024)
by: Li, Jiaqi, et al.
Published: (2024)
CompressedScaffnew: The First Theoretical Double Acceleration of Communication from Local Training and Compression in Distributed Optimization
by: Condat, Laurent, et al.
Published: (2022)
by: Condat, Laurent, et al.
Published: (2022)
Nonparametric estimation of a factorizable density using diffusion models
by: Kwon, Hyeok Kyu, et al.
Published: (2025)
by: Kwon, Hyeok Kyu, et al.
Published: (2025)
Algorithmic Universality, Low-Degree Polynomials, and Max-Cut in Sparse Random Graphs
by: Cheairi, Houssam El, et al.
Published: (2024)
by: Cheairi, Houssam El, et al.
Published: (2024)
Model Parallel Training and Transfer Learning for Convolutional Neural Networks by Domain Decomposition
by: Klawonn, Axel, et al.
Published: (2024)
by: Klawonn, Axel, et al.
Published: (2024)
Domain-decomposed image classification algorithms using linear discriminant analysis and convolutional neural networks
by: Klawonn, Axel, et al.
Published: (2024)
by: Klawonn, Axel, et al.
Published: (2024)
Estimating Coverage in Streams via a Modified CVM Method
by: Hernandez-Suarez, Carlos
Published: (2025)
by: Hernandez-Suarez, Carlos
Published: (2025)
An operator learning perspective on parameter-to-observable maps
by: Huang, Daniel Zhengyu, et al.
Published: (2024)
by: Huang, Daniel Zhengyu, et al.
Published: (2024)
Sub-Token Routing in LoRA for Adaptation and Query-Aware KV Compression
by: Jiang, Wei, et al.
Published: (2026)
by: Jiang, Wei, et al.
Published: (2026)
Adversarial Subspace Generation for Outlier Detection in High-Dimensional Data
by: Cribeiro-Ramallo, Jose, et al.
Published: (2025)
by: Cribeiro-Ramallo, Jose, et al.
Published: (2025)
A Domain Decomposition-Based CNN-DNN Architecture for Model Parallel Training Applied to Image Recognition Problems
by: Klawonn, Axel, et al.
Published: (2023)
by: Klawonn, Axel, et al.
Published: (2023)
Error Bounds for Learning with Vector-Valued Random Features
by: Lanthaler, Samuel, et al.
Published: (2023)
by: Lanthaler, Samuel, et al.
Published: (2023)
Functional Similarity Metric for Neural Networks: Overcoming Parametric Ambiguity via Activation Region Analysis
by: Hennadii, Kutomanov
Published: (2026)
by: Hennadii, Kutomanov
Published: (2026)
Subgroups of Cyclically Amalgamated Free Products
by: Kreuzer, Martin, et al.
Published: (2025)
by: Kreuzer, Martin, et al.
Published: (2025)
Adaptive monotonicity testing in sublinear time
by: Li, Housen, et al.
Published: (2025)
by: Li, Housen, et al.
Published: (2025)
Exact recovery for seeded graph matching
by: Fraiman, Nicolas, et al.
Published: (2026)
by: Fraiman, Nicolas, et al.
Published: (2026)
Progressive Feedforward Collapse of ResNet Training
by: Wang, Sicong, et al.
Published: (2024)
by: Wang, Sicong, et al.
Published: (2024)
Shortest Paths without a Map, but with an Entropic Regularizer
by: Bubeck, Sébastien, et al.
Published: (2022)
by: Bubeck, Sébastien, et al.
Published: (2022)
Optimal In-context Adaptivity and Distributional Robustness of Transformers
by: Ma, Tianyi, et al.
Published: (2025)
by: Ma, Tianyi, et al.
Published: (2025)
Similar Items
-
Classification by Separating Hypersurfaces: An Entropic Approach
by: Arratia, Argimiro, et al.
Published: (2025) -
Randomized Matrix Sketching for Neural Network Training and Gradient Monitoring
by: Antil, Harbir, et al.
Published: (2025) -
Random feature-based double Vovk-Azoury-Warmuth algorithm for online multi-kernel learning
by: Rokhlin, Dmitry B., et al.
Published: (2025) -
A hierarchical Vovk-Azoury-Warmuth forecaster with discounting for online regression in RKHS
by: Rokhlin, Dmitry B.
Published: (2025) -
Distributed Sparse Linear Regression under Communication Constraints
by: Fonseca, Rodney, et al.
Published: (2023)