Beyond IID weights: sparse and low-rank deep Neural Networks are also Gaussian Processes
Fuente:
arXiv
Saved in:
| Main Authors: | Nait-Saada, Thiziri, Naderi, Alireza, Tanner, Jared |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mind the Gap: a Spectral Analysis of Rank Collapse and Signal Propagation in Attention Layers
by: Saada, Thiziri Nait, et al.
Published: (2024)
by: Saada, Thiziri Nait, et al.
Published: (2024)
A simple proof of almost sure convergence for the largest singular value of a product of Gaussian matrices
by: Saada, Thiziri Nait, et al.
Published: (2024)
by: Saada, Thiziri Nait, et al.
Published: (2024)
The Data-Quality Illusion: Rethinking Classifier-Based Quality Filtering for LLM Pretraining
by: Saada, Thiziri Nait, et al.
Published: (2025)
by: Saada, Thiziri Nait, et al.
Published: (2025)
GraphEval: A Knowledge-Graph Based LLM Hallucination Evaluation Framework
by: Sansford, Hannah, et al.
Published: (2024)
by: Sansford, Hannah, et al.
Published: (2024)
Theory of Minimal Weight Perturbations in Deep Networks and its Applications for Low-Rank Activated Backdoor Attacks
by: Evans, Bethan, et al.
Published: (2026)
by: Evans, Bethan, et al.
Published: (2026)
SLTrain: a sparse plus low-rank approach for parameter and memory efficient pretraining
by: Han, Andi, et al.
Published: (2024)
by: Han, Andi, et al.
Published: (2024)
Deep Neural Network Initialization with Sparsity Inducing Activations
by: Price, Ilan, et al.
Published: (2024)
by: Price, Ilan, et al.
Published: (2024)
How Controlling the Variance can Improve Training Stability of Sparsely Activated DNNs and CNNs
by: Dent, Emily, et al.
Published: (2026)
by: Dent, Emily, et al.
Published: (2026)
CARMIL: Context-Aware Regularization on Multiple Instance Learning models for Whole Slide Images
by: Saada, Thiziri Nait, et al.
Published: (2024)
by: Saada, Thiziri Nait, et al.
Published: (2024)
PCA recovery thresholds in low-rank matrix inference with sparse noise
by: Adomaityte, Urte, et al.
Published: (2025)
by: Adomaityte, Urte, et al.
Published: (2025)
Generalization Bounds for Rank-sparse Neural Networks
by: Ledent, Antoine, et al.
Published: (2025)
by: Ledent, Antoine, et al.
Published: (2025)
Beyond IID: data-driven decision-making in heterogeneous environments
by: Besbes, Omar, et al.
Published: (2022)
by: Besbes, Omar, et al.
Published: (2022)
Revisiting Generalization Measures Beyond IID: An Empirical Study under Distributional Shift
by: Nakai, Sora, et al.
Published: (2026)
by: Nakai, Sora, et al.
Published: (2026)
Reinforcement Learning from LLM Feedback to Counteract Goal Misgeneralization
by: Barj, Houda Nait El, et al.
Published: (2024)
by: Barj, Houda Nait El, et al.
Published: (2024)
Tighter sparse variational Gaussian processes
by: Bui, Thang D., et al.
Published: (2025)
by: Bui, Thang D., et al.
Published: (2025)
Sparse Gaussian Neural Processes
by: Rochussen, Tommy, et al.
Published: (2025)
by: Rochussen, Tommy, et al.
Published: (2025)
Vecchia Gaussian Process Ensembles on Internal Representations of Deep Neural Networks
by: Jimenez, Felix, et al.
Published: (2023)
by: Jimenez, Felix, et al.
Published: (2023)
Gaussian Process Regression -- Neural Network Hybrid with Optimized Redundant Coordinates
by: Manzhos, Sergei, et al.
Published: (2025)
by: Manzhos, Sergei, et al.
Published: (2025)
Random ReLU Neural Networks as Non-Gaussian Processes
by: Parhi, Rahul, et al.
Published: (2024)
by: Parhi, Rahul, et al.
Published: (2024)
Motif distribution and function of sparse deep neural networks
by: Zahn, Olivia T., et al.
Published: (2024)
by: Zahn, Olivia T., et al.
Published: (2024)
On efficiently computable functions, deep networks and sparse compositionality
by: Poggio, Tomaso
Published: (2025)
by: Poggio, Tomaso
Published: (2025)
Approximate Gaussianity Beyond Initialisation in Neural Networks
by: Hirst, Edward, et al.
Published: (2025)
by: Hirst, Edward, et al.
Published: (2025)
A Gaussian Process View on Observation Noise and Initialization in Wide Neural Networks
by: Calvo-Ordoñez, Sergio, et al.
Published: (2025)
by: Calvo-Ordoñez, Sergio, et al.
Published: (2025)
Revisiting the Equivalence of Bayesian Neural Networks and Gaussian Processes: On the Importance of Learning Activations
by: Sendera, Marcin, et al.
Published: (2024)
by: Sendera, Marcin, et al.
Published: (2024)
Label Propagation Training Schemes for Physics-Informed Neural Networks and Gaussian Processes
by: Zhong, Ming, et al.
Published: (2024)
by: Zhong, Ming, et al.
Published: (2024)
Understanding Federated Learning from IID to Non-IID dataset: An Experimental Study
by: Seo, Jungwon, et al.
Published: (2025)
by: Seo, Jungwon, et al.
Published: (2025)
Low-rank computation of the posterior mean in Multi-Output Gaussian Processes
by: Esche, Sebastian, et al.
Published: (2025)
by: Esche, Sebastian, et al.
Published: (2025)
Probabilistic Neural Networks (PNNs) with t-Distributed Outputs: Adaptive Prediction Intervals Beyond Gaussian Assumptions
by: Pourkamali-Anaraki, Farhad
Published: (2025)
by: Pourkamali-Anaraki, Farhad
Published: (2025)
Evaluating Federated Kolmogorov-Arnold Networks on Non-IID Data
by: Sasse, Arthur Mendonça, et al.
Published: (2024)
by: Sasse, Arthur Mendonça, et al.
Published: (2024)
Three Costs of Amortizing Gaussian Process Inference with Neural Processes
by: Young, Robin
Published: (2026)
by: Young, Robin
Published: (2026)
Decentralized Learning Strategies for Estimation Error Minimization with Graph Neural Networks
by: Chen, Xingran, et al.
Published: (2026)
by: Chen, Xingran, et al.
Published: (2026)
A Framework for Nonstationary Gaussian Processes with Neural Network Parameters
by: James, Zachary, et al.
Published: (2025)
by: James, Zachary, et al.
Published: (2025)
Federated Spectral Graph Transformers Meet Neural Ordinary Differential Equations for Non-IID Graphs
by: Gurumurthy, Kishan, et al.
Published: (2025)
by: Gurumurthy, Kishan, et al.
Published: (2025)
Model Merging by Output-Space Projection
by: Evans, Bethan, et al.
Published: (2026)
by: Evans, Bethan, et al.
Published: (2026)
Beyond Feature Fusion: Contextual Bayesian PEFT for Multimodal Uncertainty Estimation
by: Naderi, Habibeh, et al.
Published: (2026)
by: Naderi, Habibeh, et al.
Published: (2026)
Enhancing convolutional neural network generalizability via low-rank weight approximation
by: Gao, Chenyin, et al.
Published: (2022)
by: Gao, Chenyin, et al.
Published: (2022)
Misclassification bounds for PAC-Bayesian sparse deep learning
by: Mai, The Tien
Published: (2024)
by: Mai, The Tien
Published: (2024)
Gaussian Process Kolmogorov-Arnold Networks
by: Chen, Andrew Siyuan
Published: (2024)
by: Chen, Andrew Siyuan
Published: (2024)
Learning Beyond the Gaussian Data: Learning Dynamics of Neural Networks on an Expressive and Cumulant-Controllable Data Model
by: Ure, Onat, et al.
Published: (2026)
by: Ure, Onat, et al.
Published: (2026)
DDP-SA: Scalable Privacy-Preserving Federated Learning via Distributed Differential Privacy and Secure Aggregation
by: Wei, Wenjing, et al.
Published: (2026)
by: Wei, Wenjing, et al.
Published: (2026)
Similar Items
-
Mind the Gap: a Spectral Analysis of Rank Collapse and Signal Propagation in Attention Layers
by: Saada, Thiziri Nait, et al.
Published: (2024) -
A simple proof of almost sure convergence for the largest singular value of a product of Gaussian matrices
by: Saada, Thiziri Nait, et al.
Published: (2024) -
The Data-Quality Illusion: Rethinking Classifier-Based Quality Filtering for LLM Pretraining
by: Saada, Thiziri Nait, et al.
Published: (2025) -
GraphEval: A Knowledge-Graph Based LLM Hallucination Evaluation Framework
by: Sansford, Hannah, et al.
Published: (2024) -
Theory of Minimal Weight Perturbations in Deep Networks and its Applications for Low-Rank Activated Backdoor Attacks
by: Evans, Bethan, et al.
Published: (2026)