Saved in:
| Main Authors: | Ghosh, Avrajit, Cong, Bai, Yokota, Rio, Ravishankar, Saiprasad, Wang, Rongrong, Tao, Molei, Khan, Mohammad Emtiyaz, Möllenhoff, Thomas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2506.12903 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Dynamics of Deep Linear Networks Beyond the Edge of Stability
by: Ghosh, Avrajit, et al.
Published: (2025)
by: Ghosh, Avrajit, et al.
Published: (2025)
Improving LoRA with Variational Learning
by: Cong, Bai, et al.
Published: (2025)
by: Cong, Bai, et al.
Published: (2025)
Variational Low-Rank Adaptation Using IVON
by: Cong, Bai, et al.
Published: (2024)
by: Cong, Bai, et al.
Published: (2024)
Optimal Eye Surgeon: Finding Image Priors through Sparse Generators at Initialization
by: Ghosh, Avrajit, et al.
Published: (2024)
by: Ghosh, Avrajit, et al.
Published: (2024)
Pruning Unrolled Networks (PUN) at Initialization for MRI Reconstruction Improves Generalization
by: Liang, Shijun, et al.
Published: (2024)
by: Liang, Shijun, et al.
Published: (2024)
Optimization Guarantees for Square-Root Natural-Gradient Variational Inference
by: Kumar, Navish, et al.
Published: (2025)
by: Kumar, Navish, et al.
Published: (2025)
Variational Learning is Effective for Large Deep Networks
by: Shen, Yuesong, et al.
Published: (2024)
by: Shen, Yuesong, et al.
Published: (2024)
Natural Variational Annealing for Multimodal Optimization
by: LeMinh, Tâm, et al.
Published: (2025)
by: LeMinh, Tâm, et al.
Published: (2025)
Information Geometry of Variational Bayes
by: Khan, Mohammad Emtiyaz
Published: (2025)
by: Khan, Mohammad Emtiyaz
Published: (2025)
Federated ADMM from Bayesian Duality
by: Möllenhoff, Thomas, et al.
Published: (2025)
by: Möllenhoff, Thomas, et al.
Published: (2025)
SVRG and Beyond via Posterior Correction
by: Daheim, Nico, et al.
Published: (2025)
by: Daheim, Nico, et al.
Published: (2025)
Conformal Prediction via Regression-as-Classification
by: Guha, Etash, et al.
Published: (2024)
by: Guha, Etash, et al.
Published: (2024)
The Memory Perturbation Equation: Understanding Model's Sensitivity to Data
by: Nickl, Peter, et al.
Published: (2023)
by: Nickl, Peter, et al.
Published: (2023)
Log-Normal Multiplicative Dynamics for Stable Low-Precision Training of Large Networks
by: Nishida, Keigo, et al.
Published: (2025)
by: Nishida, Keigo, et al.
Published: (2025)
Joint Model and Data Sparsification via the Marginal Likelihood
by: Timans, Alexander, et al.
Published: (2026)
by: Timans, Alexander, et al.
Published: (2026)
How to Weight Multitask Finetuning? Fast Previews via Bayesian Model-Merging
by: Maldonado, Hugo Monzón, et al.
Published: (2024)
by: Maldonado, Hugo Monzón, et al.
Published: (2024)
Model Merging by Uncertainty-Based Gradient Matching
by: Daheim, Nico, et al.
Published: (2023)
by: Daheim, Nico, et al.
Published: (2023)
Compact Memory for Continual Logistic Regression
by: Jung, Yohan, et al.
Published: (2025)
by: Jung, Yohan, et al.
Published: (2025)
The Bayesian Learning Rule
by: Khan, Mohammad Emtiyaz, et al.
Published: (2021)
by: Khan, Mohammad Emtiyaz, et al.
Published: (2021)
Knowledge Adaptation as Posterior Correction
by: Khan, Mohammad Emtiyaz
Published: (2025)
by: Khan, Mohammad Emtiyaz
Published: (2025)
Understanding Untrained Deep Models for Inverse Problems: Algorithms and Theory
by: Alkhouri, Ismail, et al.
Published: (2025)
by: Alkhouri, Ismail, et al.
Published: (2025)
Improving Generalization of Complex Models under Unbounded Loss Using PAC-Bayes Bounds
by: Zhang, Xitong, et al.
Published: (2023)
by: Zhang, Xitong, et al.
Published: (2023)
Variational Learning Induces Adaptive Label Smoothing
by: Yang, Sin-Han, et al.
Published: (2025)
by: Yang, Sin-Han, et al.
Published: (2025)
Adaptive Local Neighborhood-based Neural Networks for MR Image Reconstruction from Undersampled Data
by: Liang, Shijun, et al.
Published: (2022)
by: Liang, Shijun, et al.
Published: (2022)
Stein's Lemma for the Reparameterization Trick with Exponential Family Mixtures
by: Lin, Wu, et al.
Published: (2019)
by: Lin, Wu, et al.
Published: (2019)
A Dataless Reinforcement Learning Approach to Rounding Hyperplane Optimization for Max-Cut
by: Maliakal, Gabriel, et al.
Published: (2025)
by: Maliakal, Gabriel, et al.
Published: (2025)
Gradient Descent with Polyak's Momentum Finds Flatter Minima via Large Catapults
by: Phunyaphibarn, Prin, et al.
Published: (2023)
by: Phunyaphibarn, Prin, et al.
Published: (2023)
Connecting Federated ADMM to Bayes
by: Swaroop, Siddharth, et al.
Published: (2025)
by: Swaroop, Siddharth, et al.
Published: (2025)
Tada-DIP: Input-adaptive Deep Image Prior for One-shot 3D Image Reconstruction
by: Bell, Evan, et al.
Published: (2025)
by: Bell, Evan, et al.
Published: (2025)
SITCOM: Step-wise Triple-Consistent Diffusion Sampling for Inverse Problems
by: Alkhouri, Ismail, et al.
Published: (2024)
by: Alkhouri, Ismail, et al.
Published: (2024)
Balancing Speed and Stability: The Trade-offs of FP8 vs. BF16 Training in LLMs
by: Fujii, Kazuki, et al.
Published: (2024)
by: Fujii, Kazuki, et al.
Published: (2024)
Variational Schrödinger Momentum Diffusion
by: Rojas, Kevin, et al.
Published: (2025)
by: Rojas, Kevin, et al.
Published: (2025)
Simplifying Momentum-based Positive-definite Submanifold Optimization with Applications to Deep Learning
by: Lin, Wu, et al.
Published: (2023)
by: Lin, Wu, et al.
Published: (2023)
A Stein Identity for q-Gaussians with Bounded Support
by: Sklaviadis, Sophia, et al.
Published: (2026)
by: Sklaviadis, Sophia, et al.
Published: (2026)
Hard labels sampled from sparse targets mislead rotation invariant algorithms
by: Ghosh, Avrajit, et al.
Published: (2026)
by: Ghosh, Avrajit, et al.
Published: (2026)
DeepTTV: Deep Learning Prediction of Hidden Exoplanet From Transit Timing Variations
by: Chen, Chen, et al.
Published: (2024)
by: Chen, Chen, et al.
Published: (2024)
Learning Gradient-based Mixup with Extrapolation toward Flatter Minima for Domain Generalization
by: Peng, Danni, et al.
Published: (2022)
by: Peng, Danni, et al.
Published: (2022)
Bilateral Sharpness-Aware Minimization for Flatter Minima
by: Deng, Jiaxin, et al.
Published: (2024)
by: Deng, Jiaxin, et al.
Published: (2024)
UGoDIT: Unsupervised Group Deep Image Prior Via Transferable Weights
by: Liang, Shijun, et al.
Published: (2025)
by: Liang, Shijun, et al.
Published: (2025)
Sharpness-Aware Minimization Efficiently Selects Flatter Minima Late in Training
by: Zhou, Zhanpeng, et al.
Published: (2024)
by: Zhou, Zhanpeng, et al.
Published: (2024)
Similar Items
-
Learning Dynamics of Deep Linear Networks Beyond the Edge of Stability
by: Ghosh, Avrajit, et al.
Published: (2025) -
Improving LoRA with Variational Learning
by: Cong, Bai, et al.
Published: (2025) -
Variational Low-Rank Adaptation Using IVON
by: Cong, Bai, et al.
Published: (2024) -
Optimal Eye Surgeon: Finding Image Priors through Sparse Generators at Initialization
by: Ghosh, Avrajit, et al.
Published: (2024) -
Pruning Unrolled Networks (PUN) at Initialization for MRI Reconstruction Improves Generalization
by: Liang, Shijun, et al.
Published: (2024)