Optimization Guarantees for Square-Root Natural-Gradient Variational Inference
Fuente:
arXiv
Saved in:
| Main Authors: | Kumar, Navish, Möllenhoff, Thomas, Khan, Mohammad Emtiyaz, Lucchi, Aurelien |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Information Geometry of Variational Bayes
by: Khan, Mohammad Emtiyaz
Published: (2025)
by: Khan, Mohammad Emtiyaz
Published: (2025)
Model Merging by Uncertainty-Based Gradient Matching
by: Daheim, Nico, et al.
Published: (2023)
by: Daheim, Nico, et al.
Published: (2023)
SVRG and Beyond via Posterior Correction
by: Daheim, Nico, et al.
Published: (2025)
by: Daheim, Nico, et al.
Published: (2025)
Improving LoRA with Variational Learning
by: Cong, Bai, et al.
Published: (2025)
by: Cong, Bai, et al.
Published: (2025)
The Memory Perturbation Equation: Understanding Model's Sensitivity to Data
by: Nickl, Peter, et al.
Published: (2023)
by: Nickl, Peter, et al.
Published: (2023)
Variational Low-Rank Adaptation Using IVON
by: Cong, Bai, et al.
Published: (2024)
by: Cong, Bai, et al.
Published: (2024)
How to Weight Multitask Finetuning? Fast Previews via Bayesian Model-Merging
by: Maldonado, Hugo Monzón, et al.
Published: (2024)
by: Maldonado, Hugo Monzón, et al.
Published: (2024)
Natural Variational Annealing for Multimodal Optimization
by: LeMinh, Tâm, et al.
Published: (2025)
by: LeMinh, Tâm, et al.
Published: (2025)
Knowledge Adaptation as Posterior Correction
by: Khan, Mohammad Emtiyaz
Published: (2025)
by: Khan, Mohammad Emtiyaz
Published: (2025)
Variational Learning Induces Adaptive Label Smoothing
by: Yang, Sin-Han, et al.
Published: (2025)
by: Yang, Sin-Han, et al.
Published: (2025)
Variational Learning is Effective for Large Deep Networks
by: Shen, Yuesong, et al.
Published: (2024)
by: Shen, Yuesong, et al.
Published: (2024)
Connecting Federated ADMM to Bayes
by: Swaroop, Siddharth, et al.
Published: (2025)
by: Swaroop, Siddharth, et al.
Published: (2025)
Federated ADMM from Bayesian Duality
by: Möllenhoff, Thomas, et al.
Published: (2025)
by: Möllenhoff, Thomas, et al.
Published: (2025)
Conformal Prediction via Regression-as-Classification
by: Guha, Etash, et al.
Published: (2024)
by: Guha, Etash, et al.
Published: (2024)
Gradient Extrapolation-Based Policy Optimization
by: Swapnil, Ismam Nur, et al.
Published: (2026)
by: Swapnil, Ismam Nur, et al.
Published: (2026)
When Bias Meets Trainability: Connecting Theories of Initialization
by: Bassi, Alberto, et al.
Published: (2025)
by: Bassi, Alberto, et al.
Published: (2025)
Variational Learning Finds Flatter Solutions at the Edge of Stability
by: Ghosh, Avrajit, et al.
Published: (2025)
by: Ghosh, Avrajit, et al.
Published: (2025)
Log-Normal Multiplicative Dynamics for Stable Low-Precision Training of Large Networks
by: Nishida, Keigo, et al.
Published: (2025)
by: Nishida, Keigo, et al.
Published: (2025)
Joint Model and Data Sparsification via the Marginal Likelihood
by: Timans, Alexander, et al.
Published: (2026)
by: Timans, Alexander, et al.
Published: (2026)
Local Reinforcement Learning with Action-Conditioned Root Mean Squared Q-Functions
by: Wu, Frank, et al.
Published: (2025)
by: Wu, Frank, et al.
Published: (2025)
Uncertainty-Aware Decoding with Minimum Bayes Risk
by: Daheim, Nico, et al.
Published: (2025)
by: Daheim, Nico, et al.
Published: (2025)
Gradient Regularized Natural Gradients
by: Dash, Satya Prakash, et al.
Published: (2026)
by: Dash, Satya Prakash, et al.
Published: (2026)
Gradient Flow Convergence Guarantee for General Neural Network Architectures
by: Jakhmola, Yash
Published: (2025)
by: Jakhmola, Yash
Published: (2025)
Score-based Integrated Gradient for Root Cause Explanations of Outliers
by: Nguyen, Phuoc, et al.
Published: (2026)
by: Nguyen, Phuoc, et al.
Published: (2026)
Reconstructing Deep Neural Networks: Unleashing the Optimization Potential of Natural Gradient Descent
by: Liu, Weihua, et al.
Published: (2024)
by: Liu, Weihua, et al.
Published: (2024)
Iterative Amortized Inference: Unifying In-Context Learning and Learned Optimizers
by: Mittal, Sarthak, et al.
Published: (2025)
by: Mittal, Sarthak, et al.
Published: (2025)
UnIT: Scalable Unstructured Inference-Time Pruning for MAC-efficient Neural Inference on MCUs
by: Neth, Ashe, et al.
Published: (2025)
by: Neth, Ashe, et al.
Published: (2025)
Fast and Robust Likelihood-Guided Diffusion Posterior Sampling with Amortized Variational Inference
by: Zheng, Léon, et al.
Published: (2026)
by: Zheng, Léon, et al.
Published: (2026)
Compact Memory for Continual Logistic Regression
by: Jung, Yohan, et al.
Published: (2025)
by: Jung, Yohan, et al.
Published: (2025)
ONG: Orthogonal Natural Gradient Descent
by: Yadav, Yajat, et al.
Published: (2025)
by: Yadav, Yajat, et al.
Published: (2025)
Remove that Square Root: A New Efficient Scale-Invariant Version of AdaGrad
by: Choudhury, Sayantan, et al.
Published: (2024)
by: Choudhury, Sayantan, et al.
Published: (2024)
SMMF: Square-Matricized Momentum Factorization for Memory-Efficient Optimization
by: Park, Kwangryeol, et al.
Published: (2024)
by: Park, Kwangryeol, et al.
Published: (2024)
Variational Inference via Smoothed Particle Hydrodynamics
by: Huang, Yongchao
Published: (2024)
by: Huang, Yongchao
Published: (2024)
Instance-Adaptive Parametrization for Amortized Variational Inference
by: Pollastro, Andrea, et al.
Published: (2026)
by: Pollastro, Andrea, et al.
Published: (2026)
Variational Autoencoders for Efficient Simulation-Based Inference
by: Nautiyal, Mayank, et al.
Published: (2024)
by: Nautiyal, Mayank, et al.
Published: (2024)
Stochastic Re-weighted Gradient Descent via Distributionally Robust Optimization
by: Kumar, Ramnath, et al.
Published: (2023)
by: Kumar, Ramnath, et al.
Published: (2023)
Global Convergence Guarantees for Federated Policy Gradient Methods with Adversaries
by: Ganesh, Swetha, et al.
Published: (2024)
by: Ganesh, Swetha, et al.
Published: (2024)
KernelSHAP-IQ: Weighted Least-Square Optimization for Shapley Interactions
by: Fumagalli, Fabian, et al.
Published: (2024)
by: Fumagalli, Fabian, et al.
Published: (2024)
MAVRL: Learning Reward Functions from Multiple Feedback Types with Amortized Variational Inference
by: Baur, Raphaël, et al.
Published: (2026)
by: Baur, Raphaël, et al.
Published: (2026)
How to Square Tensor Networks and Circuits Without Squaring Them
by: Loconte, Lorenzo, et al.
Published: (2025)
by: Loconte, Lorenzo, et al.
Published: (2025)
Similar Items
-
Information Geometry of Variational Bayes
by: Khan, Mohammad Emtiyaz
Published: (2025) -
Model Merging by Uncertainty-Based Gradient Matching
by: Daheim, Nico, et al.
Published: (2023) -
SVRG and Beyond via Posterior Correction
by: Daheim, Nico, et al.
Published: (2025) -
Improving LoRA with Variational Learning
by: Cong, Bai, et al.
Published: (2025) -
The Memory Perturbation Equation: Understanding Model's Sensitivity to Data
by: Nickl, Peter, et al.
Published: (2023)