Saved in:
| Main Authors: | Oldewage, Elre T., Clarke, Ross M., Hernández-Lobato, José Miguel |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2310.14901 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Studying K-FAC Heuristics by Viewing Adam through a Second-Order Lens
by: Clarke, Ross M., et al.
Published: (2023)
by: Clarke, Ross M., et al.
Published: (2023)
Warm Start Marginal Likelihood Optimisation for Iterative Gaussian Processes
by: Lin, Jihao Andreas, et al.
Published: (2024)
by: Lin, Jihao Andreas, et al.
Published: (2024)
Leveraging Task Structures for Improved Identifiability in Neural Network Representations
by: Chen, Wenlin, et al.
Published: (2023)
by: Chen, Wenlin, et al.
Published: (2023)
Improving Linear System Solvers for Hyperparameter Optimisation in Iterative Gaussian Processes
by: Lin, Jihao Andreas, et al.
Published: (2024)
by: Lin, Jihao Andreas, et al.
Published: (2024)
Training-Free Vector Quantization via Gaussian VAEs
by: Xu, Tongda, et al.
Published: (2025)
by: Xu, Tongda, et al.
Published: (2025)
Getting Free Bits Back from Rotational Symmetries in LLMs
by: He, Jiajun, et al.
Published: (2024)
by: He, Jiajun, et al.
Published: (2024)
Hessian-guided Perturbed Wasserstein Gradient Flows for Escaping Saddle Points
by: Yamamoto, Naoya, et al.
Published: (2025)
by: Yamamoto, Naoya, et al.
Published: (2025)
Inertial Newton Algorithms Avoiding Strict Saddle Points
by: Castera, Camille
Published: (2021)
by: Castera, Camille
Published: (2021)
Saddle-to-Saddle Dynamics Explains A Simplicity Bias Across Neural Network Architectures
by: Zhang, Yedi, et al.
Published: (2025)
by: Zhang, Yedi, et al.
Published: (2025)
RECOMBINER: Robust and Enhanced Compression with Bayesian Implicit Neural Representations
by: He, Jiajun, et al.
Published: (2023)
by: He, Jiajun, et al.
Published: (2023)
Uncertainty Modeling in Graph Neural Networks via Stochastic Differential Equations
by: Bergna, Richard, et al.
Published: (2024)
by: Bergna, Richard, et al.
Published: (2024)
Diagnosing and fixing common problems in Bayesian optimization for molecule design
by: Tripp, Austin, et al.
Published: (2024)
by: Tripp, Austin, et al.
Published: (2024)
Neural Network-based High-index Saddle Dynamics Method for Searching Saddle Points and Solution Landscape
by: Liu, Yuankai, et al.
Published: (2024)
by: Liu, Yuankai, et al.
Published: (2024)
Training Neural Samplers with Reverse Diffusive KL Divergence
by: He, Jiajun, et al.
Published: (2024)
by: He, Jiajun, et al.
Published: (2024)
Causal Effect Estimation under Networked Interference without Networked Unconfoundedness Assumption
by: Chen, Weilin, et al.
Published: (2025)
by: Chen, Weilin, et al.
Published: (2025)
Better Training Data Attribution via Better Inverse Hessian-Vector Products
by: Wang, Andrew, et al.
Published: (2025)
by: Wang, Andrew, et al.
Published: (2025)
A Gaussian Process View on Observation Noise and Initialization in Wide Neural Networks
by: Calvo-Ordoñez, Sergio, et al.
Published: (2025)
by: Calvo-Ordoñez, Sergio, et al.
Published: (2025)
Decoupled PFNs: Identifiable Epistemic-Aleatoric Decomposition via Structured Synthetic Priors
by: Bergna, Richard, et al.
Published: (2026)
by: Bergna, Richard, et al.
Published: (2026)
There Was Never a Bottleneck in Concept Bottleneck Models
by: Almudévar, Antonio, et al.
Published: (2025)
by: Almudévar, Antonio, et al.
Published: (2025)
Stochastic Gradient Descent in the Saddle-to-Saddle Regime of Deep Linear Networks
by: Corlouer, Guillaume, et al.
Published: (2026)
by: Corlouer, Guillaume, et al.
Published: (2026)
Post-Hoc Uncertainty Quantification in Pre-Trained Neural Networks via Activation-Level Gaussian Processes
by: Bergna, Richard, et al.
Published: (2025)
by: Bergna, Richard, et al.
Published: (2025)
Dimension-Free Saddle-Point Escape in Muon
by: Long, Yanlin, et al.
Published: (2026)
by: Long, Yanlin, et al.
Published: (2026)
Saddle-To-Saddle Dynamics in Deep ReLU Networks: Low-Rank Bias in the First Saddle Escape
by: Bantzis, Ioannis, et al.
Published: (2025)
by: Bantzis, Ioannis, et al.
Published: (2025)
Online Laplace Model Selection Revisited
by: Lin, Jihao Andreas, et al.
Published: (2023)
by: Lin, Jihao Andreas, et al.
Published: (2023)
Conditional Diffusion Sampling
by: Castro-Macías, Francisco M., et al.
Published: (2026)
by: Castro-Macías, Francisco M., et al.
Published: (2026)
Efficient and Unbiased Sampling from Boltzmann Distributions via Variance-Tuned Diffusion Models
by: Zhang, Fengzhe, et al.
Published: (2025)
by: Zhang, Fengzhe, et al.
Published: (2025)
Online Newton Method for Bandit Convex Optimisation
by: Fokkema, Hidde, et al.
Published: (2024)
by: Fokkema, Hidde, et al.
Published: (2024)
Accelerating Relative Entropy Coding with Space Partitioning
by: He, Jiajun, et al.
Published: (2024)
by: He, Jiajun, et al.
Published: (2024)
Wiener Chaos Expansion based Neural Operator for Singular Stochastic Partial Differential Equations
by: Shi, Dai, et al.
Published: (2026)
by: Shi, Dai, et al.
Published: (2026)
Towards Quantifying the Hessian Structure of Neural Networks
by: Dong, Zhaorui, et al.
Published: (2025)
by: Dong, Zhaorui, et al.
Published: (2025)
Newton-CG methods for nonconvex unconstrained optimization with Hölder continuous Hessian
by: He, Chuan, et al.
Published: (2023)
by: He, Chuan, et al.
Published: (2023)
Exact, Tractable Gauss-Newton Optimization in Deep Reversible Architectures Reveal Poor Generalization
by: Buffelli, Davide, et al.
Published: (2024)
by: Buffelli, Davide, et al.
Published: (2024)
Expanding the Chaos: Neural Operator for Stochastic (Partial) Differential Equations
by: Shi, Dai, et al.
Published: (2026)
by: Shi, Dai, et al.
Published: (2026)
Activation-Space Uncertainty Quantification for Pretrained Networks
by: Bergna, Richard, et al.
Published: (2026)
by: Bergna, Richard, et al.
Published: (2026)
Improving Iterative Gaussian Processes via Warm Starting Sequential Posteriors
by: Dong, Alan Yufei, et al.
Published: (2025)
by: Dong, Alan Yufei, et al.
Published: (2025)
Mitigating Forgetting in Low Rank Adaptation
by: Sliwa, Joanna, et al.
Published: (2025)
by: Sliwa, Joanna, et al.
Published: (2025)
A Diffusive Classification Loss for Learning Energy-based Generative Models
by: OuYang, RuiKang, et al.
Published: (2026)
by: OuYang, RuiKang, et al.
Published: (2026)
RNE: plug-and-play diffusion inference-time control and energy-based training
by: He, Jiajun, et al.
Published: (2025)
by: He, Jiajun, et al.
Published: (2025)
Characterizing Learning in Deep Neural Networks using Tractable Algorithmic Complexity Analysis
by: Bakhtiarifard, Pedram, et al.
Published: (2026)
by: Bakhtiarifard, Pedram, et al.
Published: (2026)
From Saddle Points Toward Global Minima: A Newton-Type Method on Wasserstein Space
by: Lascu, Razvan-Andrei, et al.
Published: (2026)
by: Lascu, Razvan-Andrei, et al.
Published: (2026)
Similar Items
-
Studying K-FAC Heuristics by Viewing Adam through a Second-Order Lens
by: Clarke, Ross M., et al.
Published: (2023) -
Warm Start Marginal Likelihood Optimisation for Iterative Gaussian Processes
by: Lin, Jihao Andreas, et al.
Published: (2024) -
Leveraging Task Structures for Improved Identifiability in Neural Network Representations
by: Chen, Wenlin, et al.
Published: (2023) -
Improving Linear System Solvers for Hyperparameter Optimisation in Iterative Gaussian Processes
by: Lin, Jihao Andreas, et al.
Published: (2024) -
Training-Free Vector Quantization via Gaussian VAEs
by: Xu, Tongda, et al.
Published: (2025)