Convergence of projected stochastic natural gradient variational inference for various step size and sample or batch size schedules
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Guilmeau, Thomas, Hendrikx, Hadrien, Forbes, Florence |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Convergence rates of stochastic gradient method with independent sequences of step-size and momentum weight
von: Hwang, Wen-Liang
Veröffentlicht: (2024)
von: Hwang, Wen-Liang
Veröffentlicht: (2024)
New logarithmic step size for stochastic gradient descent
von: Shamaee, M. Soheil, et al.
Veröffentlicht: (2024)
von: Shamaee, M. Soheil, et al.
Veröffentlicht: (2024)
From Inexact Gradients to Byzantine Robustness: Acceleration and Optimization under Similarity
von: Gaucher, Renaud, et al.
Veröffentlicht: (2026)
von: Gaucher, Renaud, et al.
Veröffentlicht: (2026)
Unified Breakdown Analysis for Byzantine Robust Gossip
von: Gaucher, Renaud, et al.
Veröffentlicht: (2024)
von: Gaucher, Renaud, et al.
Veröffentlicht: (2024)
The Relative Gaussian Mechanism and its Application to Private Gradient Descent
von: Hendrikx, Hadrien, et al.
Veröffentlicht: (2023)
von: Hendrikx, Hadrien, et al.
Veröffentlicht: (2023)
Investigating Variance Definitions for Mirror Descent with Relative Smoothness
von: Hendrikx, Hadrien
Veröffentlicht: (2024)
von: Hendrikx, Hadrien
Veröffentlicht: (2024)
Simultaneous estimation of multiple discrete unimodal distributions under stochastic order constraints
von: Yoshida, Yasuhiro, et al.
Veröffentlicht: (2026)
von: Yoshida, Yasuhiro, et al.
Veröffentlicht: (2026)
SplitVAEs: Decentralized scenario generation from siloed data for stochastic optimization problems
von: Islam, H M Mohaimanul, et al.
Veröffentlicht: (2024)
von: Islam, H M Mohaimanul, et al.
Veröffentlicht: (2024)
Convergence and concentration properties of constant step-size SGD through Markov chains
von: Merad, Ibrahim, et al.
Veröffentlicht: (2023)
von: Merad, Ibrahim, et al.
Veröffentlicht: (2023)
Is your batch size the problem? Revisiting the Adam-SGD gap in language modeling
von: Srećković, Teodora, et al.
Veröffentlicht: (2025)
von: Srećković, Teodora, et al.
Veröffentlicht: (2025)
Bayesian dynamic scheduling of multipurpose batch processes under incomplete look-ahead information
von: Zheng, Taicheng, et al.
Veröffentlicht: (2025)
von: Zheng, Taicheng, et al.
Veröffentlicht: (2025)
Convergence of continuous-time stochastic gradient descent with applications to deep neural networks
von: Lugosi, Gabor, et al.
Veröffentlicht: (2024)
von: Lugosi, Gabor, et al.
Veröffentlicht: (2024)
Robust stochastic first order methods in heavy-tailed noise via medoid mini-batch gradient sampling
von: Vukovic, Manojlo, et al.
Veröffentlicht: (2026)
von: Vukovic, Manojlo, et al.
Veröffentlicht: (2026)
Shuffling the Data, Stretching the Step-size: Sharper Bias in constant step-size SGD
von: Emmanouilidis, Konstantinos, et al.
Veröffentlicht: (2026)
von: Emmanouilidis, Konstantinos, et al.
Veröffentlicht: (2026)
Learning Sequential Decisions from Multiple Sources via Group-Robust Markov Decision Processes
von: Xu, Mingyuan, et al.
Veröffentlicht: (2026)
von: Xu, Mingyuan, et al.
Veröffentlicht: (2026)
Locally Adaptive Multi-Objective Learning
von: Kaur, Jivat Neet, et al.
Veröffentlicht: (2026)
von: Kaur, Jivat Neet, et al.
Veröffentlicht: (2026)
Automatic Outlier Rectification via Optimal Transport
von: Blanchet, Jose, et al.
Veröffentlicht: (2024)
von: Blanchet, Jose, et al.
Veröffentlicht: (2024)
Learning Large Causal Structures from Inverse Covariance Matrix via Sparse Matrix Decomposition
von: Dong, Shuyu, et al.
Veröffentlicht: (2022)
von: Dong, Shuyu, et al.
Veröffentlicht: (2022)
Robust Data-driven Prescriptiveness Optimization
von: Poursoltani, Mehran, et al.
Veröffentlicht: (2023)
von: Poursoltani, Mehran, et al.
Veröffentlicht: (2023)
Fast Robust Kernel Regression through Sign Gradient Descent with Early Stopping
von: Allerbo, Oskar
Veröffentlicht: (2023)
von: Allerbo, Oskar
Veröffentlicht: (2023)
AltGDmin: Alternating GD and Minimization for Partly-Decoupled (Federated) Optimization
von: Vaswani, Namrata
Veröffentlicht: (2025)
von: Vaswani, Namrata
Veröffentlicht: (2025)
A distribution-free mixed-integer optimization approach to hierarchical modelling of clustered and longitudinal data
von: Sankaranarayanan, Madhav, et al.
Veröffentlicht: (2023)
von: Sankaranarayanan, Madhav, et al.
Veröffentlicht: (2023)
Random Pareto front surfaces
von: Tu, Ben, et al.
Veröffentlicht: (2024)
von: Tu, Ben, et al.
Veröffentlicht: (2024)
Optimal Cross-Validation for Sparse Linear Regression
von: Cory-Wright, Ryan, et al.
Veröffentlicht: (2023)
von: Cory-Wright, Ryan, et al.
Veröffentlicht: (2023)
Global Group Fairness in Federated Learning via Function Tracking
von: Rychener, Yves, et al.
Veröffentlicht: (2025)
von: Rychener, Yves, et al.
Veröffentlicht: (2025)
Changing the Kernel During Training Leads to Double Descent in Kernel Regression
von: Allerbo, Oskar
Veröffentlicht: (2023)
von: Allerbo, Oskar
Veröffentlicht: (2023)
Optimizing Input Data Collection for Ranking and Selection
von: Song, Eunhye, et al.
Veröffentlicht: (2025)
von: Song, Eunhye, et al.
Veröffentlicht: (2025)
DCILP: A Distributed Approach for Large-Scale Causal Structure Learning
von: Dong, Shuyu, et al.
Veröffentlicht: (2024)
von: Dong, Shuyu, et al.
Veröffentlicht: (2024)
Calibrated Multi-Level Quantile Forecasting
von: Ding, Tiffany, et al.
Veröffentlicht: (2025)
von: Ding, Tiffany, et al.
Veröffentlicht: (2025)
From Contextual Data to Newsvendor Decisions: On the Actual Performance of Data-Driven Algorithms
von: Besbes, Omar, et al.
Veröffentlicht: (2023)
von: Besbes, Omar, et al.
Veröffentlicht: (2023)
Data-Driven Sequential Sampling for Tail Risk Mitigation
von: Ahn, Dohyun, et al.
Veröffentlicht: (2025)
von: Ahn, Dohyun, et al.
Veröffentlicht: (2025)
Arc travel time and path choice model estimation subsumed
von: Mohammadpour, Sobhan, et al.
Veröffentlicht: (2022)
von: Mohammadpour, Sobhan, et al.
Veröffentlicht: (2022)
Data-Driven Influence Functions for Optimization-Based Causal Inference
von: Jordan, Michael I., et al.
Veröffentlicht: (2022)
von: Jordan, Michael I., et al.
Veröffentlicht: (2022)
Bridging Constraints and Stochasticity: A Fully First-Order Method for Stochastic Bilevel Optimization with Linear Constraints
von: Phan, Cac, et al.
Veröffentlicht: (2025)
von: Phan, Cac, et al.
Veröffentlicht: (2025)
Bayesian optimization for mixed variables using an adaptive dimension reduction process: applications to aircraft design
von: Saves, Paul, et al.
Veröffentlicht: (2025)
von: Saves, Paul, et al.
Veröffentlicht: (2025)
Structured Difference-of-Q via Orthogonal Learning
von: Cao, Defu, et al.
Veröffentlicht: (2024)
von: Cao, Defu, et al.
Veröffentlicht: (2024)
The Stochastic Proximal Distance Algorithm
von: Jiang, Haoyu, et al.
Veröffentlicht: (2022)
von: Jiang, Haoyu, et al.
Veröffentlicht: (2022)
"Over-optimizing" for Normality: Budget-constrained Uncertainty Quantification for Contextual Decision-making
von: Wang, Yanyuan, et al.
Veröffentlicht: (2025)
von: Wang, Yanyuan, et al.
Veröffentlicht: (2025)
Robust Learning Rate Selection for Stochastic Optimization via Splitting Diagnostic
von: Sordello, Matteo, et al.
Veröffentlicht: (2019)
von: Sordello, Matteo, et al.
Veröffentlicht: (2019)
Gradient descent inference in empirical risk minimization
von: Han, Qiyang, et al.
Veröffentlicht: (2024)
von: Han, Qiyang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Convergence rates of stochastic gradient method with independent sequences of step-size and momentum weight
von: Hwang, Wen-Liang
Veröffentlicht: (2024) -
New logarithmic step size for stochastic gradient descent
von: Shamaee, M. Soheil, et al.
Veröffentlicht: (2024) -
From Inexact Gradients to Byzantine Robustness: Acceleration and Optimization under Similarity
von: Gaucher, Renaud, et al.
Veröffentlicht: (2026) -
Unified Breakdown Analysis for Byzantine Robust Gossip
von: Gaucher, Renaud, et al.
Veröffentlicht: (2024) -
The Relative Gaussian Mechanism and its Application to Private Gradient Descent
von: Hendrikx, Hadrien, et al.
Veröffentlicht: (2023)