Bootstrap SGD: Algorithmic Stability and Robustness
Fuente:
arXiv
Salvato in:
| Autori principali: | Christmann, Andreas, Lei, Yunwen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Minibatch and Local SGD: Algorithmic Stability and Linear Speedup in Generalization
di: Lei, Yunwen, et al.
Pubblicazione: (2023)
di: Lei, Yunwen, et al.
Pubblicazione: (2023)
Stochastic Gradient Descent with Momentum is Algorithmically Stable
di: Lei, Yunwen, et al.
Pubblicazione: (2026)
di: Lei, Yunwen, et al.
Pubblicazione: (2026)
Towards Initialization-dependent and Non-vacuous Generalization Bounds for Overparameterized Shallow Neural Networks
di: Lei, Yunwen, et al.
Pubblicazione: (2026)
di: Lei, Yunwen, et al.
Pubblicazione: (2026)
Learning Theory of the SVRG: Generalization and Convergence Analysis
di: Lei, Yunwen, et al.
Pubblicazione: (2026)
di: Lei, Yunwen, et al.
Pubblicazione: (2026)
Generalization and Optimization of SGD with Lookahead
di: Li, Kangcheng, et al.
Pubblicazione: (2025)
di: Li, Kangcheng, et al.
Pubblicazione: (2025)
SGD at the Edge of Stability: The Stochastic Sharpness Gap
di: Liao, Fangshuo, et al.
Pubblicazione: (2026)
di: Liao, Fangshuo, et al.
Pubblicazione: (2026)
Target Networks and Over-parameterization Stabilize Off-policy Bootstrapping with Function Approximation
di: Che, Fengdi, et al.
Pubblicazione: (2024)
di: Che, Fengdi, et al.
Pubblicazione: (2024)
Accumulative SGD Influence Estimation for Data Attribution
di: Shi, Yunxiao, et al.
Pubblicazione: (2025)
di: Shi, Yunxiao, et al.
Pubblicazione: (2025)
Anon: Extrapolating Adaptivity Beyond SGD and Adam
di: Zhang, Yiheng, et al.
Pubblicazione: (2026)
di: Zhang, Yiheng, et al.
Pubblicazione: (2026)
The Stabilizer Bootstrap of Quantum Machine Learning with up to 10000 qubits
di: Li, Yuqing, et al.
Pubblicazione: (2024)
di: Li, Yuqing, et al.
Pubblicazione: (2024)
Bootstrapped Reward Shaping
di: Adamczyk, Jacob, et al.
Pubblicazione: (2025)
di: Adamczyk, Jacob, et al.
Pubblicazione: (2025)
Scaling Laws of SignSGD in Linear Regression: When Does It Outperform SGD?
di: Kim, Jihwan, et al.
Pubblicazione: (2026)
di: Kim, Jihwan, et al.
Pubblicazione: (2026)
Worker Disagreement Reveals Sharp Directions in Local SGD
di: Dimlioglu, Tolga, et al.
Pubblicazione: (2026)
di: Dimlioglu, Tolga, et al.
Pubblicazione: (2026)
Bootstrapping Expectiles in Reinforcement Learning
di: Clavier, Pierre, et al.
Pubblicazione: (2024)
di: Clavier, Pierre, et al.
Pubblicazione: (2024)
Imitation Bootstrapped Reinforcement Learning
di: Hu, Hengyuan, et al.
Pubblicazione: (2023)
di: Hu, Hengyuan, et al.
Pubblicazione: (2023)
RQP-SGD: Differential Private Machine Learning through Noisy SGD and Randomized Quantization
di: Feng, Ce, et al.
Pubblicazione: (2024)
di: Feng, Ce, et al.
Pubblicazione: (2024)
Diagonalisation SGD: Fast & Convergent SGD for Non-Differentiable Models via Reparameterisation and Smoothing
di: Wagner, Dominik, et al.
Pubblicazione: (2024)
di: Wagner, Dominik, et al.
Pubblicazione: (2024)
SWAN: SGD with Normalization and Whitening Enables Stateless LLM Training
di: Ma, Chao, et al.
Pubblicazione: (2024)
di: Ma, Chao, et al.
Pubblicazione: (2024)
DC-SGD: Differentially Private SGD with Dynamic Clipping through Gradient Norm Distribution Estimation
di: Wei, Chengkun, et al.
Pubblicazione: (2025)
di: Wei, Chengkun, et al.
Pubblicazione: (2025)
Mixed-Sample SGD: an End-to-end Analysis of Supervised Transfer Learning
di: Deng, Yuyang, et al.
Pubblicazione: (2025)
di: Deng, Yuyang, et al.
Pubblicazione: (2025)
INO-SGD: Addressing Utility Imbalance under Individualized Differential Privacy
di: Tian, Xiao, et al.
Pubblicazione: (2026)
di: Tian, Xiao, et al.
Pubblicazione: (2026)
On the Learning Dynamics of Two-layer Linear Networks with Label Noise SGD
di: Zhang, Tongcheng, et al.
Pubblicazione: (2026)
di: Zhang, Tongcheng, et al.
Pubblicazione: (2026)
StoSignSGD: Unbiased Structural Stochasticity Fixes SignSGD for Training Large Language Models
di: Yu, Dingzhi, et al.
Pubblicazione: (2026)
di: Yu, Dingzhi, et al.
Pubblicazione: (2026)
Enhancing DP-SGD through Non-monotonous Adaptive Scaling Gradient Weight
di: Huang, Tao, et al.
Pubblicazione: (2024)
di: Huang, Tao, et al.
Pubblicazione: (2024)
Value of Information-Enhanced Exploration in Bootstrapped DQN
di: Plataniotis, Stergios, et al.
Pubblicazione: (2025)
di: Plataniotis, Stergios, et al.
Pubblicazione: (2025)
Connections between Schedule-Free Optimizers, AdEMAMix, and Accelerated SGD Variants
di: Morwani, Depen, et al.
Pubblicazione: (2025)
di: Morwani, Depen, et al.
Pubblicazione: (2025)
Bootstrapped Model Predictive Control
di: Wang, Yuhang, et al.
Pubblicazione: (2025)
di: Wang, Yuhang, et al.
Pubblicazione: (2025)
Flattening Hierarchies with Policy Bootstrapping
di: Zhou, John L., et al.
Pubblicazione: (2025)
di: Zhou, John L., et al.
Pubblicazione: (2025)
Algorithms for Adversarially Robust Deep Learning
di: Robey, Alexander
Pubblicazione: (2025)
di: Robey, Alexander
Pubblicazione: (2025)
APOLLO: SGD-like Memory, AdamW-level Performance
di: Zhu, Hanqing, et al.
Pubblicazione: (2024)
di: Zhu, Hanqing, et al.
Pubblicazione: (2024)
Do We Need Adam? Surprisingly Strong and Sparse Reinforcement Learning with SGD in LLMs
di: Mukherjee, Sagnik, et al.
Pubblicazione: (2026)
di: Mukherjee, Sagnik, et al.
Pubblicazione: (2026)
Stochastic Collapse: How Gradient Noise Attracts SGD Dynamics Towards Simpler Subnetworks
di: Chen, Feng, et al.
Pubblicazione: (2023)
di: Chen, Feng, et al.
Pubblicazione: (2023)
Why Adam Can Beat SGD: Second-Moment Normalization Yields Sharper Tails
di: Jin, Ruinan, et al.
Pubblicazione: (2026)
di: Jin, Ruinan, et al.
Pubblicazione: (2026)
Improving the Diversity of Bootstrapped DQN by Replacing Priors With Noise
di: Meng, Li, et al.
Pubblicazione: (2022)
di: Meng, Li, et al.
Pubblicazione: (2022)
Bootstrapping LLMs via Preference-Based Policy Optimization
di: Jia, Chen
Pubblicazione: (2025)
di: Jia, Chen
Pubblicazione: (2025)
Bootstrap Off-policy with World Model
di: Zhan, Guojian, et al.
Pubblicazione: (2025)
di: Zhan, Guojian, et al.
Pubblicazione: (2025)
Representations learnt by SGD and Adaptive learning rules: Conditions that vary sparsity and selectivity in neural networks
di: Park, Jin Hyun
Pubblicazione: (2022)
di: Park, Jin Hyun
Pubblicazione: (2022)
Decentralised Resource Sharing in TinyML: Wireless Bilayer Gossip Parallel SGD for Collaborative Learning
di: Bao, Ziyuan, et al.
Pubblicazione: (2025)
di: Bao, Ziyuan, et al.
Pubblicazione: (2025)
LANPO: Bootstrapping Language and Numerical Feedback for Reinforcement Learning in LLMs
di: Li, Ang, et al.
Pubblicazione: (2025)
di: Li, Ang, et al.
Pubblicazione: (2025)
Whispering to a Blackbox: Bootstrapping Frozen OCR with Visual Prompts
di: Samandarov, Samandar, et al.
Pubblicazione: (2026)
di: Samandarov, Samandar, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Minibatch and Local SGD: Algorithmic Stability and Linear Speedup in Generalization
di: Lei, Yunwen, et al.
Pubblicazione: (2023) -
Stochastic Gradient Descent with Momentum is Algorithmically Stable
di: Lei, Yunwen, et al.
Pubblicazione: (2026) -
Towards Initialization-dependent and Non-vacuous Generalization Bounds for Overparameterized Shallow Neural Networks
di: Lei, Yunwen, et al.
Pubblicazione: (2026) -
Learning Theory of the SVRG: Generalization and Convergence Analysis
di: Lei, Yunwen, et al.
Pubblicazione: (2026) -
Generalization and Optimization of SGD with Lookahead
di: Li, Kangcheng, et al.
Pubblicazione: (2025)