Flat Minima and Generalization: Insights from Stochastic Convex Optimization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Schliserman, Matan, Vansover-Hager, Shira, Koren, Tomer |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Rapid Overfitting of Multi-Pass Stochastic Gradient Descent in Stochastic Convex Optimization
von: Vansover-Hager, Shira, et al.
Veröffentlicht: (2025)
von: Vansover-Hager, Shira, et al.
Veröffentlicht: (2025)
Complexity of Vector-valued Prediction: From Linear Models to Stochastic Convex Optimization
von: Schliserman, Matan, et al.
Veröffentlicht: (2024)
von: Schliserman, Matan, et al.
Veröffentlicht: (2024)
The Dimension Strikes Back with Gradients: Generalization of Gradient Methods in Stochastic Convex Optimization
von: Schliserman, Matan, et al.
Veröffentlicht: (2024)
von: Schliserman, Matan, et al.
Veröffentlicht: (2024)
Multiclass Loss Geometry Matters for Generalization of Gradient Descent in Separable Classification
von: Schliserman, Matan, et al.
Veröffentlicht: (2025)
von: Schliserman, Matan, et al.
Veröffentlicht: (2025)
Fast Last-Iterate Convergence of SGD in the Smooth Interpolation Regime
von: Attia, Amit, et al.
Veröffentlicht: (2025)
von: Attia, Amit, et al.
Veröffentlicht: (2025)
Optimal Rates in Continual Linear Regression via Increasing Regularization
von: Levinstein, Ran, et al.
Veröffentlicht: (2025)
von: Levinstein, Ran, et al.
Veröffentlicht: (2025)
From Continual Learning to SGD and Back: Better Rates for Continual Linear Models
von: Evron, Itay, et al.
Veröffentlicht: (2025)
von: Evron, Itay, et al.
Veröffentlicht: (2025)
How Free is Parameter-Free Stochastic Optimization?
von: Attia, Amit, et al.
Veröffentlicht: (2024)
von: Attia, Amit, et al.
Veröffentlicht: (2024)
Learning Rate Annealing Improves Tuning Robustness in Stochastic Optimization
von: Attia, Amit, et al.
Veröffentlicht: (2025)
von: Attia, Amit, et al.
Veröffentlicht: (2025)
Faster Stochastic Optimization with Arbitrary Delays via Asynchronous Mini-Batching
von: Attia, Amit, et al.
Veröffentlicht: (2024)
von: Attia, Amit, et al.
Veröffentlicht: (2024)
Are Flat Minima an Illusion?
von: Bennett, Michael Timothy
Veröffentlicht: (2026)
von: Bennett, Michael Timothy
Veröffentlicht: (2026)
A General Reduction for High-Probability Analysis with General Light-Tailed Distributions
von: Attia, Amit, et al.
Veröffentlicht: (2024)
von: Attia, Amit, et al.
Veröffentlicht: (2024)
Zeroth-Order Optimization Finds Flat Minima
von: Zhang, Liang, et al.
Veröffentlicht: (2025)
von: Zhang, Liang, et al.
Veröffentlicht: (2025)
Multiplicative Reweighting for Robust Neural Network Optimization
von: Bar, Noga, et al.
Veröffentlicht: (2021)
von: Bar, Noga, et al.
Veröffentlicht: (2021)
From Contextual Combinatorial Semi-Bandits to Bandit List Classification: Improved Sample Complexity with Sparse Rewards
von: Erez, Liad, et al.
Veröffentlicht: (2025)
von: Erez, Liad, et al.
Veröffentlicht: (2025)
Towards Robust Influence Functions with Flat Validation Minima
von: Ye, Xichen, et al.
Veröffentlicht: (2025)
von: Ye, Xichen, et al.
Veröffentlicht: (2025)
Rate-Optimal Policy Optimization for Linear Markov Decision Processes
von: Sherman, Uri, et al.
Veröffentlicht: (2023)
von: Sherman, Uri, et al.
Veröffentlicht: (2023)
Towards Fully Parameter-Free Stochastic Optimization: Grid Search with Self-Bounding Analysis
von: Zhao, Yuheng, et al.
Veröffentlicht: (2026)
von: Zhao, Yuheng, et al.
Veröffentlicht: (2026)
A PAC-Bayesian Link Between Generalisation and Flat Minima
von: Haddouche, Maxime, et al.
Veröffentlicht: (2024)
von: Haddouche, Maxime, et al.
Veröffentlicht: (2024)
SAFE: Finding Sparse and Flat Minima to Improve Pruning
von: Lee, Dongyeop, et al.
Veröffentlicht: (2025)
von: Lee, Dongyeop, et al.
Veröffentlicht: (2025)
Towards the Connection between Activation Sparsity and Flat Minima
von: Peng, Ze, et al.
Veröffentlicht: (2026)
von: Peng, Ze, et al.
Veröffentlicht: (2026)
Optimal Learning from Label Proportions with General Loss Functions
von: Applebaum, Lorne, et al.
Veröffentlicht: (2025)
von: Applebaum, Lorne, et al.
Veröffentlicht: (2025)
A Flat Minima Perspective on Understanding Augmentations and Model Robustness
von: Yoo, Weebum, et al.
Veröffentlicht: (2025)
von: Yoo, Weebum, et al.
Veröffentlicht: (2025)
Convergence of Policy Mirror Descent Beyond Compatible Function Approximation
von: Sherman, Uri, et al.
Veröffentlicht: (2025)
von: Sherman, Uri, et al.
Veröffentlicht: (2025)
Convergence and Sample Complexity of First-Order Methods for Agnostic Reinforcement Learning
von: Sherman, Uri, et al.
Veröffentlicht: (2025)
von: Sherman, Uri, et al.
Veröffentlicht: (2025)
Noise Stability Optimization for Finding Flat Minima: A Hessian-based Regularization Approach
von: Zhang, Hongyang R., et al.
Veröffentlicht: (2023)
von: Zhang, Hongyang R., et al.
Veröffentlicht: (2023)
A Function-Centric Perspective on Flat and Sharp Minima
von: Mason-Williams, Israel, et al.
Veröffentlicht: (2025)
von: Mason-Williams, Israel, et al.
Veröffentlicht: (2025)
Information Complexity of Stochastic Convex Optimization: Applications to Generalization and Memorization
von: Attias, Idan, et al.
Veröffentlicht: (2024)
von: Attias, Idan, et al.
Veröffentlicht: (2024)
The Hidden Cost of Approximation in Online Mirror Descent
von: Schlisselberg, Ofir, et al.
Veröffentlicht: (2025)
von: Schlisselberg, Ofir, et al.
Veröffentlicht: (2025)
On Traceability in $\ell_p$ Stochastic Convex Optimization
von: Voitovych, Sasha, et al.
Veröffentlicht: (2025)
von: Voitovych, Sasha, et al.
Veröffentlicht: (2025)
Enhancing Parallelism in Decentralized Stochastic Convex Optimization
von: Eisen, Ofri, et al.
Veröffentlicht: (2025)
von: Eisen, Ofri, et al.
Veröffentlicht: (2025)
When Flat Minima Fail: Characterizing INT4 Quantization Collapse After FP32 Convergence
von: Armstrong, Marcus
Veröffentlicht: (2026)
von: Armstrong, Marcus
Veröffentlicht: (2026)
Stochastic Difference-of-Convex Optimization with Momentum
von: Chayti, El Mahdi, et al.
Veröffentlicht: (2025)
von: Chayti, El Mahdi, et al.
Veröffentlicht: (2025)
The Price of Adaptivity in Stochastic Convex Optimization
von: Carmon, Yair, et al.
Veröffentlicht: (2024)
von: Carmon, Yair, et al.
Veröffentlicht: (2024)
Coherence Awareness in Diffractive Neural Networks
von: Kleiner, Matan, et al.
Veröffentlicht: (2024)
von: Kleiner, Matan, et al.
Veröffentlicht: (2024)
Illumination Angular Spectrum Encoding for Controlling the Functionality of Diffractive Networks
von: Kleiner, Matan, et al.
Veröffentlicht: (2026)
von: Kleiner, Matan, et al.
Veröffentlicht: (2026)
Mirror Gradient: Towards Robust Multimodal Recommender Systems via Exploring Flat Local Minima
von: Zhong, Shanshan, et al.
Veröffentlicht: (2024)
von: Zhong, Shanshan, et al.
Veröffentlicht: (2024)
Cost-Aware Learning
von: Mohri, Clara, et al.
Veröffentlicht: (2026)
von: Mohri, Clara, et al.
Veröffentlicht: (2026)
Optimal Rates for Robust Stochastic Convex Optimization
von: Gao, Changyu, et al.
Veröffentlicht: (2024)
von: Gao, Changyu, et al.
Veröffentlicht: (2024)
ContactNet: Geometric-Based Deep Learning Model for Predicting Protein-Protein Interactions
von: Halfon, Matan, et al.
Veröffentlicht: (2024)
von: Halfon, Matan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Rapid Overfitting of Multi-Pass Stochastic Gradient Descent in Stochastic Convex Optimization
von: Vansover-Hager, Shira, et al.
Veröffentlicht: (2025) -
Complexity of Vector-valued Prediction: From Linear Models to Stochastic Convex Optimization
von: Schliserman, Matan, et al.
Veröffentlicht: (2024) -
The Dimension Strikes Back with Gradients: Generalization of Gradient Methods in Stochastic Convex Optimization
von: Schliserman, Matan, et al.
Veröffentlicht: (2024) -
Multiclass Loss Geometry Matters for Generalization of Gradient Descent in Separable Classification
von: Schliserman, Matan, et al.
Veröffentlicht: (2025) -
Fast Last-Iterate Convergence of SGD in the Smooth Interpolation Regime
von: Attia, Amit, et al.
Veröffentlicht: (2025)