Sharpness-Aware Minimization Can Hallucinate Minimizers
Fuente:
arXiv
Saved in:
| Main Authors: | Park, Chanwoong, Jang, Uijeong, Ryu, Ernest K., Yang, Insoon |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Generalized Continuous-Time Models for Nesterov's Accelerated Gradient Methods
by: Park, Chanwoong, et al.
Published: (2024)
by: Park, Chanwoong, et al.
Published: (2024)
LoRA Training in the NTK Regime has No Spurious Local Minima
by: Jang, Uijeong, et al.
Published: (2024)
by: Jang, Uijeong, et al.
Published: (2024)
Optimal Acceleration for Proximal Minimization of the Sum of Convex and Strongly Convex Functions
by: Chari, Govind M., et al.
Published: (2026)
by: Chari, Govind M., et al.
Published: (2026)
Point Convergence of Nesterov's Accelerated Gradient Method: An AI-Assisted Proof
by: Jang, Uijeong, et al.
Published: (2025)
by: Jang, Uijeong, et al.
Published: (2025)
Sharpness-Aware Minimization: General Analysis and Improved Rates
by: Oikonomou, Dimitris, et al.
Published: (2025)
by: Oikonomou, Dimitris, et al.
Published: (2025)
Computer-Assisted Design of Accelerated Composite Optimization Methods: OptISTA
by: Jang, Uijeong, et al.
Published: (2023)
by: Jang, Uijeong, et al.
Published: (2023)
DGSAM: Domain Generalization via Individual Sharpness-Aware Minimization
by: Song, Youngjun, et al.
Published: (2025)
by: Song, Youngjun, et al.
Published: (2025)
ALiA: Adaptive Linearized ADMM
by: Jang, Uijeong, et al.
Published: (2026)
by: Jang, Uijeong, et al.
Published: (2026)
Critical Influence of Overparameterization on Sharpness-aware Minimization
by: Shin, Sungbin, et al.
Published: (2023)
by: Shin, Sungbin, et al.
Published: (2023)
Convergence of Sharpness-Aware Minimization Algorithms using Increasing Batch Size and Decaying Learning Rate
by: Harada, Hinata, et al.
Published: (2024)
by: Harada, Hinata, et al.
Published: (2024)
Adaptive Sharpness-Aware Minimization with a Polyak-type Step size: A Theory-Grounded Scheduler
by: Oikonomou, Dimitris, et al.
Published: (2026)
by: Oikonomou, Dimitris, et al.
Published: (2026)
Nesterov Flow May Travel Infinitely Long to Converge to a Minimizer
by: Ryu, Ernest K.
Published: (2026)
by: Ryu, Ernest K.
Published: (2026)
On the Duality Between Sharpness-Aware Minimization and Adversarial Training
by: Zhang, Yihao, et al.
Published: (2024)
by: Zhang, Yihao, et al.
Published: (2024)
Dynamic Regularized Sharpness Aware Minimization in Federated Learning: Approaching Global Consistency and Smooth Landscape
by: Sun, Yan, et al.
Published: (2023)
by: Sun, Yan, et al.
Published: (2023)
Deflated Dynamics Value Iteration
by: Lee, Jongmin, et al.
Published: (2024)
by: Lee, Jongmin, et al.
Published: (2024)
Parameter-Free Algorithms for Performative Regret Minimization under Decision-Dependent Distributions
by: Park, Sungwoo, et al.
Published: (2024)
by: Park, Sungwoo, et al.
Published: (2024)
Efficient Continual Finite-Sum Minimization
by: Mavrothalassitis, Ioannis, et al.
Published: (2024)
by: Mavrothalassitis, Ioannis, et al.
Published: (2024)
Global Optimization via Softmin Energy Minimization
by: Agazzi, Andrea, et al.
Published: (2025)
by: Agazzi, Andrea, et al.
Published: (2025)
Optimal Algorithms for Stochastic Complementary Composite Minimization
by: d'Aspremont, Alexandre, et al.
Published: (2022)
by: d'Aspremont, Alexandre, et al.
Published: (2022)
An Equivalence Between Static and Dynamic Regret Minimization
by: Jacobsen, Andrew, et al.
Published: (2024)
by: Jacobsen, Andrew, et al.
Published: (2024)
Fundamental Convergence Analysis of Sharpness-Aware Minimization
by: Khanh, Pham Duy, et al.
Published: (2024)
by: Khanh, Pham Duy, et al.
Published: (2024)
Efficient Low-rank Identification via Accelerated Iteratively Reweighted Nuclear Norm Minimization
by: Wang, Hao, et al.
Published: (2024)
by: Wang, Hao, et al.
Published: (2024)
SOREL: A Stochastic Algorithm for Spectral Risks Minimization
by: Ge, Yuze, et al.
Published: (2024)
by: Ge, Yuze, et al.
Published: (2024)
Distributionally Robust Kalman Filter
by: Jang, Minhyuk, et al.
Published: (2025)
by: Jang, Minhyuk, et al.
Published: (2025)
Wasserstein Distributionally Robust Control and State Estimation for Partially Observable Linear Systems
by: Jang, Minhyuk, et al.
Published: (2024)
by: Jang, Minhyuk, et al.
Published: (2024)
Convergence Properties of Natural Gradient Descent for Minimizing KL Divergence
by: Datar, Adwait, et al.
Published: (2025)
by: Datar, Adwait, et al.
Published: (2025)
Inertial Quadratic Majorization Minimization with Application to Kernel Regularized Learning
by: Heng, Qiang, et al.
Published: (2025)
by: Heng, Qiang, et al.
Published: (2025)
Efficient Alternating Minimization with Applications to Weighted Low Rank Approximation
by: Song, Zhao, et al.
Published: (2023)
by: Song, Zhao, et al.
Published: (2023)
A Minimal-Assumption Analysis of Q-Learning with Time-Varying Policies
by: Nanda, Phalguni, et al.
Published: (2025)
by: Nanda, Phalguni, et al.
Published: (2025)
Specifying and Solving Robust Empirical Risk Minimization Problems Using CVXPY
by: Luxenberg, Eric, et al.
Published: (2023)
by: Luxenberg, Eric, et al.
Published: (2023)
StoTAM: Stochastic Alternating Minimization for Tucker-Structured Tensor Sensing
by: Li, Shuang
Published: (2026)
by: Li, Shuang
Published: (2026)
ADMM for Structured Fractional Minimization
by: Yuan, Ganzhao
Published: (2024)
by: Yuan, Ganzhao
Published: (2024)
Fast Stochastic Composite Minimization and an Accelerated Frank-Wolfe Algorithm under Parallelization
by: Dubois-Taine, Benjamin, et al.
Published: (2022)
by: Dubois-Taine, Benjamin, et al.
Published: (2022)
Exact Instance Compression for Convex Empirical Risk Minimization via Color Refinement
by: Zhu, Bryan, et al.
Published: (2026)
by: Zhu, Bryan, et al.
Published: (2026)
Minimizing UCB: a Better Local Search Strategy in Local Bayesian Optimization
by: Fan, Zheyi, et al.
Published: (2024)
by: Fan, Zheyi, et al.
Published: (2024)
Stochastic Variance-Reduced Newton: Accelerating Finite-Sum Minimization with Large Batches
by: Dereziński, Michał
Published: (2022)
by: Dereziński, Michał
Published: (2022)
Inclusive KL Minimization: A Wasserstein-Fisher-Rao Gradient Flow Perspective
by: Zhu, Jia-Jie
Published: (2024)
by: Zhu, Jia-Jie
Published: (2024)
Minimizing the Weighted Number of Tardy Jobs: Data-Driven Heuristic for Single-Machine Scheduling
by: Antonov, Nikolai, et al.
Published: (2025)
by: Antonov, Nikolai, et al.
Published: (2025)
Variable Bregman Majorization-Minimization Algorithm and its Application to Dirichlet Maximum Likelihood Estimation
by: Martin, Ségolène, et al.
Published: (2025)
by: Martin, Ségolène, et al.
Published: (2025)
Local LMO: Constrained Gradient Optimization via a Local Linear Minimization Oracle
by: Richtárik, Peter, et al.
Published: (2026)
by: Richtárik, Peter, et al.
Published: (2026)
Similar Items
-
Generalized Continuous-Time Models for Nesterov's Accelerated Gradient Methods
by: Park, Chanwoong, et al.
Published: (2024) -
LoRA Training in the NTK Regime has No Spurious Local Minima
by: Jang, Uijeong, et al.
Published: (2024) -
Optimal Acceleration for Proximal Minimization of the Sum of Convex and Strongly Convex Functions
by: Chari, Govind M., et al.
Published: (2026) -
Point Convergence of Nesterov's Accelerated Gradient Method: An AI-Assisted Proof
by: Jang, Uijeong, et al.
Published: (2025) -
Sharpness-Aware Minimization: General Analysis and Improved Rates
by: Oikonomou, Dimitris, et al.
Published: (2025)