Navigating Potholes with Geometry-Aware Sharpness Minimization
Fuente:
arXiv
Saved in:
| Main Authors: | Dufort-Labbé, Simon, Hamidi, Mehrab, Pascanu, Razvan, Mitliagkas, Ioannis, Scieur, Damien, Baratin, Aristide |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Layerwise LQR for Geometry-Aware Optimization of Deep Networks
by: Dufort-Labbé, Simon, et al.
Published: (2026)
by: Dufort-Labbé, Simon, et al.
Published: (2026)
Maxwell's Demon at Work: Efficient Pruning by Leveraging Saturation of Neurons
by: Dufort-Labbé, Simon, et al.
Published: (2024)
by: Dufort-Labbé, Simon, et al.
Published: (2024)
Torque-Aware Momentum
by: Malviya, Pranshu, et al.
Published: (2024)
by: Malviya, Pranshu, et al.
Published: (2024)
Promoting Exploration in Memory-Augmented Adam using Critical Momenta
by: Malviya, Pranshu, et al.
Published: (2023)
by: Malviya, Pranshu, et al.
Published: (2023)
Understanding Adam Requires Better Rotation Dependent Assumptions
by: Zhang, Tianyue H., et al.
Published: (2024)
by: Zhang, Tianyue H., et al.
Published: (2024)
Softmax is not Enough (for Sharp Size Generalisation)
by: Veličković, Petar, et al.
Published: (2024)
by: Veličković, Petar, et al.
Published: (2024)
Lookbehind-SAM: k steps back, 1 step forward
by: Mordido, Gonçalo, et al.
Published: (2023)
by: Mordido, Gonçalo, et al.
Published: (2023)
Compositional Risk Minimization
by: Mahajan, Divyat, et al.
Published: (2024)
by: Mahajan, Divyat, et al.
Published: (2024)
Lattice: Learning to Efficiently Compress the Memory
by: Karami, Mahdi, et al.
Published: (2025)
by: Karami, Mahdi, et al.
Published: (2025)
Revisiting Adam for Streaming Reinforcement Learning
by: Gogianu, Florin, et al.
Published: (2026)
by: Gogianu, Florin, et al.
Published: (2026)
Meta-learning how to Share Credit among Macro-Actions
by: Hosu, Ionel-Alexandru, et al.
Published: (2025)
by: Hosu, Ionel-Alexandru, et al.
Published: (2025)
Feature learning as alignment: a structural property of gradient descent in non-linear neural networks
by: Beaglehole, Daniel, et al.
Published: (2024)
by: Beaglehole, Daniel, et al.
Published: (2024)
Stability Analysis of Sharpness-Aware Minimization
by: Kim, Hoki, et al.
Published: (2023)
by: Kim, Hoki, et al.
Published: (2023)
Perplexity Cannot Always Tell Right from Wrong
by: Veličković, Petar, et al.
Published: (2026)
by: Veličković, Petar, et al.
Published: (2026)
An Empirical Study of Pre-trained Model Selection for Out-of-Distribution Generalization and Calibration
by: Naganuma, Hiroki, et al.
Published: (2023)
by: Naganuma, Hiroki, et al.
Published: (2023)
Towards efficient representation identification in supervised learning
by: Ahuja, Kartik, et al.
Published: (2022)
by: Ahuja, Kartik, et al.
Published: (2022)
Membership Privacy Risks of Sharpness Aware Minimization
by: Kim, Young In, et al.
Published: (2023)
by: Kim, Young In, et al.
Published: (2023)
LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
by: Maes, Lucas, et al.
Published: (2026)
by: Maes, Lucas, et al.
Published: (2026)
Towards Understanding The Calibration Benefits of Sharpness-Aware Minimization
by: Tan, Chengli, et al.
Published: (2025)
by: Tan, Chengli, et al.
Published: (2025)
Empirical Analysis of Model Selection for Heterogeneous Causal Effect Estimation
by: Mahajan, Divyat, et al.
Published: (2022)
by: Mahajan, Divyat, et al.
Published: (2022)
LLMs are Greedy Agents: Effects of RL Fine-tuning on Decision-Making Abilities
by: Schmied, Thomas, et al.
Published: (2025)
by: Schmied, Thomas, et al.
Published: (2025)
State Soup: In-Context Skill Learning, Retrieval and Mixing
by: Pióro, Maciej, et al.
Published: (2024)
by: Pióro, Maciej, et al.
Published: (2024)
SAMOSA: Sharpness Aware Minimization for Open Set Active learning
by: Kim, Young In, et al.
Published: (2025)
by: Kim, Young In, et al.
Published: (2025)
Performance Control in Early Exiting to Deploy Large Models at the Same Cost of Smaller Ones
by: Mofakhami, Mehrnaz, et al.
Published: (2024)
by: Mofakhami, Mehrnaz, et al.
Published: (2024)
Revisiting Sharpness-Aware Minimization: A More Faithful and Effective Implementation
by: Chen, Jianlong, et al.
Published: (2026)
by: Chen, Jianlong, et al.
Published: (2026)
Mining Generalizable Activation Functions
by: Vitvitskyi, Alex, et al.
Published: (2026)
by: Vitvitskyi, Alex, et al.
Published: (2026)
Unpacking the Implicit Norm Dynamics of Sharpness-Aware Minimization in Tensorized Models
by: Cao, Tianxiao, et al.
Published: (2025)
by: Cao, Tianxiao, et al.
Published: (2025)
Retrieval-Augmented Decision Transformer: External Memory for In-context RL
by: Schmied, Thomas, et al.
Published: (2024)
by: Schmied, Thomas, et al.
Published: (2024)
PotholeGuard: A Pothole Detection Approach by Point Cloud Semantic Segmentation
by: Nawale, Sahil, et al.
Published: (2023)
by: Nawale, Sahil, et al.
Published: (2023)
X-SAM: Boosting Sharpness-Aware Minimization with Dominant-Eigenvector Gradient Correction
by: Duan, Hongru, et al.
Published: (2026)
by: Duan, Hongru, et al.
Published: (2026)
Sharpness-Aware Minimization in Logit Space Efficiently Enhances Direct Preference Optimization
by: Luo, Haocheng, et al.
Published: (2026)
by: Luo, Haocheng, et al.
Published: (2026)
FairSAM: Fair Classification on Corrupted Data Through Sharpness-Aware Minimization
by: Dai, Yucong, et al.
Published: (2025)
by: Dai, Yucong, et al.
Published: (2025)
Towards Understanding the Role of Sharpness-Aware Minimization Algorithms for Out-of-Distribution Generalization
by: Schapiro, Samuel, et al.
Published: (2024)
by: Schapiro, Samuel, et al.
Published: (2024)
Attributing Data for Sharpness-Aware Minimization
by: Ren, Chenyang, et al.
Published: (2025)
by: Ren, Chenyang, et al.
Published: (2025)
Fast Graph Sharpness-Aware Minimization for Enhancing and Accelerating Few-Shot Node Classification
by: Luo, Yihong, et al.
Published: (2024)
by: Luo, Yihong, et al.
Published: (2024)
Fine-Tuned In-Context Learners for Efficient Adaptation
by: Bornschein, Jorg, et al.
Published: (2025)
by: Bornschein, Jorg, et al.
Published: (2025)
Normalization and effective learning rates in reinforcement learning
by: Lyle, Clare, et al.
Published: (2024)
by: Lyle, Clare, et al.
Published: (2024)
Minor First, Major Last: A Depth-Induced Implicit Bias of Sharpness-Aware Minimization
by: Moon, Chaewon, et al.
Published: (2026)
by: Moon, Chaewon, et al.
Published: (2026)
Bi-LoRA: Efficient Sharpness-Aware Minimization for Fine-Tuning Large-Scale Models
by: Liu, Yuhang, et al.
Published: (2025)
by: Liu, Yuhang, et al.
Published: (2025)
Enhancing Robustness of Offline Reinforcement Learning Under Data Corruption via Sharpness-Aware Minimization
by: Xu, Le, et al.
Published: (2025)
by: Xu, Le, et al.
Published: (2025)
Similar Items
-
Layerwise LQR for Geometry-Aware Optimization of Deep Networks
by: Dufort-Labbé, Simon, et al.
Published: (2026) -
Maxwell's Demon at Work: Efficient Pruning by Leveraging Saturation of Neurons
by: Dufort-Labbé, Simon, et al.
Published: (2024) -
Torque-Aware Momentum
by: Malviya, Pranshu, et al.
Published: (2024) -
Promoting Exploration in Memory-Augmented Adam using Critical Momenta
by: Malviya, Pranshu, et al.
Published: (2023) -
Understanding Adam Requires Better Rotation Dependent Assumptions
by: Zhang, Tianyue H., et al.
Published: (2024)