Are Flat Minima an Illusion?
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Bennett, Michael Timothy |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Zeroth-Order Optimization Finds Flat Minima
von: Zhang, Liang, et al.
Veröffentlicht: (2025)
von: Zhang, Liang, et al.
Veröffentlicht: (2025)
Towards the Connection between Activation Sparsity and Flat Minima
von: Peng, Ze, et al.
Veröffentlicht: (2026)
von: Peng, Ze, et al.
Veröffentlicht: (2026)
SAFE: Finding Sparse and Flat Minima to Improve Pruning
von: Lee, Dongyeop, et al.
Veröffentlicht: (2025)
von: Lee, Dongyeop, et al.
Veröffentlicht: (2025)
A Flat Minima Perspective on Understanding Augmentations and Model Robustness
von: Yoo, Weebum, et al.
Veröffentlicht: (2025)
von: Yoo, Weebum, et al.
Veröffentlicht: (2025)
A Function-Centric Perspective on Flat and Sharp Minima
von: Mason-Williams, Israel, et al.
Veröffentlicht: (2025)
von: Mason-Williams, Israel, et al.
Veröffentlicht: (2025)
Is Complexity an Illusion?
von: Bennett, Michael Timothy
Veröffentlicht: (2024)
von: Bennett, Michael Timothy
Veröffentlicht: (2024)
DP-FedPGN: Finding Global Flat Minima for Differentially Private Federated Learning via Penalizing Gradient Norm
von: Liu, Junkang, et al.
Veröffentlicht: (2025)
von: Liu, Junkang, et al.
Veröffentlicht: (2025)
The Optimal Choice of Hypothesis Is the Weakest, Not the Shortest
von: Bennett, Michael Timothy
Veröffentlicht: (2023)
von: Bennett, Michael Timothy
Veröffentlicht: (2023)
Inference Time Context Sparsity: Illusion or Opportunity?
von: Joshi, Sahil, et al.
Veröffentlicht: (2026)
von: Joshi, Sahil, et al.
Veröffentlicht: (2026)
Nested Learning: The Illusion of Deep Learning Architectures
von: Behrouz, Ali, et al.
Veröffentlicht: (2025)
von: Behrouz, Ali, et al.
Veröffentlicht: (2025)
The Illusion of Certainty: Decoupling Capability and Calibration in On-Policy Distillation
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2026)
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2026)
The Leaderboard Illusion
von: Singh, Shivalika, et al.
Veröffentlicht: (2025)
von: Singh, Shivalika, et al.
Veröffentlicht: (2025)
Beyond Sharp Minima: Robust LLM Unlearning via Feedback-Guided Multi-Point Optimization
von: Wu, Wenhan, et al.
Veröffentlicht: (2025)
von: Wu, Wenhan, et al.
Veröffentlicht: (2025)
Stable Minima Cannot Overfit in Univariate ReLU Networks: Generalization by Large Step Sizes
von: Qiao, Dan, et al.
Veröffentlicht: (2024)
von: Qiao, Dan, et al.
Veröffentlicht: (2024)
A Unified Theory of Sparse Dictionary Learning in Mechanistic Interpretability: Piecewise Biconvexity and Spurious Minima
von: Tang, Yiming, et al.
Veröffentlicht: (2025)
von: Tang, Yiming, et al.
Veröffentlicht: (2025)
The Illusion of Readiness in Health AI
von: Gu, Yu, et al.
Veröffentlicht: (2025)
von: Gu, Yu, et al.
Veröffentlicht: (2025)
Adversarial Illusions in Multi-Modal Embeddings
von: Zhang, Tingwei, et al.
Veröffentlicht: (2023)
von: Zhang, Tingwei, et al.
Veröffentlicht: (2023)
The Illusion of Specialization: Unveiling the Domain-Invariant "Standing Committee" in Mixture-of-Experts Models
von: Wang, Yan, et al.
Veröffentlicht: (2026)
von: Wang, Yan, et al.
Veröffentlicht: (2026)
The Illusion of Equivalence: Systematic FP16 Divergence in KV-Cached Autoregressive Inference
von: Chodavarapu, Ranjith, et al.
Veröffentlicht: (2026)
von: Chodavarapu, Ranjith, et al.
Veröffentlicht: (2026)
Bridging Reasoning to Learning: Unmasking Illusions using Complexity Out of Distribution Generalization
von: Paqaleh, Mohammad Mahdi Samiei, et al.
Veröffentlicht: (2025)
von: Paqaleh, Mohammad Mahdi Samiei, et al.
Veröffentlicht: (2025)
Theory-optimal Quantization Based on Flatness
von: Huang, Xiusheng, et al.
Veröffentlicht: (2026)
von: Huang, Xiusheng, et al.
Veröffentlicht: (2026)
The Illusion of Reasoning: Exposing Evasive Data Contamination in LLMs via Zero-CoT Truncation
von: Lan, Yifan, et al.
Veröffentlicht: (2026)
von: Lan, Yifan, et al.
Veröffentlicht: (2026)
Comment on The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity
von: Lawsen, A.
Veröffentlicht: (2025)
von: Lawsen, A.
Veröffentlicht: (2025)
The Illusion of Progress: Re-evaluating Hallucination Detection in LLMs
von: Janiak, Denis, et al.
Veröffentlicht: (2025)
von: Janiak, Denis, et al.
Veröffentlicht: (2025)
Illusions of reflection: open-ended task reveals systematic failures in Large Language Models' reflective reasoning
von: Weatherhead, Sion, et al.
Veröffentlicht: (2025)
von: Weatherhead, Sion, et al.
Veröffentlicht: (2025)
Not All Latent Spaces Are Flat: Hyperbolic Concept Control
von: Briglia, Maria Rosaria, et al.
Veröffentlicht: (2026)
von: Briglia, Maria Rosaria, et al.
Veröffentlicht: (2026)
Fairness is Not Flat: Geometric Phase Transitions Against Shortcut Learning
von: Rodriguez-Alvarez, Nicolas, et al.
Veröffentlicht: (2026)
von: Rodriguez-Alvarez, Nicolas, et al.
Veröffentlicht: (2026)
FedNSAM:Consistency of Local and Global Flatness for Federated Learning
von: Liu, Junkang, et al.
Veröffentlicht: (2026)
von: Liu, Junkang, et al.
Veröffentlicht: (2026)
Are UFOs Driving Innovation? The Illusion of Causality in Large Language Models
von: Carro, María Victoria, et al.
Veröffentlicht: (2024)
von: Carro, María Victoria, et al.
Veröffentlicht: (2024)
The Erasure Illusion: Stress-Testing the Generalization of LLM Forgetting Evaluation
von: Jia, Hengrui, et al.
Veröffentlicht: (2025)
von: Jia, Hengrui, et al.
Veröffentlicht: (2025)
A Reply to Makelov et al. (2023)'s "Interpretability Illusion" Arguments
von: Wu, Zhengxuan, et al.
Veröffentlicht: (2024)
von: Wu, Zhengxuan, et al.
Veröffentlicht: (2024)
A Comment On "The Illusion of Thinking": Reframing the Reasoning Cliff as an Agentic Gap
von: Khan, Sheraz, et al.
Veröffentlicht: (2025)
von: Khan, Sheraz, et al.
Veröffentlicht: (2025)
The Semantic Illusion: Certified Limits of Embedding-Based Hallucination Detection in RAG Systems
von: Sinha, Debu
Veröffentlicht: (2025)
von: Sinha, Debu
Veröffentlicht: (2025)
Consolidating TinyML Lifecycle with Large Language Models: Reality, Illusion, or Opportunity?
von: Wu, Guanghan, et al.
Veröffentlicht: (2025)
von: Wu, Guanghan, et al.
Veröffentlicht: (2025)
From Flat Facts to Sharp Hallucinations: Detecting Stubborn Errors via Gradient Sensitivity
von: Liew, Yee Zhing, et al.
Veröffentlicht: (2026)
von: Liew, Yee Zhing, et al.
Veröffentlicht: (2026)
The Personality Illusion: Revealing Dissociation Between Self-Reports & Behavior in LLMs
von: Han, Pengrui, et al.
Veröffentlicht: (2025)
von: Han, Pengrui, et al.
Veröffentlicht: (2025)
The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity
von: Shojaee, Parshin, et al.
Veröffentlicht: (2025)
von: Shojaee, Parshin, et al.
Veröffentlicht: (2025)
Does Flatness imply Generalization for Logistic Loss in Univariate Two-Layer ReLU Network?
von: Qiao, Dan, et al.
Veröffentlicht: (2025)
von: Qiao, Dan, et al.
Veröffentlicht: (2025)
Low-Rank MDPs with Continuous Action Spaces
von: Bennett, Andrew, et al.
Veröffentlicht: (2023)
von: Bennett, Andrew, et al.
Veröffentlicht: (2023)
The Narcissus Hypothesis: Descending to the Rung of Illusion
von: Cadei, Riccardo, et al.
Veröffentlicht: (2025)
von: Cadei, Riccardo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Zeroth-Order Optimization Finds Flat Minima
von: Zhang, Liang, et al.
Veröffentlicht: (2025) -
Towards the Connection between Activation Sparsity and Flat Minima
von: Peng, Ze, et al.
Veröffentlicht: (2026) -
SAFE: Finding Sparse and Flat Minima to Improve Pruning
von: Lee, Dongyeop, et al.
Veröffentlicht: (2025) -
A Flat Minima Perspective on Understanding Augmentations and Model Robustness
von: Yoo, Weebum, et al.
Veröffentlicht: (2025) -
A Function-Centric Perspective on Flat and Sharp Minima
von: Mason-Williams, Israel, et al.
Veröffentlicht: (2025)