QV May Be Enough: Toward the Essence of Attention in LLMs
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Edward, Zhang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Greedy Is Enough: Sparse Action Discovery in Agentic LLMs
von: Majumdar, Angshul
Veröffentlicht: (2026)
von: Majumdar, Angshul
Veröffentlicht: (2026)
Linear Attention is Enough in Spatial-Temporal Forecasting
von: Ning, Xinyu
Veröffentlicht: (2024)
von: Ning, Xinyu
Veröffentlicht: (2024)
Fitting Is Not Enough: Smoothness in Extremely Quantized LLMs
von: Xu, Yuzhuang, et al.
Veröffentlicht: (2026)
von: Xu, Yuzhuang, et al.
Veröffentlicht: (2026)
Reward Is Enough: LLMs Are In-Context Reinforcement Learners
von: Song, Kefan, et al.
Veröffentlicht: (2025)
von: Song, Kefan, et al.
Veröffentlicht: (2025)
On the Essence and Prospect: An Investigation of Alignment Approaches for Big Models
von: Wang, Xinpeng, et al.
Veröffentlicht: (2024)
von: Wang, Xinpeng, et al.
Veröffentlicht: (2024)
SEUF: Is Unlearning One Expert Enough for Mixture-of-Experts LLMs?
von: Zhuang, Haomin, et al.
Veröffentlicht: (2024)
von: Zhuang, Haomin, et al.
Veröffentlicht: (2024)
Automatic Feature Learning for Essence: a Case Study on Car Sequencing
von: Pellegrino, Alessio, et al.
Veröffentlicht: (2024)
von: Pellegrino, Alessio, et al.
Veröffentlicht: (2024)
Reasoning Isn't Enough: Examining Truth-Bias and Sycophancy in LLMs
von: Barkett, Emilio, et al.
Veröffentlicht: (2025)
von: Barkett, Emilio, et al.
Veröffentlicht: (2025)
Distilling the Essence: Efficient Reasoning Distillation via Sequence Truncation
von: Chen, Wei-Rui, et al.
Veröffentlicht: (2025)
von: Chen, Wei-Rui, et al.
Veröffentlicht: (2025)
Automating Reformulation of Essence Specifications via Graph Rewriting
von: Miguel, Ian, et al.
Veröffentlicht: (2024)
von: Miguel, Ian, et al.
Veröffentlicht: (2024)
Toward Better EHR Reasoning in LLMs: Reinforcement Learning with Expert Attention Guidance
von: Fang, Yue, et al.
Veröffentlicht: (2025)
von: Fang, Yue, et al.
Veröffentlicht: (2025)
When Is Enough Not Enough? Illusory Completion in Search Agents
von: Ko, Dayoon, et al.
Veröffentlicht: (2026)
von: Ko, Dayoon, et al.
Veröffentlicht: (2026)
Attention's Gravitational Field:A Power-Law Interpretation of Positional Correlation
von: Zhang, Edward
Veröffentlicht: (2026)
von: Zhang, Edward
Veröffentlicht: (2026)
Enough Coin Flips Can Make LLMs Act Bayesian
von: Gupta, Ritwik, et al.
Veröffentlicht: (2025)
von: Gupta, Ritwik, et al.
Veröffentlicht: (2025)
When Looking Is Not Enough: Visual Attention Structure Reveals Hallucination in MLLMs
von: Cao, Fanpu, et al.
Veröffentlicht: (2026)
von: Cao, Fanpu, et al.
Veröffentlicht: (2026)
The Essence of Contextual Understanding in Theory of Mind: A Study on Question Answering with Story Characters
von: Zhou, Chulun, et al.
Veröffentlicht: (2025)
von: Zhou, Chulun, et al.
Veröffentlicht: (2025)
In-House Evaluation Is Not Enough: Towards Robust Third-Party Flaw Disclosure for General-Purpose AI
von: Longpre, Shayne, et al.
Veröffentlicht: (2025)
von: Longpre, Shayne, et al.
Veröffentlicht: (2025)
Take its Essence, Discard its Dross! Debiasing for Toxic Language Detection via Counterfactual Causal Effect
von: Lu, Junyu, et al.
Veröffentlicht: (2024)
von: Lu, Junyu, et al.
Veröffentlicht: (2024)
How Many Experts Are Enough? Towards Optimal Semantic Specialization for Mixture-of-Experts
von: Park, Sumin, et al.
Veröffentlicht: (2025)
von: Park, Sumin, et al.
Veröffentlicht: (2025)
The Gatekeeper Knows Enough
von: Abebayew, Fikresilase Wondmeneh
Veröffentlicht: (2025)
von: Abebayew, Fikresilase Wondmeneh
Veröffentlicht: (2025)
Is Distance Matrix Enough for Geometric Deep Learning?
von: Li, Zian, et al.
Veröffentlicht: (2023)
von: Li, Zian, et al.
Veröffentlicht: (2023)
Single Ground Truth Is Not Enough: Adding Flexibility to Aspect-Based Sentiment Analysis Evaluation
von: Yang, Soyoung, et al.
Veröffentlicht: (2024)
von: Yang, Soyoung, et al.
Veröffentlicht: (2024)
Integrate the Essence and Eliminate the Dross: Fine-Grained Self-Consistency for Free-Form Language Generation
von: Wang, Xinglin, et al.
Veröffentlicht: (2024)
von: Wang, Xinglin, et al.
Veröffentlicht: (2024)
Chat-Based Support Alone May Not Be Enough: Comparing Conversational and Embedded LLM Feedback for Mathematical Proof Learning
von: Chen, Eason, et al.
Veröffentlicht: (2026)
von: Chen, Eason, et al.
Veröffentlicht: (2026)
Screening Is Enough
von: Nakanishi, Ken M.
Veröffentlicht: (2026)
von: Nakanishi, Ken M.
Veröffentlicht: (2026)
LLMs May Perform MCQA by Selecting the Least Incorrect Option
von: Wang, Haochun, et al.
Veröffentlicht: (2024)
von: Wang, Haochun, et al.
Veröffentlicht: (2024)
Towards Economical Inference: Enabling DeepSeek's Multi-Head Latent Attention in Any Transformer-based LLMs
von: Ji, Tao, et al.
Veröffentlicht: (2025)
von: Ji, Tao, et al.
Veröffentlicht: (2025)
Unlocking the Essence of Beauty: Advanced Aesthetic Reasoning with Relative-Absolute Policy Optimization
von: Liu, Boyang, et al.
Veröffentlicht: (2025)
von: Liu, Boyang, et al.
Veröffentlicht: (2025)
Asking Is Not Enough: Protocol Sensitivity in LLM Confidence Calibration
von: Kim, Hankyeol, et al.
Veröffentlicht: (2026)
von: Kim, Hankyeol, et al.
Veröffentlicht: (2026)
Are LLMs Enough for Hyperpartisan, Fake, Polarized and Harmful Content Detection? Evaluating In-Context Learning vs. Fine-Tuning
von: Maggini, Michele Joshua, et al.
Veröffentlicht: (2025)
von: Maggini, Michele Joshua, et al.
Veröffentlicht: (2025)
SLASH the Sink: Sharpening Structural Attention Inside LLMs
von: Liu, Yiming, et al.
Veröffentlicht: (2026)
von: Liu, Yiming, et al.
Veröffentlicht: (2026)
Agents Are Not Enough
von: Shah, Chirag, et al.
Veröffentlicht: (2024)
von: Shah, Chirag, et al.
Veröffentlicht: (2024)
State Rank Dynamics in Linear Attention LLMs
von: Sun, Ao, et al.
Veröffentlicht: (2026)
von: Sun, Ao, et al.
Veröffentlicht: (2026)
Softmax is not Enough (for Adaptive Conformal Classification)
von: Attar, Navid Akhavan, et al.
Veröffentlicht: (2026)
von: Attar, Navid Akhavan, et al.
Veröffentlicht: (2026)
When Can Model-Free Reinforcement Learning be Enough for Thinking?
von: Hanna, Josiah P., et al.
Veröffentlicht: (2025)
von: Hanna, Josiah P., et al.
Veröffentlicht: (2025)
Is Implicit Knowledge Enough for LLMs? A RAG Approach for Tree-based Structures
von: Gupte, Mihir, et al.
Veröffentlicht: (2025)
von: Gupte, Mihir, et al.
Veröffentlicht: (2025)
The Morality of Probability: How Implicit Moral Biases in LLMs May Shape the Future of Human-AI Symbiosis
von: O'Doherty, Eoin, et al.
Veröffentlicht: (2025)
von: O'Doherty, Eoin, et al.
Veröffentlicht: (2025)
Weight Patching: Toward Source-Level Mechanistic Localization in LLMs
von: Sun, Chenghao, et al.
Veröffentlicht: (2026)
von: Sun, Chenghao, et al.
Veröffentlicht: (2026)
Learning to Decide with Just Enough: Information-Theoretic Context Summarization for CMDPs
von: Liu, Peidong, et al.
Veröffentlicht: (2025)
von: Liu, Peidong, et al.
Veröffentlicht: (2025)
When Alignment Isn't Enough: Response-Path Attacks on LLM Agents
von: Luo, Mingyu, et al.
Veröffentlicht: (2026)
von: Luo, Mingyu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Greedy Is Enough: Sparse Action Discovery in Agentic LLMs
von: Majumdar, Angshul
Veröffentlicht: (2026) -
Linear Attention is Enough in Spatial-Temporal Forecasting
von: Ning, Xinyu
Veröffentlicht: (2024) -
Fitting Is Not Enough: Smoothness in Extremely Quantized LLMs
von: Xu, Yuzhuang, et al.
Veröffentlicht: (2026) -
Reward Is Enough: LLMs Are In-Context Reinforcement Learners
von: Song, Kefan, et al.
Veröffentlicht: (2025) -
On the Essence and Prospect: An Investigation of Alignment Approaches for Big Models
von: Wang, Xinpeng, et al.
Veröffentlicht: (2024)