QV May Be Enough: Toward the Essence of Attention in LLMs
Fuente:
arXiv
Guardado en:
| Autor principal: | Edward, Zhang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Greedy Is Enough: Sparse Action Discovery in Agentic LLMs
por: Majumdar, Angshul
Publicado: (2026)
por: Majumdar, Angshul
Publicado: (2026)
Linear Attention is Enough in Spatial-Temporal Forecasting
por: Ning, Xinyu
Publicado: (2024)
por: Ning, Xinyu
Publicado: (2024)
Fitting Is Not Enough: Smoothness in Extremely Quantized LLMs
por: Xu, Yuzhuang, et al.
Publicado: (2026)
por: Xu, Yuzhuang, et al.
Publicado: (2026)
Reward Is Enough: LLMs Are In-Context Reinforcement Learners
por: Song, Kefan, et al.
Publicado: (2025)
por: Song, Kefan, et al.
Publicado: (2025)
On the Essence and Prospect: An Investigation of Alignment Approaches for Big Models
por: Wang, Xinpeng, et al.
Publicado: (2024)
por: Wang, Xinpeng, et al.
Publicado: (2024)
SEUF: Is Unlearning One Expert Enough for Mixture-of-Experts LLMs?
por: Zhuang, Haomin, et al.
Publicado: (2024)
por: Zhuang, Haomin, et al.
Publicado: (2024)
Automatic Feature Learning for Essence: a Case Study on Car Sequencing
por: Pellegrino, Alessio, et al.
Publicado: (2024)
por: Pellegrino, Alessio, et al.
Publicado: (2024)
Reasoning Isn't Enough: Examining Truth-Bias and Sycophancy in LLMs
por: Barkett, Emilio, et al.
Publicado: (2025)
por: Barkett, Emilio, et al.
Publicado: (2025)
Distilling the Essence: Efficient Reasoning Distillation via Sequence Truncation
por: Chen, Wei-Rui, et al.
Publicado: (2025)
por: Chen, Wei-Rui, et al.
Publicado: (2025)
Automating Reformulation of Essence Specifications via Graph Rewriting
por: Miguel, Ian, et al.
Publicado: (2024)
por: Miguel, Ian, et al.
Publicado: (2024)
Toward Better EHR Reasoning in LLMs: Reinforcement Learning with Expert Attention Guidance
por: Fang, Yue, et al.
Publicado: (2025)
por: Fang, Yue, et al.
Publicado: (2025)
When Is Enough Not Enough? Illusory Completion in Search Agents
por: Ko, Dayoon, et al.
Publicado: (2026)
por: Ko, Dayoon, et al.
Publicado: (2026)
Attention's Gravitational Field:A Power-Law Interpretation of Positional Correlation
por: Zhang, Edward
Publicado: (2026)
por: Zhang, Edward
Publicado: (2026)
Enough Coin Flips Can Make LLMs Act Bayesian
por: Gupta, Ritwik, et al.
Publicado: (2025)
por: Gupta, Ritwik, et al.
Publicado: (2025)
When Looking Is Not Enough: Visual Attention Structure Reveals Hallucination in MLLMs
por: Cao, Fanpu, et al.
Publicado: (2026)
por: Cao, Fanpu, et al.
Publicado: (2026)
The Essence of Contextual Understanding in Theory of Mind: A Study on Question Answering with Story Characters
por: Zhou, Chulun, et al.
Publicado: (2025)
por: Zhou, Chulun, et al.
Publicado: (2025)
In-House Evaluation Is Not Enough: Towards Robust Third-Party Flaw Disclosure for General-Purpose AI
por: Longpre, Shayne, et al.
Publicado: (2025)
por: Longpre, Shayne, et al.
Publicado: (2025)
Take its Essence, Discard its Dross! Debiasing for Toxic Language Detection via Counterfactual Causal Effect
por: Lu, Junyu, et al.
Publicado: (2024)
por: Lu, Junyu, et al.
Publicado: (2024)
How Many Experts Are Enough? Towards Optimal Semantic Specialization for Mixture-of-Experts
por: Park, Sumin, et al.
Publicado: (2025)
por: Park, Sumin, et al.
Publicado: (2025)
The Gatekeeper Knows Enough
por: Abebayew, Fikresilase Wondmeneh
Publicado: (2025)
por: Abebayew, Fikresilase Wondmeneh
Publicado: (2025)
Is Distance Matrix Enough for Geometric Deep Learning?
por: Li, Zian, et al.
Publicado: (2023)
por: Li, Zian, et al.
Publicado: (2023)
Single Ground Truth Is Not Enough: Adding Flexibility to Aspect-Based Sentiment Analysis Evaluation
por: Yang, Soyoung, et al.
Publicado: (2024)
por: Yang, Soyoung, et al.
Publicado: (2024)
Integrate the Essence and Eliminate the Dross: Fine-Grained Self-Consistency for Free-Form Language Generation
por: Wang, Xinglin, et al.
Publicado: (2024)
por: Wang, Xinglin, et al.
Publicado: (2024)
Chat-Based Support Alone May Not Be Enough: Comparing Conversational and Embedded LLM Feedback for Mathematical Proof Learning
por: Chen, Eason, et al.
Publicado: (2026)
por: Chen, Eason, et al.
Publicado: (2026)
Screening Is Enough
por: Nakanishi, Ken M.
Publicado: (2026)
por: Nakanishi, Ken M.
Publicado: (2026)
LLMs May Perform MCQA by Selecting the Least Incorrect Option
por: Wang, Haochun, et al.
Publicado: (2024)
por: Wang, Haochun, et al.
Publicado: (2024)
Towards Economical Inference: Enabling DeepSeek's Multi-Head Latent Attention in Any Transformer-based LLMs
por: Ji, Tao, et al.
Publicado: (2025)
por: Ji, Tao, et al.
Publicado: (2025)
Unlocking the Essence of Beauty: Advanced Aesthetic Reasoning with Relative-Absolute Policy Optimization
por: Liu, Boyang, et al.
Publicado: (2025)
por: Liu, Boyang, et al.
Publicado: (2025)
Asking Is Not Enough: Protocol Sensitivity in LLM Confidence Calibration
por: Kim, Hankyeol, et al.
Publicado: (2026)
por: Kim, Hankyeol, et al.
Publicado: (2026)
Are LLMs Enough for Hyperpartisan, Fake, Polarized and Harmful Content Detection? Evaluating In-Context Learning vs. Fine-Tuning
por: Maggini, Michele Joshua, et al.
Publicado: (2025)
por: Maggini, Michele Joshua, et al.
Publicado: (2025)
SLASH the Sink: Sharpening Structural Attention Inside LLMs
por: Liu, Yiming, et al.
Publicado: (2026)
por: Liu, Yiming, et al.
Publicado: (2026)
Agents Are Not Enough
por: Shah, Chirag, et al.
Publicado: (2024)
por: Shah, Chirag, et al.
Publicado: (2024)
State Rank Dynamics in Linear Attention LLMs
por: Sun, Ao, et al.
Publicado: (2026)
por: Sun, Ao, et al.
Publicado: (2026)
Softmax is not Enough (for Adaptive Conformal Classification)
por: Attar, Navid Akhavan, et al.
Publicado: (2026)
por: Attar, Navid Akhavan, et al.
Publicado: (2026)
When Can Model-Free Reinforcement Learning be Enough for Thinking?
por: Hanna, Josiah P., et al.
Publicado: (2025)
por: Hanna, Josiah P., et al.
Publicado: (2025)
Is Implicit Knowledge Enough for LLMs? A RAG Approach for Tree-based Structures
por: Gupte, Mihir, et al.
Publicado: (2025)
por: Gupte, Mihir, et al.
Publicado: (2025)
The Morality of Probability: How Implicit Moral Biases in LLMs May Shape the Future of Human-AI Symbiosis
por: O'Doherty, Eoin, et al.
Publicado: (2025)
por: O'Doherty, Eoin, et al.
Publicado: (2025)
Weight Patching: Toward Source-Level Mechanistic Localization in LLMs
por: Sun, Chenghao, et al.
Publicado: (2026)
por: Sun, Chenghao, et al.
Publicado: (2026)
Learning to Decide with Just Enough: Information-Theoretic Context Summarization for CMDPs
por: Liu, Peidong, et al.
Publicado: (2025)
por: Liu, Peidong, et al.
Publicado: (2025)
When Alignment Isn't Enough: Response-Path Attacks on LLM Agents
por: Luo, Mingyu, et al.
Publicado: (2026)
por: Luo, Mingyu, et al.
Publicado: (2026)
Ejemplares similares
-
Greedy Is Enough: Sparse Action Discovery in Agentic LLMs
por: Majumdar, Angshul
Publicado: (2026) -
Linear Attention is Enough in Spatial-Temporal Forecasting
por: Ning, Xinyu
Publicado: (2024) -
Fitting Is Not Enough: Smoothness in Extremely Quantized LLMs
por: Xu, Yuzhuang, et al.
Publicado: (2026) -
Reward Is Enough: LLMs Are In-Context Reinforcement Learners
por: Song, Kefan, et al.
Publicado: (2025) -
On the Essence and Prospect: An Investigation of Alignment Approaches for Big Models
por: Wang, Xinpeng, et al.
Publicado: (2024)