Attention Consistency for LLMs Explanation
Fuente:
arXiv
Saved in:
| Main Authors: | Lan, Tian, Xu, Jinyuan, He, Xue, Hwang, Jenq-Neng, Li, Lei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bayesian Optimization for Controlled Image Editing via LLMs
by: Cai, Chengkun, et al.
Published: (2025)
by: Cai, Chengkun, et al.
Published: (2025)
Modeling LLM Agent Reviewer Dynamics in Elo-Ranked Review System
by: Huang, Hsiang-Wei, et al.
Published: (2026)
by: Huang, Hsiang-Wei, et al.
Published: (2026)
Argument-Based Consistency in Toxicity Explanations of LLMs
by: Mothilal, Ramaravind Kommiya, et al.
Published: (2025)
by: Mothilal, Ramaravind Kommiya, et al.
Published: (2025)
The Role of Deductive and Inductive Reasoning in Large Language Models
by: Cai, Chengkun, et al.
Published: (2024)
by: Cai, Chengkun, et al.
Published: (2024)
CNSocialDepress: A Chinese Social Media Dataset for Depression Risk Detection and Structured Analysis
by: Xu, Jinyuan, et al.
Published: (2025)
by: Xu, Jinyuan, et al.
Published: (2025)
Intrinsic Entropy of Context Length Scaling in LLMs
by: Shi, Jingzhe, et al.
Published: (2025)
by: Shi, Jingzhe, et al.
Published: (2025)
Towards Consistent Natural-Language Explanations via Explanation-Consistency Finetuning
by: Chen, Yanda, et al.
Published: (2024)
by: Chen, Yanda, et al.
Published: (2024)
Learning to Learn Weight Generation via Local Consistency Diffusion
by: Guan, Yunchuan, et al.
Published: (2025)
by: Guan, Yunchuan, et al.
Published: (2025)
Aligning What LLMs Do and Say: Towards Self-Consistent Explanations
by: Admoni, Sahar, et al.
Published: (2025)
by: Admoni, Sahar, et al.
Published: (2025)
Unsupervised Text Style Transfer via LLMs and Attention Masking with Multi-way Interactions
by: Pan, Lei, et al.
Published: (2024)
by: Pan, Lei, et al.
Published: (2024)
Recent Advances in Embedding Methods for Multi-Object Tracking: A Survey
by: Wang, Gaoang, et al.
Published: (2022)
by: Wang, Gaoang, et al.
Published: (2022)
Explanation Generation for Contradiction Reconciliation with LLMs
by: Chan, Jason, et al.
Published: (2026)
by: Chan, Jason, et al.
Published: (2026)
Towards Faithful Explanations for Text Classification with Robustness Improvement and Explanation Guided Training
by: Li, Dongfang, et al.
Published: (2023)
by: Li, Dongfang, et al.
Published: (2023)
DRIV-EX: Counterfactual Explanations for Driving LLMs
by: Cardiel, Amaia, et al.
Published: (2026)
by: Cardiel, Amaia, et al.
Published: (2026)
Rhea: Role-aware Heuristic Episodic Attention for Conversational LLMs
by: Hong, Wanyang, et al.
Published: (2025)
by: Hong, Wanyang, et al.
Published: (2025)
Shadows in the Attention: Contextual Perturbation and Representation Drift in the Dynamics of Hallucination in LLMs
by: Wei, Zeyu, et al.
Published: (2025)
by: Wei, Zeyu, et al.
Published: (2025)
DAug: Diffusion-based Channel Augmentation for Radiology Image Retrieval and Classification
by: Jin, Ying, et al.
Published: (2024)
by: Jin, Ying, et al.
Published: (2024)
ProTPS: Prototype-Guided Text Prompt Selection for Continual Learning
by: Mei, Jie, et al.
Published: (2026)
by: Mei, Jie, et al.
Published: (2026)
Understanding Identity Continuity in Thermal Video through Scene-Level Consistency
by: Sun, Wei-Chieh, et al.
Published: (2026)
by: Sun, Wei-Chieh, et al.
Published: (2026)
SuperCLUE-Fin: Graded Fine-Grained Analysis of Chinese LLMs on Diverse Financial Tasks and Applications
by: Xu, Liang, et al.
Published: (2024)
by: Xu, Liang, et al.
Published: (2024)
Block-Attention for Efficient Prefilling
by: Ma, Dongyang, et al.
Published: (2024)
by: Ma, Dongyang, et al.
Published: (2024)
BottleHumor: Self-Informed Humor Explanation using the Information Bottleneck Principle
by: Hwang, EunJeong, et al.
Published: (2025)
by: Hwang, EunJeong, et al.
Published: (2025)
Does Instruction Tuning Make LLMs More Consistent?
by: Fierro, Constanza, et al.
Published: (2024)
by: Fierro, Constanza, et al.
Published: (2024)
Local Explanations and Self-Explanations for Assessing Faithfulness in black-box LLMs
by: Fragkathoulas, Christos, et al.
Published: (2024)
by: Fragkathoulas, Christos, et al.
Published: (2024)
Foot-In-The-Door: A Multi-turn Jailbreak for LLMs
by: Weng, Zixuan, et al.
Published: (2025)
by: Weng, Zixuan, et al.
Published: (2025)
Last Layer Logits to Logic: Empowering LLMs with Logic-Consistent Structured Knowledge Reasoning
by: Li, Songze, et al.
Published: (2025)
by: Li, Songze, et al.
Published: (2025)
MovieChat+: Question-aware Sparse Memory for Long Video Question Answering
by: Song, Enxin, et al.
Published: (2024)
by: Song, Enxin, et al.
Published: (2024)
Too Consistent to Detect: A Study of Self-Consistent Errors in LLMs
by: Tan, Hexiang, et al.
Published: (2025)
by: Tan, Hexiang, et al.
Published: (2025)
Pointmap Association and Piecewise-Plane Constraint for Consistent and Compact 3D Gaussian Segmentation Field
by: Hu, Wenhao, et al.
Published: (2025)
by: Hu, Wenhao, et al.
Published: (2025)
Harnessing LLMs Explanations to Boost Surrogate Models in Tabular Data Classification
by: Shi, Ruxue, et al.
Published: (2025)
by: Shi, Ruxue, et al.
Published: (2025)
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache
by: Liu, Xiaoran, et al.
Published: (2025)
by: Liu, Xiaoran, et al.
Published: (2025)
Consistent Autoformalization for Constructing Mathematical Libraries
by: Zhang, Lan, et al.
Published: (2024)
by: Zhang, Lan, et al.
Published: (2024)
Attn-GS: Attention-Guided Context Compression for Efficient Personalized LLMs
by: Zeng, Shenglai, et al.
Published: (2026)
by: Zeng, Shenglai, et al.
Published: (2026)
Regularization, Semi-supervision, and Supervision for a Plausible Attention-Based Explanation
by: Nguyen, Duc Hau, et al.
Published: (2025)
by: Nguyen, Duc Hau, et al.
Published: (2025)
Generating Medically-Informed Explanations for Depression Detection using LLMs
by: Chen, Xiangyong, et al.
Published: (2025)
by: Chen, Xiangyong, et al.
Published: (2025)
ParetoRAG: Leveraging Sentence-Context Attention for Robust and Efficient Retrieval-Augmented Generation
by: Yao, Ruobing, et al.
Published: (2025)
by: Yao, Ruobing, et al.
Published: (2025)
Detecting Contextual Hallucinations in LLMs with Frequency-Aware Attention
by: Qi, Siya, et al.
Published: (2026)
by: Qi, Siya, et al.
Published: (2026)
Consistency Matters: Explore LLMs Consistency From a Black-Box Perspective
by: Zhao, Fufangchen, et al.
Published: (2024)
by: Zhao, Fufangchen, et al.
Published: (2024)
Knocking-Heads Attention
by: Zhou, Zhanchao, et al.
Published: (2025)
by: Zhou, Zhanchao, et al.
Published: (2025)
LLMs as Bridges: Reformulating Grounded Multimodal Named Entity Recognition
by: Li, Jinyuan, et al.
Published: (2024)
by: Li, Jinyuan, et al.
Published: (2024)
Similar Items
-
Bayesian Optimization for Controlled Image Editing via LLMs
by: Cai, Chengkun, et al.
Published: (2025) -
Modeling LLM Agent Reviewer Dynamics in Elo-Ranked Review System
by: Huang, Hsiang-Wei, et al.
Published: (2026) -
Argument-Based Consistency in Toxicity Explanations of LLMs
by: Mothilal, Ramaravind Kommiya, et al.
Published: (2025) -
The Role of Deductive and Inductive Reasoning in Large Language Models
by: Cai, Chengkun, et al.
Published: (2024) -
CNSocialDepress: A Chinese Social Media Dataset for Depression Risk Detection and Structured Analysis
by: Xu, Jinyuan, et al.
Published: (2025)