Cross-Layer Attention Probing for Fine-Grained Hallucination Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Suresh, Malavika, Aljundi, Rahaf, Nkisi-Orji, Ikechukwu, Wiratunga, Nirmalie |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RAPT: Retrieval-Augmented Post-hoc Thresholding for Multi-Label Classification
by: Jayawardena, Lasal, et al.
Published: (2026)
by: Jayawardena, Lasal, et al.
Published: (2026)
CBR-RAG: Case-Based Reasoning for Retrieval Augmented Generation in LLMs for Legal Question Answering
by: Wiratunga, Nirmalie, et al.
Published: (2024)
by: Wiratunga, Nirmalie, et al.
Published: (2024)
XEQ Scale for Evaluating XAI Experience Quality
by: Wijekoon, Anjana, et al.
Published: (2024)
by: Wijekoon, Anjana, et al.
Published: (2024)
iSee: Advancing Multi-Shot Explainable AI Using Case-based Recommendations
by: Wijekoon, Anjana, et al.
Published: (2024)
by: Wijekoon, Anjana, et al.
Published: (2024)
FedFT: Improving Communication Performance for Federated Learning with Frequency Space Transformation
by: Palihawadana, Chamath, et al.
Published: (2024)
by: Palihawadana, Chamath, et al.
Published: (2024)
This changes to that : Combining causal and non-causal explanations to generate disease progression in capsule endoscopy
by: Vats, Anuja, et al.
Published: (2022)
by: Vats, Anuja, et al.
Published: (2022)
ReXTrust: A Model for Fine-Grained Hallucination Detection in AI-Generated Radiology Reports
by: Hardy, Romain, et al.
Published: (2024)
by: Hardy, Romain, et al.
Published: (2024)
Imperfect Vision Encoders: Efficient and Robust Tuning for Vision-Language Models
by: Panos, Aristeidis, et al.
Published: (2024)
by: Panos, Aristeidis, et al.
Published: (2024)
Hallucination Detection with the Internal Layers of LLMs
by: Preiß, Martin
Published: (2025)
by: Preiß, Martin
Published: (2025)
Tell me more: Intent Fulfilment Framework for Enhancing User Experiences in Conversational XAI
by: Wijekoon, Anjana, et al.
Published: (2024)
by: Wijekoon, Anjana, et al.
Published: (2024)
C-FAITH: A Chinese Fine-Grained Benchmark for Automated Hallucination Evaluation
by: Zhang, Xu, et al.
Published: (2025)
by: Zhang, Xu, et al.
Published: (2025)
HausaNLP at SemEval-2025 Task 3: Towards a Fine-Grained Model-Aware Hallucination Detection
by: Bala, Maryam, et al.
Published: (2025)
by: Bala, Maryam, et al.
Published: (2025)
Hallucination Detection in LLMs with Topological Divergence on Attention Graphs
by: Bazarova, Alexandra, et al.
Published: (2025)
by: Bazarova, Alexandra, et al.
Published: (2025)
Neural Probe-Based Hallucination Detection for Large Language Models
by: Liang, Shize, et al.
Published: (2025)
by: Liang, Shize, et al.
Published: (2025)
Counterfactual Probing for Hallucination Detection and Mitigation in Large Language Models
by: Feng, Yijun
Published: (2025)
by: Feng, Yijun
Published: (2025)
Attention with Dependency Parsing Augmentation for Fine-Grained Attribution
by: Ding, Qiang, et al.
Published: (2024)
by: Ding, Qiang, et al.
Published: (2024)
ICR Probe: Tracking Hidden State Dynamics for Reliable Hallucination Detection in LLMs
by: Zhang, Zhenliang, et al.
Published: (2025)
by: Zhang, Zhenliang, et al.
Published: (2025)
AggTruth: Contextual Hallucination Detection using Aggregated Attention Scores in LLMs
by: Matys, Piotr, et al.
Published: (2025)
by: Matys, Piotr, et al.
Published: (2025)
Hallucination Detection and Hallucination Mitigation: An Investigation
by: Luo, Junliang, et al.
Published: (2024)
by: Luo, Junliang, et al.
Published: (2024)
Hallucinated Span Detection with Multi-View Attention Features
by: Ogasa, Yuya, et al.
Published: (2025)
by: Ogasa, Yuya, et al.
Published: (2025)
Semantic Entropy Probes: Robust and Cheap Hallucination Detection in LLMs
by: Kossen, Jannik, et al.
Published: (2024)
by: Kossen, Jannik, et al.
Published: (2024)
Attention-guided Self-reflection for Zero-shot Hallucination Detection in Large Language Models
by: Liu, Qiang, et al.
Published: (2025)
by: Liu, Qiang, et al.
Published: (2025)
ZINA: Multimodal Fine-grained Hallucination Detection and Editing
by: Wada, Yuiga, et al.
Published: (2025)
by: Wada, Yuiga, et al.
Published: (2025)
PFME: A Modular Approach for Fine-grained Hallucination Detection and Editing of Large Language Models
by: Deng, Kunquan, et al.
Published: (2024)
by: Deng, Kunquan, et al.
Published: (2024)
Small Updates, Big Doubts: Does Parameter-Efficient Fine-tuning Enhance Hallucination Detection ?
by: Hu, Xu, et al.
Published: (2026)
by: Hu, Xu, et al.
Published: (2026)
Detecting and Mitigating Hallucination in Large Vision Language Models via Fine-Grained AI Feedback
by: Xiao, Wenyi, et al.
Published: (2024)
by: Xiao, Wenyi, et al.
Published: (2024)
Middle-Layer Representation Alignment for Cross-Lingual Transfer in Fine-Tuned LLMs
by: Liu, Danni, et al.
Published: (2025)
by: Liu, Danni, et al.
Published: (2025)
Hallucination Detection in LLMs Using Spectral Features of Attention Maps
by: Binkowski, Jakub, et al.
Published: (2025)
by: Binkowski, Jakub, et al.
Published: (2025)
CCHall: A Novel Benchmark for Joint Cross-Lingual and Cross-Modal Hallucinations Detection in Large Language Models
by: Zhang, Yongheng, et al.
Published: (2025)
by: Zhang, Yongheng, et al.
Published: (2025)
PRISM: Probing Reasoning, Instruction, and Source Memory in LLM Hallucinations
by: Wu, Yuhe, et al.
Published: (2026)
by: Wu, Yuhe, et al.
Published: (2026)
LRP4RAG: Detecting Hallucinations in Retrieval-Augmented Generation via Layer-wise Relevance Propagation
by: Hu, Haichuan, et al.
Published: (2024)
by: Hu, Haichuan, et al.
Published: (2024)
Hallucination Detection in LLMs: Fast and Memory-Efficient Fine-Tuned Models
by: Arteaga, Gabriel Y., et al.
Published: (2024)
by: Arteaga, Gabriel Y., et al.
Published: (2024)
HICD: Hallucination-Inducing via Attention Dispersion for Contrastive Decoding to Mitigate Hallucinations in Large Language Models
by: Jiang, Xinyan, et al.
Published: (2025)
by: Jiang, Xinyan, et al.
Published: (2025)
Detecting Hallucinations in SpeechLLMs at Inference Time Using Attention Maps
by: Waldendorf, Jonas, et al.
Published: (2026)
by: Waldendorf, Jonas, et al.
Published: (2026)
Fine-Grained Detection of AI-Generated Text Using Sentence-Level Segmentation
by: Teja, Lekkala Sai, et al.
Published: (2025)
by: Teja, Lekkala Sai, et al.
Published: (2025)
Fine-Grained Bias Detection in LLM: Enhancing detection mechanisms for nuanced biases
by: Mohanty, Suvendu
Published: (2025)
by: Mohanty, Suvendu
Published: (2025)
CFMS: Towards Explainable and Fine-Grained Chinese Multimodal Sarcasm Detection Benchmark
by: Zhang, Junzhao, et al.
Published: (2026)
by: Zhang, Junzhao, et al.
Published: (2026)
Beyond Binary Classification: Detecting Fine-Grained Sexism in Social Media Videos
by: De Grazia, Laura, et al.
Published: (2026)
by: De Grazia, Laura, et al.
Published: (2026)
On Fine-Grained I/O Complexity of Attention Backward Passes
by: Li, Xiaoyu, et al.
Published: (2024)
by: Li, Xiaoyu, et al.
Published: (2024)
Opioid Named Entity Recognition (ONER-2025) from Reddit
by: Ahmad, Muhammad, et al.
Published: (2025)
by: Ahmad, Muhammad, et al.
Published: (2025)
Similar Items
-
RAPT: Retrieval-Augmented Post-hoc Thresholding for Multi-Label Classification
by: Jayawardena, Lasal, et al.
Published: (2026) -
CBR-RAG: Case-Based Reasoning for Retrieval Augmented Generation in LLMs for Legal Question Answering
by: Wiratunga, Nirmalie, et al.
Published: (2024) -
XEQ Scale for Evaluating XAI Experience Quality
by: Wijekoon, Anjana, et al.
Published: (2024) -
iSee: Advancing Multi-Shot Explainable AI Using Case-based Recommendations
by: Wijekoon, Anjana, et al.
Published: (2024) -
FedFT: Improving Communication Performance for Federated Learning with Frequency Space Transformation
by: Palihawadana, Chamath, et al.
Published: (2024)