Hallucination Detection in LLMs Using Spectral Features of Attention Maps
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Binkowski, Jakub, Janiak, Denis, Sawczyn, Albert, Gabrys, Bogdan, Kajdanowicz, Tomasz |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FactSelfCheck: Fact-Level Black-Box Hallucination Detection for LLMs
von: Sawczyn, Albert, et al.
Veröffentlicht: (2025)
von: Sawczyn, Albert, et al.
Veröffentlicht: (2025)
The Illusion of Progress: Re-evaluating Hallucination Detection in LLMs
von: Janiak, Denis, et al.
Veröffentlicht: (2025)
von: Janiak, Denis, et al.
Veröffentlicht: (2025)
Empowering Small-Scale Knowledge Graphs: A Strategy of Leveraging General-Purpose Knowledge Graphs for Enriched Embeddings
von: Sawczyn, Albert, et al.
Veröffentlicht: (2024)
von: Sawczyn, Albert, et al.
Veröffentlicht: (2024)
Attention Sinks as Internal Signals for Hallucination Detection in Large Language Models
von: Binkowski, Jakub, et al.
Veröffentlicht: (2026)
von: Binkowski, Jakub, et al.
Veröffentlicht: (2026)
A Geometry-Based View of Mahalanobis OOD Detection
von: Janiak, Denis, et al.
Veröffentlicht: (2025)
von: Janiak, Denis, et al.
Veröffentlicht: (2025)
Developing PUGG for Polish: A Modern Approach to KBQA, MRC, and IR Dataset Construction
von: Sawczyn, Albert, et al.
Veröffentlicht: (2024)
von: Sawczyn, Albert, et al.
Veröffentlicht: (2024)
KG-Guard: Graph-Based Hallucination Detection for Knowledge Base Question Answering
von: Sawczyn, Albert, et al.
Veröffentlicht: (2026)
von: Sawczyn, Albert, et al.
Veröffentlicht: (2026)
Detecting Hallucinations in SpeechLLMs at Inference Time Using Attention Maps
von: Waldendorf, Jonas, et al.
Veröffentlicht: (2026)
von: Waldendorf, Jonas, et al.
Veröffentlicht: (2026)
Hallucinated Span Detection with Multi-View Attention Features
von: Ogasa, Yuya, et al.
Veröffentlicht: (2025)
von: Ogasa, Yuya, et al.
Veröffentlicht: (2025)
Lookback Lens: Detecting and Mitigating Contextual Hallucinations in Large Language Models Using Only Attention Maps
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2024)
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2024)
Cost-Effective Hallucination Detection for LLMs
von: Valentin, Simon, et al.
Veröffentlicht: (2024)
von: Valentin, Simon, et al.
Veröffentlicht: (2024)
The Map of Misbelief: Tracing Intrinsic and Extrinsic Hallucinations Through Attention Patterns
von: Hajji, Elyes, et al.
Veröffentlicht: (2025)
von: Hajji, Elyes, et al.
Veröffentlicht: (2025)
HD-NDEs: Neural Differential Equations for Hallucination Detection in LLMs
von: Li, Qing, et al.
Veröffentlicht: (2025)
von: Li, Qing, et al.
Veröffentlicht: (2025)
Semantic Entropy Probes: Robust and Cheap Hallucination Detection in LLMs
von: Kossen, Jannik, et al.
Veröffentlicht: (2024)
von: Kossen, Jannik, et al.
Veröffentlicht: (2024)
Hallucination Detection in LLMs: Fast and Memory-Efficient Fine-Tuned Models
von: Arteaga, Gabriel Y., et al.
Veröffentlicht: (2024)
von: Arteaga, Gabriel Y., et al.
Veröffentlicht: (2024)
A Depression Detection Method Based on Multi-Modal Feature Fusion Using Cross-Attention
von: Li, Shengjie, et al.
Veröffentlicht: (2024)
von: Li, Shengjie, et al.
Veröffentlicht: (2024)
When Bias Pretends to Be Truth: How Spurious Correlations Undermine Hallucination Detection in LLMs
von: Wang, Shaowen, et al.
Veröffentlicht: (2025)
von: Wang, Shaowen, et al.
Veröffentlicht: (2025)
Learning to Reason for Hallucination Span Detection
von: Su, Hsuan, et al.
Veröffentlicht: (2025)
von: Su, Hsuan, et al.
Veröffentlicht: (2025)
Steer LLM Latents for Hallucination Detection
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
Anomaly Detection of Tabular Data Using LLMs
von: Li, Aodong, et al.
Veröffentlicht: (2024)
von: Li, Aodong, et al.
Veröffentlicht: (2024)
QueryBandits for Hallucination Mitigation: Exploiting Semantic Features for No-Regret Rewriting
von: Cho, Nicole, et al.
Veröffentlicht: (2025)
von: Cho, Nicole, et al.
Veröffentlicht: (2025)
Evaluating Evaluation Metrics -- The Mirage of Hallucination Detection
von: Kulkarni, Atharva, et al.
Veröffentlicht: (2025)
von: Kulkarni, Atharva, et al.
Veröffentlicht: (2025)
KARL: Mitigating Hallucinations in LLMs via Knowledge-Boundary-Aware Reinforcement Learning
von: Gao, Cheng, et al.
Veröffentlicht: (2026)
von: Gao, Cheng, et al.
Veröffentlicht: (2026)
Distilling LLMs' Decomposition Abilities into Compact Language Models
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
(Im)possibility of Automated Hallucination Detection in Large Language Models
von: Karbasi, Amin, et al.
Veröffentlicht: (2025)
von: Karbasi, Amin, et al.
Veröffentlicht: (2025)
Bolster Hallucination Detection via Prompt-Guided Data Augmentation
von: Li, Wenyun, et al.
Veröffentlicht: (2025)
von: Li, Wenyun, et al.
Veröffentlicht: (2025)
Real-Time Detection of Hallucinated Entities in Long-Form Generation
von: Obeso, Oscar, et al.
Veröffentlicht: (2025)
von: Obeso, Oscar, et al.
Veröffentlicht: (2025)
Hallucination Detection via Activations of Open-Weight Proxy Analyzers
von: Singh, Akshita, et al.
Veröffentlicht: (2026)
von: Singh, Akshita, et al.
Veröffentlicht: (2026)
Scalable Token-Level Hallucination Detection in Large Language Models
von: Min, Rui, et al.
Veröffentlicht: (2026)
von: Min, Rui, et al.
Veröffentlicht: (2026)
Training Language Models on the Knowledge Graph: Insights on Hallucinations and Their Detectability
von: Hron, Jiri, et al.
Veröffentlicht: (2024)
von: Hron, Jiri, et al.
Veröffentlicht: (2024)
Chain-of-Thought Reasoning In The Wild Is Not Always Faithful
von: Arcuschin, Iván, et al.
Veröffentlicht: (2025)
von: Arcuschin, Iván, et al.
Veröffentlicht: (2025)
mcdok at SemEval-2026 Task 13: Finetuning LLMs for Detection of Machine-Generated Code
von: Skurla, Adam, et al.
Veröffentlicht: (2026)
von: Skurla, Adam, et al.
Veröffentlicht: (2026)
Scaling Attention to Very Long Sequences in Linear Time with Wavelet-Enhanced Random Spectral Attention (WERSA)
von: Dentamaro, Vincenzo
Veröffentlicht: (2025)
von: Dentamaro, Vincenzo
Veröffentlicht: (2025)
FRED: Financial Retrieval-Enhanced Detection and Editing of Hallucinations in Language Models
von: Tan, Likun, et al.
Veröffentlicht: (2025)
von: Tan, Likun, et al.
Veröffentlicht: (2025)
Principled Detection of Hallucinations in Large Language Models via Multiple Testing
von: Li, Jiawei, et al.
Veröffentlicht: (2025)
von: Li, Jiawei, et al.
Veröffentlicht: (2025)
Self-contradictory Hallucinations of Large Language Models: Evaluation, Detection and Mitigation
von: Mündler, Niels, et al.
Veröffentlicht: (2023)
von: Mündler, Niels, et al.
Veröffentlicht: (2023)
SLM Meets LLM: Balancing Latency, Interpretability and Consistency in Hallucination Detection
von: Hu, Mengya, et al.
Veröffentlicht: (2024)
von: Hu, Mengya, et al.
Veröffentlicht: (2024)
Manifold-based Sampling for In-Context Hallucination Detection in Large Language Models
von: Vamshi, Bodla Krishna, et al.
Veröffentlicht: (2026)
von: Vamshi, Bodla Krishna, et al.
Veröffentlicht: (2026)
StylOch at PAN: Gradient-Boosted Trees with Frequency-Based Stylometric Features
von: Ochab, Jeremi K., et al.
Veröffentlicht: (2025)
von: Ochab, Jeremi K., et al.
Veröffentlicht: (2025)
Focus On This, Not That! Steering LLMs with Adaptive Feature Specification
von: Lamb, Tom A., et al.
Veröffentlicht: (2024)
von: Lamb, Tom A., et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
FactSelfCheck: Fact-Level Black-Box Hallucination Detection for LLMs
von: Sawczyn, Albert, et al.
Veröffentlicht: (2025) -
The Illusion of Progress: Re-evaluating Hallucination Detection in LLMs
von: Janiak, Denis, et al.
Veröffentlicht: (2025) -
Empowering Small-Scale Knowledge Graphs: A Strategy of Leveraging General-Purpose Knowledge Graphs for Enriched Embeddings
von: Sawczyn, Albert, et al.
Veröffentlicht: (2024) -
Attention Sinks as Internal Signals for Hallucination Detection in Large Language Models
von: Binkowski, Jakub, et al.
Veröffentlicht: (2026) -
A Geometry-Based View of Mahalanobis OOD Detection
von: Janiak, Denis, et al.
Veröffentlicht: (2025)