AggTruth: Contextual Hallucination Detection using Aggregated Attention Scores in LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Matys, Piotr, Eliasz, Jan, Kiełczyński, Konrad, Langner, Mikołaj, Ferdinan, Teddy, Kocoń, Jan, Kazienko, Przemysław |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Into the Unknown: Self-Learning Large Language Models
by: Ferdinan, Teddy, et al.
Published: (2024)
by: Ferdinan, Teddy, et al.
Published: (2024)
What properties of reasoning supervision are associated with improved downstream model quality?
by: Langner, Mikołaj, et al.
Published: (2026)
by: Langner, Mikołaj, et al.
Published: (2026)
Divide, Cache, Conquer: Dichotomic Prompting for Efficient Multi-Label LLM-Based Classification
by: Langner, Mikołaj, et al.
Published: (2025)
by: Langner, Mikołaj, et al.
Published: (2025)
Fortifying NLP models - dataset + code
by: Ferdinan, Teddy, et al.
Published: (2025)
by: Ferdinan, Teddy, et al.
Published: (2025)
Self-training Large Language Models through Knowledge Detection
by: Yeo, Wei Jie, et al.
Published: (2024)
by: Yeo, Wei Jie, et al.
Published: (2024)
Personalized Large Language Models
by: Woźniak, Stanisław, et al.
Published: (2024)
by: Woźniak, Stanisław, et al.
Published: (2024)
Language, Culture, and Ideology: Personalizing Offensiveness Detection in Political Tweets with Reasoning LLMs
by: Pihulski, Dzmitry, et al.
Published: (2025)
by: Pihulski, Dzmitry, et al.
Published: (2025)
Detecting Contextual Hallucinations in LLMs with Frequency-Aware Attention
by: Qi, Siya, et al.
Published: (2026)
by: Qi, Siya, et al.
Published: (2026)
Unraveling SITT: Social Influence Technique Taxonomy and Detection with LLMs
by: Mieleszczenko-Kowszewicz, Wiktoria, et al.
Published: (2025)
by: Mieleszczenko-Kowszewicz, Wiktoria, et al.
Published: (2025)
Graphing the Truth: Structured Visualizations for Automated Hallucination Detection in LLMs
by: Agrawal, Tanmay
Published: (2025)
by: Agrawal, Tanmay
Published: (2025)
Shadows in the Attention: Contextual Perturbation and Representation Drift in the Dynamics of Hallucination in LLMs
by: Wei, Zeyu, et al.
Published: (2025)
by: Wei, Zeyu, et al.
Published: (2025)
Backtranslation and paraphrasing in the LLM era? Comparing data augmentation methods for emotion classification
by: Radliński, Łukasz, et al.
Published: (2025)
by: Radliński, Łukasz, et al.
Published: (2025)
Probabilistic Guarantees for Reducing Contextual Hallucinations in LLMs
by: Rautenberg, Nils, et al.
Published: (2026)
by: Rautenberg, Nils, et al.
Published: (2026)
Hallucination Detection in LLMs with Topological Divergence on Attention Graphs
by: Bazarova, Alexandra, et al.
Published: (2025)
by: Bazarova, Alexandra, et al.
Published: (2025)
When Bias Pretends to Be Truth: How Spurious Correlations Undermine Hallucination Detection in LLMs
by: Wang, Shaowen, et al.
Published: (2025)
by: Wang, Shaowen, et al.
Published: (2025)
LLMSQL: Upgrading WikiSQL for the LLM Era of Text-to-SQL
by: Pihulski, Dzmitry, et al.
Published: (2025)
by: Pihulski, Dzmitry, et al.
Published: (2025)
BeamAggR: Beam Aggregation Reasoning over Multi-source Knowledge for Multi-hop Question Answering
by: Chu, Zheng, et al.
Published: (2024)
by: Chu, Zheng, et al.
Published: (2024)
Hallucination Detection in LLMs Using Spectral Features of Attention Maps
by: Binkowski, Jakub, et al.
Published: (2025)
by: Binkowski, Jakub, et al.
Published: (2025)
Detecting Hallucinations in SpeechLLMs at Inference Time Using Attention Maps
by: Waldendorf, Jonas, et al.
Published: (2026)
by: Waldendorf, Jonas, et al.
Published: (2026)
Lookback Lens: Detecting and Mitigating Contextual Hallucinations in Large Language Models Using Only Attention Maps
by: Chuang, Yung-Sung, et al.
Published: (2024)
by: Chuang, Yung-Sung, et al.
Published: (2024)
Can LLMs Detect Their Own Hallucinations?
by: Kadotani, Sora, et al.
Published: (2025)
by: Kadotani, Sora, et al.
Published: (2025)
Gradient-guided Attention Map Editing: Towards Efficient Contextual Hallucination Mitigation
by: Wang, Yu, et al.
Published: (2025)
by: Wang, Yu, et al.
Published: (2025)
Truth is Universal: Robust Detection of Lies in LLMs
by: Bürger, Lennart, et al.
Published: (2024)
by: Bürger, Lennart, et al.
Published: (2024)
SSN and the Hodge Conjecture – Project 1: Millennium Problems
by: Nowak, Eliasz Przemysław
Published: (2025)
by: Nowak, Eliasz Przemysław
Published: (2025)
Zero-Shot Stance Detection using Contextual Data Generation with LLMs
by: Mahmoudi, Ghazaleh, et al.
Published: (2024)
by: Mahmoudi, Ghazaleh, et al.
Published: (2024)
Eagle and Finch: RWKV with Matrix-Valued States and Dynamic Recurrence
by: Peng, Bo, et al.
Published: (2024)
by: Peng, Bo, et al.
Published: (2024)
Predicting stock prices with ChatGPT-annotated Reddit sentiment
by: Kmak, Mateusz, et al.
Published: (2025)
by: Kmak, Mateusz, et al.
Published: (2025)
Training-free Truthfulness Detection via Value Vectors in LLMs
by: Liu, Runheng, et al.
Published: (2025)
by: Liu, Runheng, et al.
Published: (2025)
Geometric Uncertainty for Detecting and Correcting Hallucinations in LLMs
by: Phillips, Edward, et al.
Published: (2025)
by: Phillips, Edward, et al.
Published: (2025)
Hallucination Detection with the Internal Layers of LLMs
by: Preiß, Martin
Published: (2025)
by: Preiß, Martin
Published: (2025)
TruthX: Alleviating Hallucinations by Editing Large Language Models in Truthful Space
by: Zhang, Shaolei, et al.
Published: (2024)
by: Zhang, Shaolei, et al.
Published: (2024)
Beyond Logit Lens: Contextual Embeddings for Robust Hallucination Detection & Grounding in VLMs
by: Phukan, Anirudh, et al.
Published: (2024)
by: Phukan, Anirudh, et al.
Published: (2024)
Two Pathways to Truthfulness: On the Intrinsic Encoding of LLM Hallucinations
by: Luo, Wen, et al.
Published: (2026)
by: Luo, Wen, et al.
Published: (2026)
Enhanced Hallucination Detection in Neural Machine Translation through Simple Detector Aggregation
by: Himmi, Anas, et al.
Published: (2024)
by: Himmi, Anas, et al.
Published: (2024)
Cost-Effective Hallucination Detection for LLMs
by: Valentin, Simon, et al.
Published: (2024)
by: Valentin, Simon, et al.
Published: (2024)
Pruning Weights but Not Truth: Safeguarding Truthfulness While Pruning LLMs
by: Fu, Yao, et al.
Published: (2025)
by: Fu, Yao, et al.
Published: (2025)
On the Universal Truthfulness Hyperplane Inside LLMs
by: Liu, Junteng, et al.
Published: (2024)
by: Liu, Junteng, et al.
Published: (2024)
Ground Truth Generation for Multilingual Historical NLP using LLMs
by: Gladstone, Clovis, et al.
Published: (2025)
by: Gladstone, Clovis, et al.
Published: (2025)
The Two Sides of the Coin: Hallucination Generation and Detection with LLMs as Evaluators for LLMs
by: Bui, Anh Thu Maria, et al.
Published: (2024)
by: Bui, Anh Thu Maria, et al.
Published: (2024)
Fine-Grained Detection of Context-Grounded Hallucinations Using LLMs
by: Peisakhovsky, Yehonatan, et al.
Published: (2025)
by: Peisakhovsky, Yehonatan, et al.
Published: (2025)
Similar Items
-
Into the Unknown: Self-Learning Large Language Models
by: Ferdinan, Teddy, et al.
Published: (2024) -
What properties of reasoning supervision are associated with improved downstream model quality?
by: Langner, Mikołaj, et al.
Published: (2026) -
Divide, Cache, Conquer: Dichotomic Prompting for Efficient Multi-Label LLM-Based Classification
by: Langner, Mikołaj, et al.
Published: (2025) -
Fortifying NLP models - dataset + code
by: Ferdinan, Teddy, et al.
Published: (2025) -
Self-training Large Language Models through Knowledge Detection
by: Yeo, Wei Jie, et al.
Published: (2024)