TraceDet: Hallucination Detection from the Decoding Trace of Diffusion Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chang, Shenxu, Yu, Junchi, Wang, Weixing, Chen, Yongqiang, Yu, Jialin, Torr, Philip, Gu, Jindong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TDGNet: Hallucination Detection in Diffusion Language Models via Temporal Dynamic Graphs
von: Hemmat, Arshia, et al.
Veröffentlicht: (2026)
von: Hemmat, Arshia, et al.
Veröffentlicht: (2026)
Safer by Diffusion, Broken by Context: Diffusion LLM's Safety Blessing and Its Failure Mode
von: He, Zeyuan, et al.
Veröffentlicht: (2026)
von: He, Zeyuan, et al.
Veröffentlicht: (2026)
The Path Matters: Learning a Token-Commitment Policy for Diffusion Language Models
von: Sun, Bohang, et al.
Veröffentlicht: (2026)
von: Sun, Bohang, et al.
Veröffentlicht: (2026)
Transcoders Trace Visual Grounding and Hallucinations in Vision-Language Models
von: Damianos, Dimitrios, et al.
Veröffentlicht: (2026)
von: Damianos, Dimitrios, et al.
Veröffentlicht: (2026)
MemTrace: Tracing and Attributing Errors in Large Language Model Memory Systems
von: Deng, Xinle, et al.
Veröffentlicht: (2026)
von: Deng, Xinle, et al.
Veröffentlicht: (2026)
Causal Fine-Tuning under Latent Confounded Shift
von: Yu, Jialin, et al.
Veröffentlicht: (2024)
von: Yu, Jialin, et al.
Veröffentlicht: (2024)
Active Layer-Contrastive Decoding Reduces Hallucination in Large Language Model Generation
von: Zhang, Hongxiang, et al.
Veröffentlicht: (2025)
von: Zhang, Hongxiang, et al.
Veröffentlicht: (2025)
SEASON: Mitigating Temporal Hallucination in Video Large Language Models via Self-Diagnostic Contrastive Decoding
von: Wu, Chang-Hsun, et al.
Veröffentlicht: (2025)
von: Wu, Chang-Hsun, et al.
Veröffentlicht: (2025)
Towards Interpretable Sequence Continuation: Analyzing Shared Circuits in Large Language Models
von: Lan, Michael, et al.
Veröffentlicht: (2023)
von: Lan, Michael, et al.
Veröffentlicht: (2023)
Red Teaming GPT-4V: Are GPT-4V Safe Against Uni/Multi-Modal Jailbreak Attacks?
von: Chen, Shuo, et al.
Veröffentlicht: (2024)
von: Chen, Shuo, et al.
Veröffentlicht: (2024)
Unified Hallucination Detection for Multimodal Large Language Models
von: Chen, Xiang, et al.
Veröffentlicht: (2024)
von: Chen, Xiang, et al.
Veröffentlicht: (2024)
Revisiting Uncertainty Estimation and Calibration of Large Language Models
von: Tao, Linwei, et al.
Veröffentlicht: (2025)
von: Tao, Linwei, et al.
Veröffentlicht: (2025)
Distinguishable Deletion: Unifying Knowledge Erasure and Refusal for Large Language Model Unlearning
von: Yang, Puning, et al.
Veröffentlicht: (2026)
von: Yang, Puning, et al.
Veröffentlicht: (2026)
Efficient Contrastive Decoding with Probabilistic Hallucination Detection - Mitigating Hallucinations in Large Vision Language Models -
von: Fieback, Laura, et al.
Veröffentlicht: (2025)
von: Fieback, Laura, et al.
Veröffentlicht: (2025)
Confidence Geometry Reveals Trace-Level Correctness in Large Language Model Reasoning
von: Liu, Shuo, et al.
Veröffentlicht: (2026)
von: Liu, Shuo, et al.
Veröffentlicht: (2026)
Leveraging Graph Structures to Detect Hallucinations in Large Language Models
von: Nonkes, Noa, et al.
Veröffentlicht: (2024)
von: Nonkes, Noa, et al.
Veröffentlicht: (2024)
On Large Language Models' Hallucination with Regard to Known Facts
von: Jiang, Che, et al.
Veröffentlicht: (2024)
von: Jiang, Che, et al.
Veröffentlicht: (2024)
Dynamic Attention-Guided Context Decoding for Mitigating Context Faithfulness Hallucinations in Large Language Models
von: Huang, Yanwen, et al.
Veröffentlicht: (2025)
von: Huang, Yanwen, et al.
Veröffentlicht: (2025)
ERBench: An Entity-Relationship based Automatically Verifiable Hallucination Benchmark for Large Language Models
von: Oh, Jio, et al.
Veröffentlicht: (2024)
von: Oh, Jio, et al.
Veröffentlicht: (2024)
Language over Content: Tracing Cultural Understanding in Multilingual Large Language Models
von: Cho, Seungho, et al.
Veröffentlicht: (2025)
von: Cho, Seungho, et al.
Veröffentlicht: (2025)
Attention Sinks as Internal Signals for Hallucination Detection in Large Language Models
von: Binkowski, Jakub, et al.
Veröffentlicht: (2026)
von: Binkowski, Jakub, et al.
Veröffentlicht: (2026)
Tracing Uncertainty in Language Model "Reasoning"
von: Grünefeld, Nils, et al.
Veröffentlicht: (2026)
von: Grünefeld, Nils, et al.
Veröffentlicht: (2026)
The Map of Misbelief: Tracing Intrinsic and Extrinsic Hallucinations Through Attention Patterns
von: Hajji, Elyes, et al.
Veröffentlicht: (2025)
von: Hajji, Elyes, et al.
Veröffentlicht: (2025)
FlashDecoding++: Faster Large Language Model Inference on GPUs
von: Hong, Ke, et al.
Veröffentlicht: (2023)
von: Hong, Ke, et al.
Veröffentlicht: (2023)
On the Thinking-Language Modeling Gap in Large Language Models
von: Liu, Chenxi, et al.
Veröffentlicht: (2025)
von: Liu, Chenxi, et al.
Veröffentlicht: (2025)
Tracing Facts or just Copies? A critical investigation of the Competitions of Mechanisms in Large Language Models
von: Campregher, Dante, et al.
Veröffentlicht: (2025)
von: Campregher, Dante, et al.
Veröffentlicht: (2025)
From Noisy Traces to Stable Gradients: Bias-Variance Optimized Preference Optimization for Aligning Large Reasoning Models
von: Zhu, Mingkang, et al.
Veröffentlicht: (2025)
von: Zhu, Mingkang, et al.
Veröffentlicht: (2025)
Do as I do (Safely): Mitigating Task-Specific Fine-tuning Risks in Large Language Models
von: Eiras, Francisco, et al.
Veröffentlicht: (2024)
von: Eiras, Francisco, et al.
Veröffentlicht: (2024)
(Im)possibility of Automated Hallucination Detection in Large Language Models
von: Karbasi, Amin, et al.
Veröffentlicht: (2025)
von: Karbasi, Amin, et al.
Veröffentlicht: (2025)
Scalable Token-Level Hallucination Detection in Large Language Models
von: Min, Rui, et al.
Veröffentlicht: (2026)
von: Min, Rui, et al.
Veröffentlicht: (2026)
TruthX: Alleviating Hallucinations by Editing Large Language Models in Truthful Space
von: Zhang, Shaolei, et al.
Veröffentlicht: (2024)
von: Zhang, Shaolei, et al.
Veröffentlicht: (2024)
Large Language Models are Skeptics: False Negative Problem of Input-conflicting Hallucination
von: Song, Jongyoon, et al.
Veröffentlicht: (2024)
von: Song, Jongyoon, et al.
Veröffentlicht: (2024)
Scaling Policy Compliance Assessment in Language Models with Policy Reasoning Traces
von: Imperial, Joseph Marvin, et al.
Veröffentlicht: (2025)
von: Imperial, Joseph Marvin, et al.
Veröffentlicht: (2025)
Lookback Lens: Detecting and Mitigating Contextual Hallucinations in Large Language Models Using Only Attention Maps
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2024)
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2024)
CharED: Character-wise Ensemble Decoding for Large Language Models
von: Gu, Kevin, et al.
Veröffentlicht: (2024)
von: Gu, Kevin, et al.
Veröffentlicht: (2024)
Time Travel in LLMs: Tracing Data Contamination in Large Language Models
von: Golchin, Shahriar, et al.
Veröffentlicht: (2023)
von: Golchin, Shahriar, et al.
Veröffentlicht: (2023)
Anatomy of an Idiom: Tracing Non-Compositionality in Language Models
von: Gomes, Andrew
Veröffentlicht: (2025)
von: Gomes, Andrew
Veröffentlicht: (2025)
Stability-Weighted Decoding for Diffusion Language Models
von: Wu, Yue, et al.
Veröffentlicht: (2026)
von: Wu, Yue, et al.
Veröffentlicht: (2026)
Principled Detection of Hallucinations in Large Language Models via Multiple Testing
von: Li, Jiawei, et al.
Veröffentlicht: (2025)
von: Li, Jiawei, et al.
Veröffentlicht: (2025)
Self-contradictory Hallucinations of Large Language Models: Evaluation, Detection and Mitigation
von: Mündler, Niels, et al.
Veröffentlicht: (2023)
von: Mündler, Niels, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
TDGNet: Hallucination Detection in Diffusion Language Models via Temporal Dynamic Graphs
von: Hemmat, Arshia, et al.
Veröffentlicht: (2026) -
Safer by Diffusion, Broken by Context: Diffusion LLM's Safety Blessing and Its Failure Mode
von: He, Zeyuan, et al.
Veröffentlicht: (2026) -
The Path Matters: Learning a Token-Commitment Policy for Diffusion Language Models
von: Sun, Bohang, et al.
Veröffentlicht: (2026) -
Transcoders Trace Visual Grounding and Hallucinations in Vision-Language Models
von: Damianos, Dimitrios, et al.
Veröffentlicht: (2026) -
MemTrace: Tracing and Attributing Errors in Large Language Model Memory Systems
von: Deng, Xinle, et al.
Veröffentlicht: (2026)