HaloScope: Harnessing Unlabeled LLM Generations for Hallucination Detection
Fuente:
arXiv
Guardado en:
| Autores principales: | Du, Xuefeng, Xiao, Chaowei, Li, Yixuan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Steer LLM Latents for Hallucination Detection
por: Park, Seongheon, et al.
Publicado: (2025)
por: Park, Seongheon, et al.
Publicado: (2025)
InterrogateLLM: Zero-Resource Hallucination Detection in LLM-Generated Answers
por: Yehuda, Yakir, et al.
Publicado: (2024)
por: Yehuda, Yakir, et al.
Publicado: (2024)
How Does Unlabeled Data Provably Help Out-of-Distribution Detection?
por: Du, Xuefeng, et al.
Publicado: (2024)
por: Du, Xuefeng, et al.
Publicado: (2024)
Beyond ROUGE: N-Gram Subspace Features for LLM Hallucination Detection
por: Li, Jerry, et al.
Publicado: (2025)
por: Li, Jerry, et al.
Publicado: (2025)
ROM: Real-time Overthinking Mitigation via Streaming Detection and Intervention
por: Wang, Xinyan, et al.
Publicado: (2026)
por: Wang, Xinyan, et al.
Publicado: (2026)
Understanding Multimodal LLMs Under Distribution Shifts: An Information-Theoretic Approach
por: Oh, Changdae, et al.
Publicado: (2025)
por: Oh, Changdae, et al.
Publicado: (2025)
Smoothing Out Hallucinations: Mitigating LLM Hallucination with Smoothed Knowledge Distillation
por: Nguyen, Hieu, et al.
Publicado: (2025)
por: Nguyen, Hieu, et al.
Publicado: (2025)
SLM Meets LLM: Balancing Latency, Interpretability and Consistency in Hallucination Detection
por: Hu, Mengya, et al.
Publicado: (2024)
por: Hu, Mengya, et al.
Publicado: (2024)
Supernodes and Halos: Loss-Critical Hubs in LLM Feed-Forward Layers
por: Cherilyn, Audrey, et al.
Publicado: (2026)
por: Cherilyn, Audrey, et al.
Publicado: (2026)
MALTO at SemEval-2024 Task 6: Leveraging Synthetic Data for LLM Hallucination Detection
por: Borra, Federico, et al.
Publicado: (2024)
por: Borra, Federico, et al.
Publicado: (2024)
Consistency Is the Key: Detecting Hallucinations in LLM Generated Text By Checking Inconsistencies About Key Facts
por: Gupta, Raavi, et al.
Publicado: (2025)
por: Gupta, Raavi, et al.
Publicado: (2025)
The HalluRAG Dataset: Detecting Closed-Domain Hallucinations in RAG Applications Using an LLM's Internal States
por: Ridder, Fabian, et al.
Publicado: (2024)
por: Ridder, Fabian, et al.
Publicado: (2024)
Scalable Token-Level Hallucination Detection in Large Language Models
por: Min, Rui, et al.
Publicado: (2026)
por: Min, Rui, et al.
Publicado: (2026)
Harnessing Uncertainty: Entropy-Modulated Policy Gradients for Long-Horizon LLM Agents
por: Wang, Jiawei, et al.
Publicado: (2025)
por: Wang, Jiawei, et al.
Publicado: (2025)
Detecting AI Hallucinations in Finance: An Information-Theoretic Method Cuts Hallucination Rate by 92%
por: Singha, Mainak
Publicado: (2025)
por: Singha, Mainak
Publicado: (2025)
Real-Time Detection of Hallucinated Entities in Long-Form Generation
por: Obeso, Oscar, et al.
Publicado: (2025)
por: Obeso, Oscar, et al.
Publicado: (2025)
Leveraging Graph Structures to Detect Hallucinations in Large Language Models
por: Nonkes, Noa, et al.
Publicado: (2024)
por: Nonkes, Noa, et al.
Publicado: (2024)
Temporal Graph Network: Hallucination Detection in Multi-Turn Conversation
por: Rathore, Vidhi, et al.
Publicado: (2026)
por: Rathore, Vidhi, et al.
Publicado: (2026)
MMD-Flagger: Leveraging Maximum Mean Discrepancy to Detect Hallucinations
por: Mitsuzawa, Kensuke, et al.
Publicado: (2025)
por: Mitsuzawa, Kensuke, et al.
Publicado: (2025)
Synthesizing Conversations from Unlabeled Documents using Automatic Response Segmentation
por: Wu, Fanyou, et al.
Publicado: (2024)
por: Wu, Fanyou, et al.
Publicado: (2024)
Model-Agnostic Sentiment Distribution Stability Analysis for Robust LLM-Generated Texts Detection
por: Li, Siyuan, et al.
Publicado: (2025)
por: Li, Siyuan, et al.
Publicado: (2025)
Generalization or Hallucination? Understanding Out-of-Context Reasoning in Transformers
por: Huang, Yixiao, et al.
Publicado: (2025)
por: Huang, Yixiao, et al.
Publicado: (2025)
Mitigating LLM Hallucinations via Conformal Abstention
por: Yadkori, Yasin Abbasi, et al.
Publicado: (2024)
por: Yadkori, Yasin Abbasi, et al.
Publicado: (2024)
On Mitigating Code LLM Hallucinations with API Documentation
por: Jain, Nihal, et al.
Publicado: (2024)
por: Jain, Nihal, et al.
Publicado: (2024)
Evaluating the Relevance of Uncertainty Estimators for LLM Hallucination
por: Agnimo, Yedidia, et al.
Publicado: (2026)
por: Agnimo, Yedidia, et al.
Publicado: (2026)
RAGognizer: Hallucination-Aware Fine-Tuning via Detection Head Integration
por: Ridder, Fabian, et al.
Publicado: (2026)
por: Ridder, Fabian, et al.
Publicado: (2026)
Attention Sinks as Internal Signals for Hallucination Detection in Large Language Models
por: Binkowski, Jakub, et al.
Publicado: (2026)
por: Binkowski, Jakub, et al.
Publicado: (2026)
Enhancing LLM Factual Accuracy with RAG to Counter Hallucinations: A Case Study on Domain-Specific Queries in Private Knowledge-Bases
por: Li, Jiarui, et al.
Publicado: (2024)
por: Li, Jiarui, et al.
Publicado: (2024)
Cost-Effective Hallucination Detection for LLMs
por: Valentin, Simon, et al.
Publicado: (2024)
por: Valentin, Simon, et al.
Publicado: (2024)
Learning to Reason for Hallucination Span Detection
por: Su, Hsuan, et al.
Publicado: (2025)
por: Su, Hsuan, et al.
Publicado: (2025)
TwinGate: Stateful Defense against Decompositional Jailbreaks in Untraceable Traffic via Asymmetric Contrastive Learning
por: Sun, Bowen, et al.
Publicado: (2026)
por: Sun, Bowen, et al.
Publicado: (2026)
Hallucination Detection and Mitigation with Diffusion in Multi-Variate Time-Series Foundation Models
por: Wichitwechkarn, Vijja, et al.
Publicado: (2025)
por: Wichitwechkarn, Vijja, et al.
Publicado: (2025)
Detection Without Correction: A Robust Asymmetry in Activation-Based Hallucination Probing
por: Roy, Dip, et al.
Publicado: (2026)
por: Roy, Dip, et al.
Publicado: (2026)
Bolster Hallucination Detection via Prompt-Guided Data Augmentation
por: Li, Wenyun, et al.
Publicado: (2025)
por: Li, Wenyun, et al.
Publicado: (2025)
Dynamic Attention-Guided Context Decoding for Mitigating Context Faithfulness Hallucinations in Large Language Models
por: Huang, Yanwen, et al.
Publicado: (2025)
por: Huang, Yanwen, et al.
Publicado: (2025)
Importing Phantoms: Measuring LLM Package Hallucination Vulnerabilities
por: Krishna, Arjun, et al.
Publicado: (2025)
por: Krishna, Arjun, et al.
Publicado: (2025)
DeepReview: Improving LLM-based Paper Review with Human-like Deep Thinking Process
por: Zhu, Minjun, et al.
Publicado: (2025)
por: Zhu, Minjun, et al.
Publicado: (2025)
TraceDet: Hallucination Detection from the Decoding Trace of Diffusion Large Language Models
por: Chang, Shenxu, et al.
Publicado: (2025)
por: Chang, Shenxu, et al.
Publicado: (2025)
Detecting Token-Level Hallucinations Using Variance Signals: A Reference-Free Approach
por: Kumar, Keshav
Publicado: (2025)
por: Kumar, Keshav
Publicado: (2025)
Intent Recognition and Out-of-Scope Detection using LLMs in Multi-party Conversations
por: Castillo-López, Galo, et al.
Publicado: (2025)
por: Castillo-López, Galo, et al.
Publicado: (2025)
Ejemplares similares
-
Steer LLM Latents for Hallucination Detection
por: Park, Seongheon, et al.
Publicado: (2025) -
InterrogateLLM: Zero-Resource Hallucination Detection in LLM-Generated Answers
por: Yehuda, Yakir, et al.
Publicado: (2024) -
How Does Unlabeled Data Provably Help Out-of-Distribution Detection?
por: Du, Xuefeng, et al.
Publicado: (2024) -
Beyond ROUGE: N-Gram Subspace Features for LLM Hallucination Detection
por: Li, Jerry, et al.
Publicado: (2025) -
ROM: Real-time Overthinking Mitigation via Streaming Detection and Intervention
por: Wang, Xinyan, et al.
Publicado: (2026)