Towards Long Context Hallucination Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Siyi, Halder, Kishaloy, Qi, Zheng, Xiao, Wei, Pappas, Nikolaos, Htut, Phu Mon, John, Neha Anna, Benajiba, Yassine, Roth, Dan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Open Domain Question Answering with Conflicting Contexts
von: Liu, Siyi, et al.
Veröffentlicht: (2024)
von: Liu, Siyi, et al.
Veröffentlicht: (2024)
Rethinking LLM Uncertainty: A Multi-Agent Approach to Estimating Black-Box Model Uncertainty
von: Feng, Yu, et al.
Veröffentlicht: (2024)
von: Feng, Yu, et al.
Veröffentlicht: (2024)
Inference time LLM alignment in single and multidomain preference spectrum
von: Shahriar, Sadat, et al.
Veröffentlicht: (2024)
von: Shahriar, Sadat, et al.
Veröffentlicht: (2024)
Journey Before Destination: On the importance of Visual Faithfulness in Slow Thinking
von: Uppaal, Rheeya, et al.
Veröffentlicht: (2025)
von: Uppaal, Rheeya, et al.
Veröffentlicht: (2025)
Unraveling and Mitigating Safety Alignment Degradation of Vision-Language Models
von: Liu, Qin, et al.
Veröffentlicht: (2024)
von: Liu, Qin, et al.
Veröffentlicht: (2024)
General Purpose Verification for Chain of Thought Prompting
von: Vacareanu, Robert, et al.
Veröffentlicht: (2024)
von: Vacareanu, Robert, et al.
Veröffentlicht: (2024)
Self-supervised Analogical Learning using Language Models
von: Zhou, Ben, et al.
Veröffentlicht: (2025)
von: Zhou, Ben, et al.
Veröffentlicht: (2025)
Can LLMs Narrate Tabular Data? An Evaluation Framework for Natural Language Representations of Text-to-SQL System Outputs
von: Singh, Jyotika, et al.
Veröffentlicht: (2025)
von: Singh, Jyotika, et al.
Veröffentlicht: (2025)
Conflicts in Texts: Data, Implications and Challenges
von: Liu, Siyi, et al.
Veröffentlicht: (2025)
von: Liu, Siyi, et al.
Veröffentlicht: (2025)
DiffuMask: Diffusion Language Model for Token-level Prompt Pruning
von: Zheng, Caleb, et al.
Veröffentlicht: (2026)
von: Zheng, Caleb, et al.
Veröffentlicht: (2026)
Arabic Named Entity Recognition
von: Yassine Benajiba
Veröffentlicht: (2010)
von: Yassine Benajiba
Veröffentlicht: (2010)
Exploration of Plan-Guided Summarization for Narrative Texts: the Case of Small Language Models
von: Grenander, Matt, et al.
Veröffentlicht: (2025)
von: Grenander, Matt, et al.
Veröffentlicht: (2025)
MT-OSC: Path for LLMs that Get Lost in Multi-Turn Conversation
von: Singh, Jyotika, et al.
Veröffentlicht: (2026)
von: Singh, Jyotika, et al.
Veröffentlicht: (2026)
From Instructions to Constraints: Language Model Alignment with Automatic Constraint Verification
von: Wang, Fei, et al.
Veröffentlicht: (2024)
von: Wang, Fei, et al.
Veröffentlicht: (2024)
MemInsight: Autonomous Memory Augmentation for LLM Agents
von: Salama, Rana, et al.
Veröffentlicht: (2025)
von: Salama, Rana, et al.
Veröffentlicht: (2025)
Diable: Efficient Dialogue State Tracking as Operations on Tables
von: Lesci, Pietro, et al.
Veröffentlicht: (2023)
von: Lesci, Pietro, et al.
Veröffentlicht: (2023)
Detecting Contextual Hallucinations in LLMs with Frequency-Aware Attention
von: Qi, Siya, et al.
Veröffentlicht: (2026)
von: Qi, Siya, et al.
Veröffentlicht: (2026)
Enhancing Hallucination Detection via Future Context
von: Lee, Joosung, et al.
Veröffentlicht: (2025)
von: Lee, Joosung, et al.
Veröffentlicht: (2025)
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation
von: Qi, Zheng, et al.
Veröffentlicht: (2025)
von: Qi, Zheng, et al.
Veröffentlicht: (2025)
LUMINA: Detecting Hallucinations in RAG System with Context-Knowledge Signals
von: Yeh, Samuel, et al.
Veröffentlicht: (2025)
von: Yeh, Samuel, et al.
Veröffentlicht: (2025)
Fine-Grained Detection of Context-Grounded Hallucinations Using LLMs
von: Peisakhovsky, Yehonatan, et al.
Veröffentlicht: (2025)
von: Peisakhovsky, Yehonatan, et al.
Veröffentlicht: (2025)
Sanity Checks for Long-Form Hallucination Detection
von: Zollicoffer, Geigh, et al.
Veröffentlicht: (2026)
von: Zollicoffer, Geigh, et al.
Veröffentlicht: (2026)
ESG-Bench: Benchmarking Long-Context ESG Reports for Hallucination Mitigation
von: Sun, Siqi, et al.
Veröffentlicht: (2026)
von: Sun, Siqi, et al.
Veröffentlicht: (2026)
DeAL: Decoding-time Alignment for Large Language Models
von: Huang, James Y., et al.
Veröffentlicht: (2024)
von: Huang, James Y., et al.
Veröffentlicht: (2024)
NewsQs: Multi-Source Question Generation for the Inquiring Mind
von: Hwang, Alyssa, et al.
Veröffentlicht: (2024)
von: Hwang, Alyssa, et al.
Veröffentlicht: (2024)
Semantic Reformulation Entropy for Robust Hallucination Detection in QA Tasks
von: Tong, Chaodong, et al.
Veröffentlicht: (2025)
von: Tong, Chaodong, et al.
Veröffentlicht: (2025)
Talking Point based Ideological Discourse Analysis in News Events
von: Nakshatri, Nishanth, et al.
Veröffentlicht: (2025)
von: Nakshatri, Nishanth, et al.
Veröffentlicht: (2025)
Evaluating LLMs' Assessment of Mixed-Context Hallucination Through the Lens of Summarization
von: Qi, Siya, et al.
Veröffentlicht: (2025)
von: Qi, Siya, et al.
Veröffentlicht: (2025)
HalluGraph: Auditable Hallucination Detection for Legal RAG Systems via Knowledge Graph Alignment
von: Noël, Valentin, et al.
Veröffentlicht: (2025)
von: Noël, Valentin, et al.
Veröffentlicht: (2025)
HalluciNot: Hallucination Detection Through Context and Common Knowledge Verification
von: Paudel, Bibek, et al.
Veröffentlicht: (2025)
von: Paudel, Bibek, et al.
Veröffentlicht: (2025)
From BERT to Qwen: Hate Detection across architectures
von: Mon, Ariadna, et al.
Veröffentlicht: (2025)
von: Mon, Ariadna, et al.
Veröffentlicht: (2025)
Barriers to Discrete Reasoning with Transformers: A Survey Across Depth, Exactness, and Bandwidth
von: Yuan, Michelle, et al.
Veröffentlicht: (2026)
von: Yuan, Michelle, et al.
Veröffentlicht: (2026)
GSM-SEM: Benchmark and Framework for Generating Semantically Variant Augmentations
von: Singh, Jyotika, et al.
Veröffentlicht: (2026)
von: Singh, Jyotika, et al.
Veröffentlicht: (2026)
COMPASS: Context-Modulated PID Attention Steering System for Hallucination Mitigation
von: Sahay, Kenji, et al.
Veröffentlicht: (2025)
von: Sahay, Kenji, et al.
Veröffentlicht: (2025)
Deceptive Semantic Shortcuts on Reasoning Chains: How Far Can Models Go without Hallucination?
von: Li, Bangzheng, et al.
Veröffentlicht: (2023)
von: Li, Bangzheng, et al.
Veröffentlicht: (2023)
Long Context Compression with Activation Beacon
von: Zhang, Peitian, et al.
Veröffentlicht: (2024)
von: Zhang, Peitian, et al.
Veröffentlicht: (2024)
SEGMENT+: Long Text Processing with Short-Context Language Models
von: Shi, Wei, et al.
Veröffentlicht: (2024)
von: Shi, Wei, et al.
Veröffentlicht: (2024)
Long$^2$RAG: Evaluating Long-Context & Long-Form Retrieval-Augmented Generation with Key Point Recall
von: Qi, Zehan, et al.
Veröffentlicht: (2024)
von: Qi, Zehan, et al.
Veröffentlicht: (2024)
HMT: Hierarchical Memory Transformer for Efficient Long Context Language Processing
von: He, Zifan, et al.
Veröffentlicht: (2024)
von: He, Zifan, et al.
Veröffentlicht: (2024)
ReEval: Automatic Hallucination Evaluation for Retrieval-Augmented Large Language Models via Transferable Adversarial Attacks
von: Yu, Xiaodong, et al.
Veröffentlicht: (2023)
von: Yu, Xiaodong, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Open Domain Question Answering with Conflicting Contexts
von: Liu, Siyi, et al.
Veröffentlicht: (2024) -
Rethinking LLM Uncertainty: A Multi-Agent Approach to Estimating Black-Box Model Uncertainty
von: Feng, Yu, et al.
Veröffentlicht: (2024) -
Inference time LLM alignment in single and multidomain preference spectrum
von: Shahriar, Sadat, et al.
Veröffentlicht: (2024) -
Journey Before Destination: On the importance of Visual Faithfulness in Slow Thinking
von: Uppaal, Rheeya, et al.
Veröffentlicht: (2025) -
Unraveling and Mitigating Safety Alignment Degradation of Vision-Language Models
von: Liu, Qin, et al.
Veröffentlicht: (2024)