Gespeichert in:
| Hauptverfasser: | Zhao, Yaxin, Zhang, Yu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2511.08916 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AraHalluEval: A Fine-grained Hallucination Evaluation Framework for Arabic LLMs
von: Alansari, Aisha, et al.
Veröffentlicht: (2025)
von: Alansari, Aisha, et al.
Veröffentlicht: (2025)
HalluLens: LLM Hallucination Benchmark
von: Bang, Yejin, et al.
Veröffentlicht: (2025)
von: Bang, Yejin, et al.
Veröffentlicht: (2025)
DSC2025 -- ViHallu Challenge: Detecting Hallucination in Vietnamese LLMs
von: Nguyen, Anh Thi-Hoang, et al.
Veröffentlicht: (2026)
von: Nguyen, Anh Thi-Hoang, et al.
Veröffentlicht: (2026)
HalluShift: Measuring Distribution Shifts towards Hallucination Detection in LLMs
von: Dasgupta, Sharanya, et al.
Veröffentlicht: (2025)
von: Dasgupta, Sharanya, et al.
Veröffentlicht: (2025)
HalluZig: Hallucination Detection using Zigzag Persistence
von: Samaga, Shreyas N., et al.
Veröffentlicht: (2026)
von: Samaga, Shreyas N., et al.
Veröffentlicht: (2026)
HalluCana: Fixing LLM Hallucination with A Canary Lookahead
von: Li, Tianyi, et al.
Veröffentlicht: (2024)
von: Li, Tianyi, et al.
Veröffentlicht: (2024)
HalluCounter: Reference-free LLM Hallucination Detection in the Wild!
von: Urlana, Ashok, et al.
Veröffentlicht: (2025)
von: Urlana, Ashok, et al.
Veröffentlicht: (2025)
HalluEntity: Benchmarking and Understanding Entity-Level Hallucination Detection
von: Yeh, Min-Hsuan, et al.
Veröffentlicht: (2025)
von: Yeh, Min-Hsuan, et al.
Veröffentlicht: (2025)
HalluHard: A Hard Multi-Turn Hallucination Benchmark
von: Fan, Dongyang, et al.
Veröffentlicht: (2026)
von: Fan, Dongyang, et al.
Veröffentlicht: (2026)
HalluScan: A Systematic Benchmark for Detecting and Mitigating Hallucinations in Instruction-Following LLMs
von: Cherif, Ahmed
Veröffentlicht: (2026)
von: Cherif, Ahmed
Veröffentlicht: (2026)
Constructing Benchmarks and Interventions for Combating Hallucinations in LLMs
von: Simhi, Adi, et al.
Veröffentlicht: (2024)
von: Simhi, Adi, et al.
Veröffentlicht: (2024)
AgentHallu: Benchmarking Automated Hallucination Attribution of LLM-based Agents
von: Liu, Xuannan, et al.
Veröffentlicht: (2026)
von: Liu, Xuannan, et al.
Veröffentlicht: (2026)
HalluScore: Large Language Model Hallucination Question Answering Benchmark
von: Alansari, Aisha, et al.
Veröffentlicht: (2026)
von: Alansari, Aisha, et al.
Veröffentlicht: (2026)
HalluWorld: A Controlled Benchmark for Hallucination via Reference World Models
von: Liu, Emmy, et al.
Veröffentlicht: (2026)
von: Liu, Emmy, et al.
Veröffentlicht: (2026)
BenHalluEval: A Multi-Task Hallucination Evaluation Framework for Large Language Models on Bengali
von: Adib, Shefayat E Shams, et al.
Veröffentlicht: (2026)
von: Adib, Shefayat E Shams, et al.
Veröffentlicht: (2026)
HalluDetect: Detecting, Mitigating, and Benchmarking Hallucinations in Conversational Systems in the Legal Domain
von: Anaokar, Spandan, et al.
Veröffentlicht: (2025)
von: Anaokar, Spandan, et al.
Veröffentlicht: (2025)
PerHalluEval: Persian Hallucination Evaluation Benchmark for Large Language Models
von: Hosseini, Mohammad, et al.
Veröffentlicht: (2025)
von: Hosseini, Mohammad, et al.
Veröffentlicht: (2025)
HalluDial: A Large-Scale Benchmark for Automatic Dialogue-Level Hallucination Evaluation
von: Luo, Wen, et al.
Veröffentlicht: (2024)
von: Luo, Wen, et al.
Veröffentlicht: (2024)
FFE-Hallu:Hallucinations in Fixed Figurative Expressions:Benchmark of Idioms and Proverbs in the Persian Language
von: Hosseini, Faezeh, et al.
Veröffentlicht: (2026)
von: Hosseini, Faezeh, et al.
Veröffentlicht: (2026)
HalluCitation Matters: Revealing the Impact of Hallucinated References with 300 Hallucinated Papers in ACL Conferences
von: Sakai, Yusuke, et al.
Veröffentlicht: (2026)
von: Sakai, Yusuke, et al.
Veröffentlicht: (2026)
HalluVerse25: Fine-grained Multilingual Benchmark Dataset for LLM Hallucinations
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
HalluGuard: Evidence-Grounded Small Reasoning Models to Mitigate Hallucinations in Retrieval-Augmented Generation
von: Bergeron, Loris, et al.
Veröffentlicht: (2025)
von: Bergeron, Loris, et al.
Veröffentlicht: (2025)
HalluMix: A Task-Agnostic, Multi-Domain Benchmark for Real-World Hallucination Detection
von: Emery, Deanna, et al.
Veröffentlicht: (2025)
von: Emery, Deanna, et al.
Veröffentlicht: (2025)
HalluSAE: Detecting Hallucinations in Large Language Models via Sparse Auto-Encoders
von: Chen, Boshui, et al.
Veröffentlicht: (2026)
von: Chen, Boshui, et al.
Veröffentlicht: (2026)
MedHallu: A Comprehensive Benchmark for Detecting Medical Hallucinations in Large Language Models
von: Pandit, Shrey, et al.
Veröffentlicht: (2025)
von: Pandit, Shrey, et al.
Veröffentlicht: (2025)
FinReflectKG -- HalluBench: GraphRAG Hallucination Benchmark for Financial Question Answering Systems
von: Kumar, Mahesh, et al.
Veröffentlicht: (2026)
von: Kumar, Mahesh, et al.
Veröffentlicht: (2026)
HalluShift++: Bridging Language and Vision through Internal Representation Shifts for Hierarchical Hallucinations in MLLMs
von: Nath, Sujoy, et al.
Veröffentlicht: (2025)
von: Nath, Sujoy, et al.
Veröffentlicht: (2025)
HalluSearch at SemEval-2025 Task 3: A Search-Enhanced RAG Pipeline for Hallucination Detection
von: Abdallah, Mohamed A., et al.
Veröffentlicht: (2025)
von: Abdallah, Mohamed A., et al.
Veröffentlicht: (2025)
The HalluRAG Dataset: Detecting Closed-Domain Hallucinations in RAG Applications Using an LLM's Internal States
von: Ridder, Fabian, et al.
Veröffentlicht: (2024)
von: Ridder, Fabian, et al.
Veröffentlicht: (2024)
SHARP: Unlocking Interactive Hallucination via Stance Transfer in Role-Playing LLMs
von: Kong, Chuyi, et al.
Veröffentlicht: (2024)
von: Kong, Chuyi, et al.
Veröffentlicht: (2024)
HalluCiteChecker: A Lightweight Toolkit for Hallucinated Citation Detection and Verification in the Era of AI Scientists
von: Sakai, Yusuke, et al.
Veröffentlicht: (2026)
von: Sakai, Yusuke, et al.
Veröffentlicht: (2026)
HalluGraph: Auditable Hallucination Detection for Legal RAG Systems via Knowledge Graph Alignment
von: Noël, Valentin, et al.
Veröffentlicht: (2025)
von: Noël, Valentin, et al.
Veröffentlicht: (2025)
A Unified Hallucination Mitigation Framework for Large Vision-Language Models
von: Chang, Yue, et al.
Veröffentlicht: (2024)
von: Chang, Yue, et al.
Veröffentlicht: (2024)
Combating Confirmation Bias: A Unified Pseudo-Labeling Framework for Entity Alignment
von: Ding, Qijie, et al.
Veröffentlicht: (2023)
von: Ding, Qijie, et al.
Veröffentlicht: (2023)
Towards Detecting LLMs Hallucination via Markov Chain-based Multi-agent Debate Framework
von: Sun, Xiaoxi, et al.
Veröffentlicht: (2024)
von: Sun, Xiaoxi, et al.
Veröffentlicht: (2024)
Acting Flatterers via LLMs Sycophancy: Combating Clickbait with LLMs Opposing-Stance Reasoning
von: Zhang, Chaowei, et al.
Veröffentlicht: (2026)
von: Zhang, Chaowei, et al.
Veröffentlicht: (2026)
Challenging Multilingual LLMs: A New Taxonomy and Benchmark for Unraveling Hallucination in Translation
von: Wu, Xinwei, et al.
Veröffentlicht: (2025)
von: Wu, Xinwei, et al.
Veröffentlicht: (2025)
DRAG: Distilling RAG for SLMs from LLMs to Transfer Knowledge and Mitigate Hallucination via Evidence and Graph-based Distillation
von: Chen, Jennifer, et al.
Veröffentlicht: (2025)
von: Chen, Jennifer, et al.
Veröffentlicht: (2025)
Detecting Contextual Hallucinations in LLMs with Frequency-Aware Attention
von: Qi, Siya, et al.
Veröffentlicht: (2026)
von: Qi, Siya, et al.
Veröffentlicht: (2026)
MathClean: A Benchmark for Synthetic Mathematical Data Cleaning
von: Liang, Hao, et al.
Veröffentlicht: (2025)
von: Liang, Hao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
AraHalluEval: A Fine-grained Hallucination Evaluation Framework for Arabic LLMs
von: Alansari, Aisha, et al.
Veröffentlicht: (2025) -
HalluLens: LLM Hallucination Benchmark
von: Bang, Yejin, et al.
Veröffentlicht: (2025) -
DSC2025 -- ViHallu Challenge: Detecting Hallucination in Vietnamese LLMs
von: Nguyen, Anh Thi-Hoang, et al.
Veröffentlicht: (2026) -
HalluShift: Measuring Distribution Shifts towards Hallucination Detection in LLMs
von: Dasgupta, Sharanya, et al.
Veröffentlicht: (2025) -
HalluZig: Hallucination Detection using Zigzag Persistence
von: Samaga, Shreyas N., et al.
Veröffentlicht: (2026)