Hallucination Detection: Robustly Discerning Reliable Answers in Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Yuyan, Fu, Qiang, Yuan, Yichen, Wen, Zhihao, Fan, Ge, Liu, Dayiheng, Zhang, Dongmei, Li, Zhixu, Xiao, Yanghua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MAPO: Boosting Large Language Model Performance with Model-Adaptive Prompt Optimization
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
Hadamard Adapter: An Extreme Parameter-Efficient Adapter Tuning Method for Pre-trained Language Models
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
Recent Advancement of Emotion Cognition in Large Language Models
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
Can Pre-trained Language Models Understand Chinese Humor?
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
OVEL: Large Language Model as Memory Manager for Online Video Entity Linking
von: Zhao, Haiquan, et al.
Veröffentlicht: (2024)
von: Zhao, Haiquan, et al.
Veröffentlicht: (2024)
Dr.Academy: A Benchmark for Evaluating Questioning Capability in Education for Large Language Models
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
Why Did Apple Fall: Evaluating Curiosity in Large Language Models
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
EmotionQueen: A Benchmark for Evaluating Empathy of Large Language Models
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
Is There a One-Model-Fits-All Approach to Information Extraction? Revisiting Task Definition Biases
von: Huang, Wenhao, et al.
Veröffentlicht: (2024)
von: Huang, Wenhao, et al.
Veröffentlicht: (2024)
Do Large Language Models have Problem-Solving Capability under Incomplete Information Scenarios?
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
Hallucination Detection via Internal States and Structured Reasoning Consistency in Large Language Models
von: Song, Yusheng, et al.
Veröffentlicht: (2025)
von: Song, Yusheng, et al.
Veröffentlicht: (2025)
Joint Evaluation of Answer and Reasoning Consistency for Hallucination Detection in Large Reasoning Models
von: Wang, Changyue, et al.
Veröffentlicht: (2025)
von: Wang, Changyue, et al.
Veröffentlicht: (2025)
Adaptive Ordered Information Extraction with Deep Reinforcement Learning
von: Huang, Wenhao, et al.
Veröffentlicht: (2023)
von: Huang, Wenhao, et al.
Veröffentlicht: (2023)
Unified Hallucination Detection for Multimodal Large Language Models
von: Chen, Xiang, et al.
Veröffentlicht: (2024)
von: Chen, Xiang, et al.
Veröffentlicht: (2024)
Judge Before Answer: Can MLLM Discern the False Premise in Question?
von: Li, Jidong, et al.
Veröffentlicht: (2025)
von: Li, Jidong, et al.
Veröffentlicht: (2025)
Enhancing Answer Reliability Through Inter-Model Consensus of Large Language Models
von: Amiri-Margavi, Alireza, et al.
Veröffentlicht: (2024)
von: Amiri-Margavi, Alireza, et al.
Veröffentlicht: (2024)
Mitigating Entity-Level Hallucination in Large Language Models
von: Su, Weihang, et al.
Veröffentlicht: (2024)
von: Su, Weihang, et al.
Veröffentlicht: (2024)
XMeCap: Meme Caption Generation with Sub-Image Adaptability
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
ESC-Eval: Evaluating Emotion Support Conversations in Large Language Models
von: Zhao, Haiquan, et al.
Veröffentlicht: (2024)
von: Zhao, Haiquan, et al.
Veröffentlicht: (2024)
Span-Level Hallucination Detection for LLM-Generated Answers
von: Elchafei, Passant, et al.
Veröffentlicht: (2025)
von: Elchafei, Passant, et al.
Veröffentlicht: (2025)
Attention-guided Self-reflection for Zero-shot Hallucination Detection in Large Language Models
von: Liu, Qiang, et al.
Veröffentlicht: (2025)
von: Liu, Qiang, et al.
Veröffentlicht: (2025)
AutoScraper: A Progressive Understanding Web Agent for Web Scraper Generation
von: Huang, Wenhao, et al.
Veröffentlicht: (2024)
von: Huang, Wenhao, et al.
Veröffentlicht: (2024)
RECKON: Large-scale Reference-based Efficient Knowledge Evaluation for Large Language Model
von: Zhang, Lin, et al.
Veröffentlicht: (2025)
von: Zhang, Lin, et al.
Veröffentlicht: (2025)
TravelAgent: An AI Assistant for Personalized Travel Planning
von: Chen, Aili, et al.
Veröffentlicht: (2024)
von: Chen, Aili, et al.
Veröffentlicht: (2024)
Hallucination Detection and Evaluation of Large Language Model
von: Zhang, Chenggong, et al.
Veröffentlicht: (2025)
von: Zhang, Chenggong, et al.
Veröffentlicht: (2025)
Diagnosing Failures in Large Language Models' Answers: Integrating Error Attribution into Evaluation Framework
von: Xu, Zishan, et al.
Veröffentlicht: (2025)
von: Xu, Zishan, et al.
Veröffentlicht: (2025)
DetectBench: Can Large Language Model Detect and Piece Together Implicit Evidence?
von: Gu, Zhouhong, et al.
Veröffentlicht: (2024)
von: Gu, Zhouhong, et al.
Veröffentlicht: (2024)
How Easily do Irrelevant Inputs Skew the Responses of Large Language Models?
von: Wu, Siye, et al.
Veröffentlicht: (2024)
von: Wu, Siye, et al.
Veröffentlicht: (2024)
Improving Recall of Large Language Models: A Model Collaboration Approach for Relational Triple Extraction
von: Ding, Zepeng, et al.
Veröffentlicht: (2024)
von: Ding, Zepeng, et al.
Veröffentlicht: (2024)
Framework for Hallucination Detection in Large Language Models
von: Dheeraj Sundaragiri, et al.
Veröffentlicht: (2026)
von: Dheeraj Sundaragiri, et al.
Veröffentlicht: (2026)
GumbelSoft: Diversified Language Model Watermarking via the GumbelMax-trick
von: Fu, Jiayi, et al.
Veröffentlicht: (2024)
von: Fu, Jiayi, et al.
Veröffentlicht: (2024)
HalluSAE: Detecting Hallucinations in Large Language Models via Sparse Auto-Encoders
von: Chen, Boshui, et al.
Veröffentlicht: (2026)
von: Chen, Boshui, et al.
Veröffentlicht: (2026)
Delta -- Contrastive Decoding Mitigates Text Hallucinations in Large Language Models
von: Huang, Cheng Peng, et al.
Veröffentlicht: (2025)
von: Huang, Cheng Peng, et al.
Veröffentlicht: (2025)
Past Meets Present: Creating Historical Analogy with Large Language Models
von: Li, Nianqi, et al.
Veröffentlicht: (2024)
von: Li, Nianqi, et al.
Veröffentlicht: (2024)
Structured Reasoning for Large Language Models
von: Han, Jinyi, et al.
Veröffentlicht: (2026)
von: Han, Jinyi, et al.
Veröffentlicht: (2026)
Poly-FEVER: A Multilingual Fact Verification Benchmark for Hallucination Detection in Large Language Models
von: Zhang, Hanzhi, et al.
Veröffentlicht: (2025)
von: Zhang, Hanzhi, et al.
Veröffentlicht: (2025)
Exploring and Mitigating Fawning Hallucinations in Large Language Models
von: Shangguan, Zixuan, et al.
Veröffentlicht: (2025)
von: Shangguan, Zixuan, et al.
Veröffentlicht: (2025)
Towards Lightweight Reliability: Using Soft Prompts for Hallucination Mitigation in Large Language Models
von: Siddiqui, S M Tahmid, et al.
Veröffentlicht: (2026)
von: Siddiqui, S M Tahmid, et al.
Veröffentlicht: (2026)
Toward Reliable Scientific Hypothesis Generation: Evaluating Truthfulness and Hallucination in Large Language Models
von: Xiong, Guangzhi, et al.
Veröffentlicht: (2025)
von: Xiong, Guangzhi, et al.
Veröffentlicht: (2025)
MetaGPT: Merging Large Language Models Using Model Exclusive Task Arithmetic
von: Zhou, Yuyan, et al.
Veröffentlicht: (2024)
von: Zhou, Yuyan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MAPO: Boosting Large Language Model Performance with Model-Adaptive Prompt Optimization
von: Chen, Yuyan, et al.
Veröffentlicht: (2024) -
Hadamard Adapter: An Extreme Parameter-Efficient Adapter Tuning Method for Pre-trained Language Models
von: Chen, Yuyan, et al.
Veröffentlicht: (2024) -
Recent Advancement of Emotion Cognition in Large Language Models
von: Chen, Yuyan, et al.
Veröffentlicht: (2024) -
Can Pre-trained Language Models Understand Chinese Humor?
von: Chen, Yuyan, et al.
Veröffentlicht: (2024) -
OVEL: Large Language Model as Memory Manager for Online Video Entity Linking
von: Zhao, Haiquan, et al.
Veröffentlicht: (2024)