ANAH-v2: Scaling Analytical Hallucination Annotation of Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Gu, Yuzhe, Ji, Ziwei, Zhang, Wenwei, Lyu, Chengqi, Lin, Dahua, Chen, Kai |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ANAH: Analytical Annotation of Hallucinations in Large Language Models
di: Ji, Ziwei, et al.
Pubblicazione: (2024)
di: Ji, Ziwei, et al.
Pubblicazione: (2024)
Mask-DPO: Generalizable Fine-grained Factuality Alignment of LLMs
di: Gu, Yuzhe, et al.
Pubblicazione: (2025)
di: Gu, Yuzhe, et al.
Pubblicazione: (2025)
CriticEval: Evaluating Large Language Model as Critic
di: Lan, Tian, et al.
Pubblicazione: (2024)
di: Lan, Tian, et al.
Pubblicazione: (2024)
CompassVerifier: A Unified and Robust Verifier for LLMs Evaluation and Outcome Reward
di: Liu, Shudong, et al.
Pubblicazione: (2025)
di: Liu, Shudong, et al.
Pubblicazione: (2025)
The Imitation Game: Turing Machine Imitator is Length Generalizable Reasoner
di: Hua, Zhouqi, et al.
Pubblicazione: (2025)
di: Hua, Zhouqi, et al.
Pubblicazione: (2025)
Long-horizon Reasoning Agent for Olympiad-Level Mathematical Problem Solving
di: Gao, Songyang, et al.
Pubblicazione: (2025)
di: Gao, Songyang, et al.
Pubblicazione: (2025)
Hallucination is Inevitable: An Innate Limitation of Large Language Models
di: Xu, Ziwei, et al.
Pubblicazione: (2024)
di: Xu, Ziwei, et al.
Pubblicazione: (2024)
InternLM2.5-StepProver: Advancing Automated Theorem Proving via Critic-Guided Search
di: Wu, Zijian, et al.
Pubblicazione: (2024)
di: Wu, Zijian, et al.
Pubblicazione: (2024)
BrainBench: Exposing the Commonsense Reasoning Gap in Large Language Models
di: Tang, Yuzhe
Pubblicazione: (2026)
di: Tang, Yuzhe
Pubblicazione: (2026)
Reference-free Hallucination Detection for Large Vision-Language Models
di: Li, Qing, et al.
Pubblicazione: (2024)
di: Li, Qing, et al.
Pubblicazione: (2024)
Alleviating Hallucinations of Large Language Models through Induced Hallucinations
di: Zhang, Yue, et al.
Pubblicazione: (2023)
di: Zhang, Yue, et al.
Pubblicazione: (2023)
Unveiling the Misuse Potential of Base Large Language Models via In-Context Learning
di: Wang, Xiao, et al.
Pubblicazione: (2024)
di: Wang, Xiao, et al.
Pubblicazione: (2024)
Delusions of Large Language Models
di: Xu, Hongshen, et al.
Pubblicazione: (2025)
di: Xu, Hongshen, et al.
Pubblicazione: (2025)
Large Language Models Are Effective Human Annotation Assistants, But Not Good Independent Annotators
di: Gu, Feng, et al.
Pubblicazione: (2025)
di: Gu, Feng, et al.
Pubblicazione: (2025)
Beyond Fine-Tuning: Effective Strategies for Mitigating Hallucinations in Large Language Models for Data Analytics
di: Rumiantsau, Mikhail, et al.
Pubblicazione: (2024)
di: Rumiantsau, Mikhail, et al.
Pubblicazione: (2024)
MALM: A Multi-Information Adapter for Large Language Models to Mitigate Hallucination
di: Jia, Ao, et al.
Pubblicazione: (2025)
di: Jia, Ao, et al.
Pubblicazione: (2025)
Reducing Hallucinations of Medical Multimodal Large Language Models with Visual Retrieval-Augmented Generation
di: Chu, Yun-Wei, et al.
Pubblicazione: (2025)
di: Chu, Yun-Wei, et al.
Pubblicazione: (2025)
Credal Transformer: A Principled Approach for Quantifying and Mitigating Hallucinations in Large Language Models
di: Ji, Shihao, et al.
Pubblicazione: (2025)
di: Ji, Shihao, et al.
Pubblicazione: (2025)
FoundaBench: Evaluating Chinese Fundamental Knowledge Capabilities of Large Language Models
di: Li, Wei, et al.
Pubblicazione: (2024)
di: Li, Wei, et al.
Pubblicazione: (2024)
Copy-Paste to Mitigate Large Language Model Hallucinations
di: Long, Yongchao, et al.
Pubblicazione: (2025)
di: Long, Yongchao, et al.
Pubblicazione: (2025)
Coarse-to-Fine Highlighting: Reducing Knowledge Hallucination in Large Language Models
di: Lv, Qitan, et al.
Pubblicazione: (2024)
di: Lv, Qitan, et al.
Pubblicazione: (2024)
Entity Alignment with Noisy Annotations from Large Language Models
di: Chen, Shengyuan, et al.
Pubblicazione: (2024)
di: Chen, Shengyuan, et al.
Pubblicazione: (2024)
An Evolutionary Large Language Model for Hallucination Mitigation
di: Boulesnane, Abdennour, et al.
Pubblicazione: (2024)
di: Boulesnane, Abdennour, et al.
Pubblicazione: (2024)
Hallucination Detection: Robustly Discerning Reliable Answers in Large Language Models
di: Chen, Yuyan, et al.
Pubblicazione: (2024)
di: Chen, Yuyan, et al.
Pubblicazione: (2024)
Hal-Eval: A Universal and Fine-grained Hallucination Evaluation Framework for Large Vision Language Models
di: Jiang, Chaoya, et al.
Pubblicazione: (2024)
di: Jiang, Chaoya, et al.
Pubblicazione: (2024)
HICD: Hallucination-Inducing via Attention Dispersion for Contrastive Decoding to Mitigate Hallucinations in Large Language Models
di: Jiang, Xinyan, et al.
Pubblicazione: (2025)
di: Jiang, Xinyan, et al.
Pubblicazione: (2025)
HalluLens: LLM Hallucination Benchmark
di: Bang, Yejin, et al.
Pubblicazione: (2025)
di: Bang, Yejin, et al.
Pubblicazione: (2025)
Delta -- Contrastive Decoding Mitigates Text Hallucinations in Large Language Models
di: Huang, Cheng Peng, et al.
Pubblicazione: (2025)
di: Huang, Cheng Peng, et al.
Pubblicazione: (2025)
LEAN-GitHub: Compiling GitHub LEAN repositories for a versatile LEAN prover
di: Wu, Zijian, et al.
Pubblicazione: (2024)
di: Wu, Zijian, et al.
Pubblicazione: (2024)
PFME: A Modular Approach for Fine-grained Hallucination Detection and Editing of Large Language Models
di: Deng, Kunquan, et al.
Pubblicazione: (2024)
di: Deng, Kunquan, et al.
Pubblicazione: (2024)
Navigating the OverKill in Large Language Models
di: Shi, Chenyu, et al.
Pubblicazione: (2024)
di: Shi, Chenyu, et al.
Pubblicazione: (2024)
Unified Hallucination Detection for Multimodal Large Language Models
di: Chen, Xiang, et al.
Pubblicazione: (2024)
di: Chen, Xiang, et al.
Pubblicazione: (2024)
Active Layer-Contrastive Decoding Reduces Hallucination in Large Language Model Generation
di: Zhang, Hongxiang, et al.
Pubblicazione: (2025)
di: Zhang, Hongxiang, et al.
Pubblicazione: (2025)
The System Hallucination Scale (SHS): A Minimal yet Effective Human-Centered Instrument for Evaluating Hallucination-Related Behavior in Large Language Models
di: Müller, Heimo, et al.
Pubblicazione: (2026)
di: Müller, Heimo, et al.
Pubblicazione: (2026)
Confabulation: The Surprising Value of Large Language Model Hallucinations
di: Sui, Peiqi, et al.
Pubblicazione: (2024)
di: Sui, Peiqi, et al.
Pubblicazione: (2024)
The Impact of Negated Text on Hallucination with Large Language Models
di: Seo, Jaehyung, et al.
Pubblicazione: (2025)
di: Seo, Jaehyung, et al.
Pubblicazione: (2025)
Theoretical Foundations and Mitigation of Hallucination in Large Language Models
di: Gumaan, Esmail
Pubblicazione: (2025)
di: Gumaan, Esmail
Pubblicazione: (2025)
Ada-LEval: Evaluating long-context LLMs with length-adaptable benchmarks
di: Wang, Chonghua, et al.
Pubblicazione: (2024)
di: Wang, Chonghua, et al.
Pubblicazione: (2024)
High-Dimension Human Value Representation in Large Language Models
di: Cahyawijaya, Samuel, et al.
Pubblicazione: (2024)
di: Cahyawijaya, Samuel, et al.
Pubblicazione: (2024)
Micro-Macro Retrieval: Reducing Long-Form Hallucination in Large Language Models
di: Feng, Yujie, et al.
Pubblicazione: (2026)
di: Feng, Yujie, et al.
Pubblicazione: (2026)
Documenti analoghi
-
ANAH: Analytical Annotation of Hallucinations in Large Language Models
di: Ji, Ziwei, et al.
Pubblicazione: (2024) -
Mask-DPO: Generalizable Fine-grained Factuality Alignment of LLMs
di: Gu, Yuzhe, et al.
Pubblicazione: (2025) -
CriticEval: Evaluating Large Language Model as Critic
di: Lan, Tian, et al.
Pubblicazione: (2024) -
CompassVerifier: A Unified and Robust Verifier for LLMs Evaluation and Outcome Reward
di: Liu, Shudong, et al.
Pubblicazione: (2025) -
The Imitation Game: Turing Machine Imitator is Length Generalizable Reasoner
di: Hua, Zhouqi, et al.
Pubblicazione: (2025)