MeasHalu: Mitigation of Scientific Measurement Hallucinations for Large Language Models with Enhanced Reasoning
Fuente:
arXiv
Guardado en:
| Autores principales: | Huang, Ruijun, Kang, Zhiqiao, Zhu, Yuxuan, Li, Junxiong, Zhao, Jiahao, Tan, Minghuan, Jiang, Feng, Yang, Min |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
HaluEval-Wild: Evaluating Hallucinations of Language Models in the Wild
por: Zhu, Zhiying, et al.
Publicado: (2024)
por: Zhu, Zhiying, et al.
Publicado: (2024)
MedHalu: Hallucinations in Responses to Healthcare Queries by Large Language Models
por: Agarwal, Vibhor, et al.
Publicado: (2024)
por: Agarwal, Vibhor, et al.
Publicado: (2024)
NUMCoT: Numerals and Units of Measurement in Chain-of-Thought Reasoning using Large Language Models
por: Xu, Ancheng, et al.
Publicado: (2024)
por: Xu, Ancheng, et al.
Publicado: (2024)
DiaHalu: A Dialogue-level Hallucination Evaluation Benchmark for Large Language Models
por: Chen, Kedi, et al.
Publicado: (2024)
por: Chen, Kedi, et al.
Publicado: (2024)
Halu-J: Critique-Based Hallucination Judge
por: Wang, Binjie, et al.
Publicado: (2024)
por: Wang, Binjie, et al.
Publicado: (2024)
HaluMem: Evaluating Hallucinations in Memory Systems of Agents
por: Chen, Ding, et al.
Publicado: (2025)
por: Chen, Ding, et al.
Publicado: (2025)
HaluNet: Learning Hallucination Risk from Internal Signals in LLM Question Answering
por: Tong, Chaodong, et al.
Publicado: (2025)
por: Tong, Chaodong, et al.
Publicado: (2025)
Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models
por: Zhong, Weihong, et al.
Publicado: (2024)
por: Zhong, Weihong, et al.
Publicado: (2024)
CodeHalu: Investigating Code Hallucinations in LLMs via Execution-based Verification
por: Tian, Yuchen, et al.
Publicado: (2024)
por: Tian, Yuchen, et al.
Publicado: (2024)
RxSafeBench: Identifying Medication Safety Issues of Large Language Models in Simulated Consultation
por: Zhao, Jiahao, et al.
Publicado: (2025)
por: Zhao, Jiahao, et al.
Publicado: (2025)
Mitigating Hallucinations in Large Language Models via Causal Reasoning
por: Li, Yuangang, et al.
Publicado: (2025)
por: Li, Yuangang, et al.
Publicado: (2025)
Mea culpa
por: ARMANDO ROMERO
Publicado: (2003)
por: ARMANDO ROMERO
Publicado: (2003)
SymLoc: Symbolic Localization of Hallucination across HaluEval and TruthfulQA
por: Lamba, Naveen, et al.
Publicado: (2025)
por: Lamba, Naveen, et al.
Publicado: (2025)
Evaluating and Enhancing Large Language Models for Conversational Reasoning on Knowledge Graphs
por: Huang, Yuxuan
Publicado: (2023)
por: Huang, Yuxuan
Publicado: (2023)
Mitigating Prompt-Induced Hallucinations in Large Language Models via Structured Reasoning
por: Hao, Jinbo, et al.
Publicado: (2026)
por: Hao, Jinbo, et al.
Publicado: (2026)
CollectiveSFT: Scaling Large Language Models for Chinese Medical Benchmark with Collective Instructions in Healthcare
por: Zhu, Jingwei, et al.
Publicado: (2024)
por: Zhu, Jingwei, et al.
Publicado: (2024)
Mitigating Hallucinations in Large Vision-Language Models via Summary-Guided Decoding
por: Min, Kyungmin, et al.
Publicado: (2024)
por: Min, Kyungmin, et al.
Publicado: (2024)
KnowHalu: Hallucination Detection via Multi-Form Knowledge Based Factual Checking
por: Zhang, Jiawei, et al.
Publicado: (2024)
por: Zhang, Jiawei, et al.
Publicado: (2024)
Investigating Symbolic Triggers of Hallucination in Gemma Models Across HaluEval and TruthfulQA
por: Lamba, Naveen, et al.
Publicado: (2025)
por: Lamba, Naveen, et al.
Publicado: (2025)
SDCD: Structure-Disrupted Contrastive Decoding for Mitigating Hallucinations in Large Vision-Language Models
por: Xia, Yuxuan, et al.
Publicado: (2026)
por: Xia, Yuxuan, et al.
Publicado: (2026)
Counterfactual Probing for Hallucination Detection and Mitigation in Large Language Models
por: Feng, Yijun
Publicado: (2025)
por: Feng, Yijun
Publicado: (2025)
Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention
por: An, Wenbin, et al.
Publicado: (2024)
por: An, Wenbin, et al.
Publicado: (2024)
HICD: Hallucination-Inducing via Attention Dispersion for Contrastive Decoding to Mitigate Hallucinations in Large Language Models
por: Jiang, Xinyan, et al.
Publicado: (2025)
por: Jiang, Xinyan, et al.
Publicado: (2025)
Mitigating Hallucination in Large Vision-Language Models through Aligning Attention Distribution to Information Flow
por: Zhao, Jianfei, et al.
Publicado: (2025)
por: Zhao, Jianfei, et al.
Publicado: (2025)
Mitigating Hallucination in Multimodal Large Language Model via Hallucination-targeted Direct Preference Optimization
por: Fu, Yuhan, et al.
Publicado: (2024)
por: Fu, Yuhan, et al.
Publicado: (2024)
VERHallu: Evaluating and Mitigating Event Relation Hallucination in Video Large Language Models
por: Zhang, Zefan, et al.
Publicado: (2026)
por: Zhang, Zefan, et al.
Publicado: (2026)
LongEmotion: Measuring Emotional Intelligence of Large Language Models in Long-Context Interaction
por: Liu, Weichu, et al.
Publicado: (2025)
por: Liu, Weichu, et al.
Publicado: (2025)
Enhancing Visual Reliance in Text Generation: A Bayesian Perspective on Mitigating Hallucination in Large Vision-Language Models
por: Hu, Nanxing, et al.
Publicado: (2025)
por: Hu, Nanxing, et al.
Publicado: (2025)
Reasoning Matters: Mitigate Hallucination in Multimodal Large Reasoning Models via Reasoning-Conditioned Preference Optimization
por: Kong, Jiawei, et al.
Publicado: (2026)
por: Kong, Jiawei, et al.
Publicado: (2026)
Mitigating Hallucinations in Large Vision-Language Models via Entity-Centric Multimodal Preference Optimization
por: Wu, Jiulong, et al.
Publicado: (2025)
por: Wu, Jiulong, et al.
Publicado: (2025)
Look Carefully: Adaptive Visual Reinforcements in Multimodal Large Language Models for Hallucination Mitigation
por: Zhu, Xingyu, et al.
Publicado: (2026)
por: Zhu, Xingyu, et al.
Publicado: (2026)
SAKED: Mitigating Hallucination in Large Vision-Language Models via Stability-Aware Knowledge Enhanced Decoding
por: Li, Zhaoxu, et al.
Publicado: (2026)
por: Li, Zhaoxu, et al.
Publicado: (2026)
Whispers that Shake Foundations: Analyzing and Mitigating False Premise Hallucinations in Large Language Models
por: Yuan, Hongbang, et al.
Publicado: (2024)
por: Yuan, Hongbang, et al.
Publicado: (2024)
Retrieval Visual Contrastive Decoding to Mitigate Object Hallucinations in Large Vision-Language Models
por: Lee, Jihoon, et al.
Publicado: (2025)
por: Lee, Jihoon, et al.
Publicado: (2025)
Mitigating Hallucinations in Large Vision-Language Models without Performance Degradation
por: Zhu, Xingyu, et al.
Publicado: (2026)
por: Zhu, Xingyu, et al.
Publicado: (2026)
Detecting and Mitigating Hallucination in Large Vision Language Models via Fine-Grained AI Feedback
por: Xiao, Wenyi, et al.
Publicado: (2024)
por: Xiao, Wenyi, et al.
Publicado: (2024)
Ever: Mitigating Hallucination in Large Language Models through Real-Time Verification and Rectification
por: Kang, Haoqiang, et al.
Publicado: (2023)
por: Kang, Haoqiang, et al.
Publicado: (2023)
RandomMeas.jl: A Julia Package for Randomized Measurements in Quantum Devices
por: Elben, Andreas, et al.
Publicado: (2025)
por: Elben, Andreas, et al.
Publicado: (2025)
VEGAS: Mitigating Hallucinations in Large Vision-Language Models via Vision-Encoder Attention Guided Adaptive Steering
por: Wang, Zihu, et al.
Publicado: (2025)
por: Wang, Zihu, et al.
Publicado: (2025)
An Evolutionary Large Language Model for Hallucination Mitigation
por: Boulesnane, Abdennour, et al.
Publicado: (2024)
por: Boulesnane, Abdennour, et al.
Publicado: (2024)
Ejemplares similares
-
HaluEval-Wild: Evaluating Hallucinations of Language Models in the Wild
por: Zhu, Zhiying, et al.
Publicado: (2024) -
MedHalu: Hallucinations in Responses to Healthcare Queries by Large Language Models
por: Agarwal, Vibhor, et al.
Publicado: (2024) -
NUMCoT: Numerals and Units of Measurement in Chain-of-Thought Reasoning using Large Language Models
por: Xu, Ancheng, et al.
Publicado: (2024) -
DiaHalu: A Dialogue-level Hallucination Evaluation Benchmark for Large Language Models
por: Chen, Kedi, et al.
Publicado: (2024) -
Halu-J: Critique-Based Hallucination Judge
por: Wang, Binjie, et al.
Publicado: (2024)