DefAn: Definitive Answer Dataset for LLMs Hallucination Evaluation
Fuente:
arXiv
Guardado en:
| Autores principales: | Rahman, A B M Ashikur, Anwar, Saeed, Usman, Muhammad, Mian, Ajmal |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multistream Network for LiDAR and Camera-based 3D Object Detection in Outdoor Scenes
por: Ibrahim, Muhammad, et al.
Publicado: (2025)
por: Ibrahim, Muhammad, et al.
Publicado: (2025)
Image Colorization: A Survey and Dataset
por: Anwar, Saeed, et al.
Publicado: (2020)
por: Anwar, Saeed, et al.
Publicado: (2020)
FastBO: Fast HPO and NAS with Adaptive Fidelity Identification
por: Jiang, Jiantong, et al.
Publicado: (2024)
por: Jiang, Jiantong, et al.
Publicado: (2024)
Visual Attention Methods in Deep Learning: An In-Depth Survey
por: Hassanin, Mohammed, et al.
Publicado: (2022)
por: Hassanin, Mohammed, et al.
Publicado: (2022)
Deep Models for Multi-View 3D Object Recognition: A Review
por: Alzahrani, Mona, et al.
Publicado: (2024)
por: Alzahrani, Mona, et al.
Publicado: (2024)
Bird Eye-View to Street-View: A Survey
por: Bajbaa, Khawlah, et al.
Publicado: (2024)
por: Bajbaa, Khawlah, et al.
Publicado: (2024)
3D-GRAND: A Million-Scale Dataset for 3D-LLMs with Better Grounding and Less Hallucination
por: Yang, Jianing, et al.
Publicado: (2024)
por: Yang, Jianing, et al.
Publicado: (2024)
Beyond Superficial Unlearning: Sharpness-Aware Robust Erasure of Hallucinations in Multimodal LLMs
por: Fang, Xianya, et al.
Publicado: (2026)
por: Fang, Xianya, et al.
Publicado: (2026)
Mitigating Object and Action Hallucinations in Multimodal LLMs via Self-Augmented Contrastive Alignment
por: Chang, Kai-Po, et al.
Publicado: (2025)
por: Chang, Kai-Po, et al.
Publicado: (2025)
Attention Down-Sampling Transformer, Relative Ranking and Self-Consistency for Blind Image Quality Assessment
por: Alsaafin, Mohammed, et al.
Publicado: (2024)
por: Alsaafin, Mohammed, et al.
Publicado: (2024)
Unified Triplet-Level Hallucination Evaluation for Large Vision-Language Models
por: Wu, Junjie, et al.
Publicado: (2024)
por: Wu, Junjie, et al.
Publicado: (2024)
SegSub: Evaluating Robustness to Knowledge Conflicts and Hallucinations in Vision-Language Models
por: Carragher, Peter, et al.
Publicado: (2025)
por: Carragher, Peter, et al.
Publicado: (2025)
Attribution-Guided Model Rectification of Unreliable Neural Network Behaviors
por: Yang, Peiyu, et al.
Publicado: (2026)
por: Yang, Peiyu, et al.
Publicado: (2026)
Are Hallucinations Bad Estimations?
por: Liu, Hude, et al.
Publicado: (2025)
por: Liu, Hude, et al.
Publicado: (2025)
Improved Crop and Weed Detection with Diverse Data Ensemble Learning
por: Asad, Muhammad Hamza, et al.
Publicado: (2023)
por: Asad, Muhammad Hamza, et al.
Publicado: (2023)
Auto-Regressive Diffusion for Generating 3D Human-Object Interactions
por: Geng, Zichen, et al.
Publicado: (2025)
por: Geng, Zichen, et al.
Publicado: (2025)
Efficient Contrastive Decoding with Probabilistic Hallucination Detection - Mitigating Hallucinations in Large Vision Language Models -
por: Fieback, Laura, et al.
Publicado: (2025)
por: Fieback, Laura, et al.
Publicado: (2025)
Skip Mamba Diffusion for Monocular 3D Semantic Scene Completion
por: Liang, Li, et al.
Publicado: (2025)
por: Liang, Li, et al.
Publicado: (2025)
MediFact at MEDIQA-M3G 2024: Medical Question Answering in Dermatology with Multimodal Learning
por: Saeed, Nadia
Publicado: (2024)
por: Saeed, Nadia
Publicado: (2024)
Steering the Verifiability of Multimodal AI Hallucinations
por: Pang, Jianhong, et al.
Publicado: (2026)
por: Pang, Jianhong, et al.
Publicado: (2026)
VIST-GPT: Ushering in the Era of Visual Storytelling with LLMs?
por: Gado, Mohamed, et al.
Publicado: (2025)
por: Gado, Mohamed, et al.
Publicado: (2025)
Evaluating the Correctness of Inference Patterns Used by LLMs for Judgment
por: Chen, Lu, et al.
Publicado: (2024)
por: Chen, Lu, et al.
Publicado: (2024)
ALOHa: A New Measure for Hallucination in Captioning Models
por: Petryk, Suzanne, et al.
Publicado: (2024)
por: Petryk, Suzanne, et al.
Publicado: (2024)
Steering LVLMs via Sparse Autoencoder for Hallucination Mitigation
por: Hua, Zhenglin, et al.
Publicado: (2025)
por: Hua, Zhenglin, et al.
Publicado: (2025)
Woodpecker: Hallucination Correction for Multimodal Large Language Models
por: Yin, Shukang, et al.
Publicado: (2023)
por: Yin, Shukang, et al.
Publicado: (2023)
When Prompts Override Vision: Prompt-Induced Hallucinations in LVLMs
por: Khayatan, Pegah, et al.
Publicado: (2026)
por: Khayatan, Pegah, et al.
Publicado: (2026)
Counterfactual Segmentation Reasoning: Diagnosing and Mitigating Pixel-Grounding Hallucination
por: Li, Xinzhuo, et al.
Publicado: (2025)
por: Li, Xinzhuo, et al.
Publicado: (2025)
Tables as Texts or Images: Evaluating the Table Reasoning Ability of LLMs and MLLMs
por: Deng, Naihao, et al.
Publicado: (2024)
por: Deng, Naihao, et al.
Publicado: (2024)
How Multimodal LLMs Solve Image Tasks: A Lens on Visual Grounding, Task Reasoning, and Answer Decoding
por: Yu, Zhuoran, et al.
Publicado: (2025)
por: Yu, Zhuoran, et al.
Publicado: (2025)
MSRNet: A Multi-Scale Recursive Network for Camouflaged Object Detection
por: Alghamdi, Leena, et al.
Publicado: (2025)
por: Alghamdi, Leena, et al.
Publicado: (2025)
YOLOatr : Deep Learning Based Automatic Target Detection and Localization in Thermal Infrared Imagery
por: Safdar, Aon, et al.
Publicado: (2025)
por: Safdar, Aon, et al.
Publicado: (2025)
From Instructions to Assistance: a Dataset Aligning Instruction Manuals with Assembly Videos for Evaluating Multimodal LLMs
por: Toschi, Federico, et al.
Publicado: (2026)
por: Toschi, Federico, et al.
Publicado: (2026)
Web2Code: A Large-scale Webpage-to-Code Dataset and Evaluation Framework for Multimodal LLMs
por: Yun, Sukmin, et al.
Publicado: (2024)
por: Yun, Sukmin, et al.
Publicado: (2024)
Talk Less, Interact Better: Evaluating In-context Conversational Adaptation in Multimodal LLMs
por: Hua, Yilun, et al.
Publicado: (2024)
por: Hua, Yilun, et al.
Publicado: (2024)
DENEB: A Hallucination-Robust Automatic Evaluation Metric for Image Captioning
por: Matsuda, Kazuki, et al.
Publicado: (2024)
por: Matsuda, Kazuki, et al.
Publicado: (2024)
Toward More Reliable Artificial Intelligence: Reducing Hallucinations in Vision-Language Models
por: Sanogo, Kassoum, et al.
Publicado: (2025)
por: Sanogo, Kassoum, et al.
Publicado: (2025)
Logical Closed Loop: Uncovering Object Hallucinations in Large Vision-Language Models
por: Wu, Junfei, et al.
Publicado: (2024)
por: Wu, Junfei, et al.
Publicado: (2024)
Mitigating Object Hallucination in Large Vision-Language Models via Image-Grounded Guidance
por: Zhao, Linxi, et al.
Publicado: (2024)
por: Zhao, Linxi, et al.
Publicado: (2024)
Skip \n: A Simple Method to Reduce Hallucination in Large Vision-Language Models
por: Han, Zongbo, et al.
Publicado: (2024)
por: Han, Zongbo, et al.
Publicado: (2024)
Unification of Balti and trans-border sister dialects in the essence of LLMs and AI Technology
por: Sharif, Muhammad, et al.
Publicado: (2024)
por: Sharif, Muhammad, et al.
Publicado: (2024)
Ejemplares similares
-
Multistream Network for LiDAR and Camera-based 3D Object Detection in Outdoor Scenes
por: Ibrahim, Muhammad, et al.
Publicado: (2025) -
Image Colorization: A Survey and Dataset
por: Anwar, Saeed, et al.
Publicado: (2020) -
FastBO: Fast HPO and NAS with Adaptive Fidelity Identification
por: Jiang, Jiantong, et al.
Publicado: (2024) -
Visual Attention Methods in Deep Learning: An In-Depth Survey
por: Hassanin, Mohammed, et al.
Publicado: (2022) -
Deep Models for Multi-View 3D Object Recognition: A Review
por: Alzahrani, Mona, et al.
Publicado: (2024)