DENEB: A Hallucination-Robust Automatic Evaluation Metric for Image Captioning
Fuente:
arXiv
Salvato in:
| Autori principali: | Matsuda, Kazuki, Wada, Yuiga, Sugiura, Komei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
VELA: An LLM-Hybrid-as-a-Judge Approach for Evaluating Long Image Captions
di: Matsuda, Kazuki, et al.
Pubblicazione: (2025)
di: Matsuda, Kazuki, et al.
Pubblicazione: (2025)
ZINA: Multimodal Fine-grained Hallucination Detection and Editing
di: Wada, Yuiga, et al.
Pubblicazione: (2025)
di: Wada, Yuiga, et al.
Pubblicazione: (2025)
Polos: Multimodal Metric Learning from Human Feedback for Image Captioning
di: Wada, Yuiga, et al.
Pubblicazione: (2024)
di: Wada, Yuiga, et al.
Pubblicazione: (2024)
LLM-Free Image Captioning Evaluation in Reference-Flexible Settings
di: Hirano, Shinnosuke, et al.
Pubblicazione: (2025)
di: Hirano, Shinnosuke, et al.
Pubblicazione: (2025)
An Examination of the Robustness of Reference-Free Image Captioning Evaluation Metrics
di: Ahmadi, Saba, et al.
Pubblicazione: (2023)
di: Ahmadi, Saba, et al.
Pubblicazione: (2023)
Deep Space Weather Model: Long-Range Solar Flare Prediction from Multi-Wavelength Images
di: Nagashima, Shunya, et al.
Pubblicazione: (2025)
di: Nagashima, Shunya, et al.
Pubblicazione: (2025)
EXPERT: An Explainable Image Captioning Evaluation Metric with Structured Explanations
di: Kim, Hyunjong, et al.
Pubblicazione: (2025)
di: Kim, Hyunjong, et al.
Pubblicazione: (2025)
Attention Lattice Adapter: Visual Explanation Generation for Visual Foundation Model
di: Hirano, Shinnosuke, et al.
Pubblicazione: (2025)
di: Hirano, Shinnosuke, et al.
Pubblicazione: (2025)
MLLM-as-a-Judge Exhibits Model Preference Bias
di: Koyama, Shuitsu, et al.
Pubblicazione: (2026)
di: Koyama, Shuitsu, et al.
Pubblicazione: (2026)
G-VEval: A Versatile Metric for Evaluating Image and Video Captions Using GPT-4o
di: Tong, Tony Cheng, et al.
Pubblicazione: (2024)
di: Tong, Tony Cheng, et al.
Pubblicazione: (2024)
Capturing Fine-Grained Alignments Improves 3D Affordance Detection
di: Tokumitsu, Junsei, et al.
Pubblicazione: (2025)
di: Tokumitsu, Junsei, et al.
Pubblicazione: (2025)
FLEUR: An Explainable Reference-Free Evaluation Metric for Image Captioning Using a Large Multimodal Model
di: Lee, Yebin, et al.
Pubblicazione: (2024)
di: Lee, Yebin, et al.
Pubblicazione: (2024)
ChartCap: Mitigating Hallucination of Dense Chart Captioning
di: Lim, Junyoung, et al.
Pubblicazione: (2025)
di: Lim, Junyoung, et al.
Pubblicazione: (2025)
Toward Automatic Safe Driving Instruction: A Large-Scale Vision Language Model Approach
di: Sakajo, Haruki, et al.
Pubblicazione: (2025)
di: Sakajo, Haruki, et al.
Pubblicazione: (2025)
DISCODE: Distribution-Aware Score Decoder for Robust Automatic Evaluation of Image Captioning
di: Inoue, Nakamasa, et al.
Pubblicazione: (2025)
di: Inoue, Nakamasa, et al.
Pubblicazione: (2025)
Image Captioning Evaluation in the Age of Multimodal LLMs: Challenges and Future Perspectives
di: Sarto, Sara, et al.
Pubblicazione: (2025)
di: Sarto, Sara, et al.
Pubblicazione: (2025)
ALOHa: A New Measure for Hallucination in Captioning Models
di: Petryk, Suzanne, et al.
Pubblicazione: (2024)
di: Petryk, Suzanne, et al.
Pubblicazione: (2024)
Cortical-SSM: A Deep State Space Model for EEG and ECoG Motor Imagery Decoding
di: Suzuki, Shuntaro, et al.
Pubblicazione: (2025)
di: Suzuki, Shuntaro, et al.
Pubblicazione: (2025)
Text-only Synthesis for Image Captioning
di: Zhou, Qing, et al.
Pubblicazione: (2024)
di: Zhou, Qing, et al.
Pubblicazione: (2024)
The Role of Data Curation in Image Captioning
di: Li, Wenyan, et al.
Pubblicazione: (2023)
di: Li, Wenyan, et al.
Pubblicazione: (2023)
CIC: A Framework for Culturally-Aware Image Captioning
di: Yun, Youngsik, et al.
Pubblicazione: (2024)
di: Yun, Youngsik, et al.
Pubblicazione: (2024)
BRIDGE: Bridging Gaps in Image Captioning Evaluation with Stronger Visual Cues
di: Sarto, Sara, et al.
Pubblicazione: (2024)
di: Sarto, Sara, et al.
Pubblicazione: (2024)
Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models
di: Zhang, Huixuan, et al.
Pubblicazione: (2025)
di: Zhang, Huixuan, et al.
Pubblicazione: (2025)
CAPEEN: Image Captioning with Early Exits and Knowledge Distillation
di: Bajpai, Divya Jyoti, et al.
Pubblicazione: (2024)
di: Bajpai, Divya Jyoti, et al.
Pubblicazione: (2024)
Explaining Caption-Image Interactions in CLIP Models with Second-Order Attributions
di: Möller, Lucas, et al.
Pubblicazione: (2024)
di: Möller, Lucas, et al.
Pubblicazione: (2024)
The Power of Many: Multi-Agent Multimodal Models for Cultural Image Captioning
di: Bai, Longju, et al.
Pubblicazione: (2024)
di: Bai, Longju, et al.
Pubblicazione: (2024)
AC-Lite : A Lightweight Image Captioning Model for Low-Resource Assamese Language
di: Choudhury, Pankaj, et al.
Pubblicazione: (2025)
di: Choudhury, Pankaj, et al.
Pubblicazione: (2025)
Towards Retrieval-Augmented Architectures for Image Captioning
di: Sarto, Sara, et al.
Pubblicazione: (2024)
di: Sarto, Sara, et al.
Pubblicazione: (2024)
CulturalFrames: Assessing Cultural Expectation Alignment in Text-to-Image Models and Evaluation Metrics
di: Nayak, Shravan, et al.
Pubblicazione: (2025)
di: Nayak, Shravan, et al.
Pubblicazione: (2025)
Open-Vocabulary Mobile Manipulation Based on Double Relaxed Contrastive Learning with Dense Labeling
di: Yashima, Daichi, et al.
Pubblicazione: (2024)
di: Yashima, Daichi, et al.
Pubblicazione: (2024)
CapRL: Stimulating Dense Image Caption Capabilities via Reinforcement Learning
di: Xing, Long, et al.
Pubblicazione: (2025)
di: Xing, Long, et al.
Pubblicazione: (2025)
Who Evaluates the Evaluations? Objectively Scoring Text-to-Image Prompt Coherence Metrics with T2IScoreScore (TS2)
di: Saxon, Michael, et al.
Pubblicazione: (2024)
di: Saxon, Michael, et al.
Pubblicazione: (2024)
VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation
di: Ku, Max, et al.
Pubblicazione: (2023)
di: Ku, Max, et al.
Pubblicazione: (2023)
WsiCaption: Multiple Instance Generation of Pathology Reports for Gigapixel Whole-Slide Images
di: Chen, Pingyi, et al.
Pubblicazione: (2023)
di: Chen, Pingyi, et al.
Pubblicazione: (2023)
LAViTeR: Learning Aligned Visual and Textual Representations Assisted by Image and Caption Generation
di: Hashemi, Mohammad Abuzar, et al.
Pubblicazione: (2021)
di: Hashemi, Mohammad Abuzar, et al.
Pubblicazione: (2021)
Mitigating Behavioral Hallucination in Multimodal Large Language Models for Sequential Images
di: You, Liangliang, et al.
Pubblicazione: (2025)
di: You, Liangliang, et al.
Pubblicazione: (2025)
SegSub: Evaluating Robustness to Knowledge Conflicts and Hallucinations in Vision-Language Models
di: Carragher, Peter, et al.
Pubblicazione: (2025)
di: Carragher, Peter, et al.
Pubblicazione: (2025)
SynC: Synthetic Image Caption Dataset Refinement with One-to-many Mapping for Zero-shot Image Captioning
di: Kim, Si-Woo, et al.
Pubblicazione: (2025)
di: Kim, Si-Woo, et al.
Pubblicazione: (2025)
TIAM -- A Metric for Evaluating Alignment in Text-to-Image Generation
di: Grimal, Paul, et al.
Pubblicazione: (2023)
di: Grimal, Paul, et al.
Pubblicazione: (2023)
Unveiling the Invisible: Captioning Videos with Metaphors
di: Kalarani, Abisek Rajakumar, et al.
Pubblicazione: (2024)
di: Kalarani, Abisek Rajakumar, et al.
Pubblicazione: (2024)
Documenti analoghi
-
VELA: An LLM-Hybrid-as-a-Judge Approach for Evaluating Long Image Captions
di: Matsuda, Kazuki, et al.
Pubblicazione: (2025) -
ZINA: Multimodal Fine-grained Hallucination Detection and Editing
di: Wada, Yuiga, et al.
Pubblicazione: (2025) -
Polos: Multimodal Metric Learning from Human Feedback for Image Captioning
di: Wada, Yuiga, et al.
Pubblicazione: (2024) -
LLM-Free Image Captioning Evaluation in Reference-Flexible Settings
di: Hirano, Shinnosuke, et al.
Pubblicazione: (2025) -
An Examination of the Robustness of Reference-Free Image Captioning Evaluation Metrics
di: Ahmadi, Saba, et al.
Pubblicazione: (2023)