IRR: Image Review Ranking Framework for Evaluating Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hayashi, Kazuki, Onishi, Kazuma, Suzuki, Toma, Ide, Yusuke, Gobara, Seiji, Saito, Shigeki, Sakai, Yusuke, Kamigaito, Hidetaka, Hayashi, Katsuhiko, Watanabe, Taro |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Artwork Explanation in Large-scale Vision Language Models
von: Hayashi, Kazuki, et al.
Veröffentlicht: (2024)
von: Hayashi, Kazuki, et al.
Veröffentlicht: (2024)
Towards Cross-Lingual Explanation of Artwork in Large-scale Vision Language Models
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2024)
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2024)
Does Pre-trained Language Model Actually Infer Unseen Links in Knowledge Graph Completion?
von: Sakai, Yusuke, et al.
Veröffentlicht: (2023)
von: Sakai, Yusuke, et al.
Veröffentlicht: (2023)
Diagnosing Vision Language Models' Perception by Leveraging Human Methods for Color Vision Deficiencies
von: Hayashi, Kazuki, et al.
Veröffentlicht: (2025)
von: Hayashi, Kazuki, et al.
Veröffentlicht: (2025)
Diversity Explains Inference Scaling Laws: Through a Case Study of Minimum Bayes Risk Decoding
von: Kamigaito, Hidetaka, et al.
Veröffentlicht: (2024)
von: Kamigaito, Hidetaka, et al.
Veröffentlicht: (2024)
Do LLMs Implicitly Determine the Suitable Text Difficulty for Users?
von: Gobara, Seiji, et al.
Veröffentlicht: (2024)
von: Gobara, Seiji, et al.
Veröffentlicht: (2024)
Accurate and Diverse Recommendations via Propensity-Weighted Linear Autoencoders
von: Onishi, Kazuma, et al.
Veröffentlicht: (2025)
von: Onishi, Kazuma, et al.
Veröffentlicht: (2025)
BQA: Body Language Question Answering Dataset for Video Large Language Models
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2024)
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2024)
TextTIGER: Text-based Intelligent Generation with Entity Prompt Refinement for Text-to-Image Generation
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2025)
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2025)
Multi-label Learning with Random Circular Vectors
von: Nishida, Ken, et al.
Veröffentlicht: (2024)
von: Nishida, Ken, et al.
Veröffentlicht: (2024)
Can Impressions of Music be Extracted from Thumbnail Images?
von: Harada, Takashi, et al.
Veröffentlicht: (2025)
von: Harada, Takashi, et al.
Veröffentlicht: (2025)
Revisiting Compositional Generalization Capability of Large Language Models Considering Instruction Following Ability
von: Sakai, Yusuke, et al.
Veröffentlicht: (2025)
von: Sakai, Yusuke, et al.
Veröffentlicht: (2025)
HalluCitation Matters: Revealing the Impact of Hallucinated References with 300 Hallucinated Papers in ACL Conferences
von: Sakai, Yusuke, et al.
Veröffentlicht: (2026)
von: Sakai, Yusuke, et al.
Veröffentlicht: (2026)
HalluCiteChecker: A Lightweight Toolkit for Hallucinated Citation Detection and Verification in the Era of AI Scientists
von: Sakai, Yusuke, et al.
Veröffentlicht: (2026)
von: Sakai, Yusuke, et al.
Veröffentlicht: (2026)
mCSQA: Multilingual Commonsense Reasoning Dataset with Unified Creation Strategy by Language Models and Humans
von: Sakai, Yusuke, et al.
Veröffentlicht: (2024)
von: Sakai, Yusuke, et al.
Veröffentlicht: (2024)
Unified Interpretation of Smoothing Methods for Negative Sampling Loss Functions in Knowledge Graph Embedding
von: Feng, Xincan, et al.
Veröffentlicht: (2024)
von: Feng, Xincan, et al.
Veröffentlicht: (2024)
From Formal Language Theory to Statistical Learning: Finite Observability of Subregular Languages
von: Hayashi, Katsuhiko, et al.
Veröffentlicht: (2025)
von: Hayashi, Katsuhiko, et al.
Veröffentlicht: (2025)
Model-based Subsampling for Knowledge Graph Completion
von: Feng, Xincan, et al.
Veröffentlicht: (2023)
von: Feng, Xincan, et al.
Veröffentlicht: (2023)
IterKey: Iterative Keyword Generation with LLMs for Enhanced Retrieval Augmented Generation
von: Hayashi, Kazuki, et al.
Veröffentlicht: (2025)
von: Hayashi, Kazuki, et al.
Veröffentlicht: (2025)
mbrs: A Library for Minimum Bayes Risk Decoding
von: Deguchi, Hiroyuki, et al.
Veröffentlicht: (2024)
von: Deguchi, Hiroyuki, et al.
Veröffentlicht: (2024)
A Simple but Effective Closed-form Solution for Extreme Multi-label Learning
von: Onishi, Kazuma, et al.
Veröffentlicht: (2025)
von: Onishi, Kazuma, et al.
Veröffentlicht: (2025)
Identifying Influential N-grams in Confidence Calibration via Regression Analysis
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2026)
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2026)
Tonguescape: Exploring Language Models Understanding of Vowel Articulation
von: Sakajo, Haruki, et al.
Veröffentlicht: (2025)
von: Sakajo, Haruki, et al.
Veröffentlicht: (2025)
Multilinguality of Large Language Models From a Structural Perspective
von: Sakajo, Haruki, et al.
Veröffentlicht: (2026)
von: Sakajo, Haruki, et al.
Veröffentlicht: (2026)
How to Make the Most of LLMs' Grammatical Knowledge for Acceptability Judgments
von: Ide, Yusuke, et al.
Veröffentlicht: (2024)
von: Ide, Yusuke, et al.
Veröffentlicht: (2024)
Dictionaries to the Rescue: Cross-Lingual Vocabulary Transfer for Low-Resource Languages Using Bilingual Dictionaries
von: Sakajo, Haruki, et al.
Veröffentlicht: (2025)
von: Sakajo, Haruki, et al.
Veröffentlicht: (2025)
Simultaneous Interpretation Corpus Construction by Large Language Models in Distant Language Pair
von: Sakai, Yusuke, et al.
Veröffentlicht: (2024)
von: Sakai, Yusuke, et al.
Veröffentlicht: (2024)
StructLens: A Structural Lens for Language Models via Maximum Spanning Trees
von: Sakajo, Haruki, et al.
Veröffentlicht: (2026)
von: Sakajo, Haruki, et al.
Veröffentlicht: (2026)
Diversity of Transformer Layers: One Aspect of Parameter Scaling Laws
von: Kamigaito, Hidetaka, et al.
Veröffentlicht: (2025)
von: Kamigaito, Hidetaka, et al.
Veröffentlicht: (2025)
Agreement-Constrained Probabilistic Minimum Bayes Risk Decoding
von: Natsumi, Koki, et al.
Veröffentlicht: (2025)
von: Natsumi, Koki, et al.
Veröffentlicht: (2025)
Toward the Evaluation of Large Language Models Considering Score Variance across Instruction Templates
von: Sakai, Yusuke, et al.
Veröffentlicht: (2024)
von: Sakai, Yusuke, et al.
Veröffentlicht: (2024)
Centroid-Based Efficient Minimum Bayes Risk Decoding
von: Deguchi, Hiroyuki, et al.
Veröffentlicht: (2024)
von: Deguchi, Hiroyuki, et al.
Veröffentlicht: (2024)
CoAM: Corpus of All-Type Multiword Expressions
von: Ide, Yusuke, et al.
Veröffentlicht: (2024)
von: Ide, Yusuke, et al.
Veröffentlicht: (2024)
Understanding the Impact of Confidence in Retrieval Augmented Generation: A Case Study in the Medical Domain
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2024)
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2024)
Enhancing Factuality through Consensus and Consistency in Summarization Using Minimum Bayes Risk Decoding
von: Soetedjo, Riza Setiawan, et al.
Veröffentlicht: (2026)
von: Soetedjo, Riza Setiawan, et al.
Veröffentlicht: (2026)
How Panel Layouts Define Manga: Insights from Visual Ablation Experiments
von: Feng, Siyuan, et al.
Veröffentlicht: (2024)
von: Feng, Siyuan, et al.
Veröffentlicht: (2024)
CArtBench: Evaluating Vision-Language Models on Chinese Art Understanding, Interpretation, and Authenticity
von: Wei, Xuefeng, et al.
Veröffentlicht: (2026)
von: Wei, Xuefeng, et al.
Veröffentlicht: (2026)
Image Referenced Sketch Colorization Based on Animation Creation Workflow
von: Yan, Dingkun, et al.
Veröffentlicht: (2025)
von: Yan, Dingkun, et al.
Veröffentlicht: (2025)
COM Kitchens: An Unedited Overhead-view Video Dataset as a Vision-Language Benchmark
von: Maeda, Koki, et al.
Veröffentlicht: (2024)
von: Maeda, Koki, et al.
Veröffentlicht: (2024)
Zero-Shot Character Identification and Speaker Prediction in Comics via Iterative Multimodal Fusion
von: Li, Yingxuan, et al.
Veröffentlicht: (2024)
von: Li, Yingxuan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Towards Artwork Explanation in Large-scale Vision Language Models
von: Hayashi, Kazuki, et al.
Veröffentlicht: (2024) -
Towards Cross-Lingual Explanation of Artwork in Large-scale Vision Language Models
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2024) -
Does Pre-trained Language Model Actually Infer Unseen Links in Knowledge Graph Completion?
von: Sakai, Yusuke, et al.
Veröffentlicht: (2023) -
Diagnosing Vision Language Models' Perception by Leveraging Human Methods for Color Vision Deficiencies
von: Hayashi, Kazuki, et al.
Veröffentlicht: (2025) -
Diversity Explains Inference Scaling Laws: Through a Case Study of Minimum Bayes Risk Decoding
von: Kamigaito, Hidetaka, et al.
Veröffentlicht: (2024)