CArtBench: Evaluating Vision-Language Models on Chinese Art Understanding, Interpretation, and Authenticity
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wei, Xuefeng, Wang, Zhixuan, Zhou, Xuan, Qu, Zhi, Li, Hongyao, Sakai, Yusuke, Kamigaito, Hidetaka, Watanabe, Taro |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Simultaneous Interpretation Corpus Construction by Large Language Models in Distant Language Pair
par: Sakai, Yusuke, et autres
Publié: (2024)
par: Sakai, Yusuke, et autres
Publié: (2024)
Revisiting Compositional Generalization Capability of Large Language Models Considering Instruction Following Ability
par: Sakai, Yusuke, et autres
Publié: (2025)
par: Sakai, Yusuke, et autres
Publié: (2025)
mCSQA: Multilingual Commonsense Reasoning Dataset with Unified Creation Strategy by Language Models and Humans
par: Sakai, Yusuke, et autres
Publié: (2024)
par: Sakai, Yusuke, et autres
Publié: (2024)
Tonguescape: Exploring Language Models Understanding of Vowel Articulation
par: Sakajo, Haruki, et autres
Publié: (2025)
par: Sakajo, Haruki, et autres
Publié: (2025)
HalluCitation Matters: Revealing the Impact of Hallucinated References with 300 Hallucinated Papers in ACL Conferences
par: Sakai, Yusuke, et autres
Publié: (2026)
par: Sakai, Yusuke, et autres
Publié: (2026)
HalluCiteChecker: A Lightweight Toolkit for Hallucinated Citation Detection and Verification in the Era of AI Scientists
par: Sakai, Yusuke, et autres
Publié: (2026)
par: Sakai, Yusuke, et autres
Publié: (2026)
Diagnosing Vision Language Models' Perception by Leveraging Human Methods for Color Vision Deficiencies
par: Hayashi, Kazuki, et autres
Publié: (2025)
par: Hayashi, Kazuki, et autres
Publié: (2025)
Multilinguality of Large Language Models From a Structural Perspective
par: Sakajo, Haruki, et autres
Publié: (2026)
par: Sakajo, Haruki, et autres
Publié: (2026)
Towards Artwork Explanation in Large-scale Vision Language Models
par: Hayashi, Kazuki, et autres
Publié: (2024)
par: Hayashi, Kazuki, et autres
Publié: (2024)
Does Pre-trained Language Model Actually Infer Unseen Links in Knowledge Graph Completion?
par: Sakai, Yusuke, et autres
Publié: (2023)
par: Sakai, Yusuke, et autres
Publié: (2023)
Toward the Evaluation of Large Language Models Considering Score Variance across Instruction Templates
par: Sakai, Yusuke, et autres
Publié: (2024)
par: Sakai, Yusuke, et autres
Publié: (2024)
mbrs: A Library for Minimum Bayes Risk Decoding
par: Deguchi, Hiroyuki, et autres
Publié: (2024)
par: Deguchi, Hiroyuki, et autres
Publié: (2024)
Exploring Intrinsic Language-specific Subspaces in Fine-tuning Multilingual Neural Machine Translation
par: Cao, Zhe, et autres
Publié: (2024)
par: Cao, Zhe, et autres
Publié: (2024)
Towards Cross-Lingual Explanation of Artwork in Large-scale Vision Language Models
par: Ozaki, Shintaro, et autres
Publié: (2024)
par: Ozaki, Shintaro, et autres
Publié: (2024)
StructLens: A Structural Lens for Language Models via Maximum Spanning Trees
par: Sakajo, Haruki, et autres
Publié: (2026)
par: Sakajo, Haruki, et autres
Publié: (2026)
Agreement-Constrained Probabilistic Minimum Bayes Risk Decoding
par: Natsumi, Koki, et autres
Publié: (2025)
par: Natsumi, Koki, et autres
Publié: (2025)
Diversity Explains Inference Scaling Laws: Through a Case Study of Minimum Bayes Risk Decoding
par: Kamigaito, Hidetaka, et autres
Publié: (2024)
par: Kamigaito, Hidetaka, et autres
Publié: (2024)
BQA: Body Language Question Answering Dataset for Video Large Language Models
par: Ozaki, Shintaro, et autres
Publié: (2024)
par: Ozaki, Shintaro, et autres
Publié: (2024)
IRR: Image Review Ranking Framework for Evaluating Vision-Language Models
par: Hayashi, Kazuki, et autres
Publié: (2024)
par: Hayashi, Kazuki, et autres
Publié: (2024)
Efficient Nearest Neighbor based Uncertainty Estimation for Natural Language Processing Tasks
par: Hashimoto, Wataru, et autres
Publié: (2024)
par: Hashimoto, Wataru, et autres
Publié: (2024)
Decoding Uncertainty: The Impact of Decoding Strategies for Uncertainty Estimation in Large Language Models
par: Hashimoto, Wataru, et autres
Publié: (2025)
par: Hashimoto, Wataru, et autres
Publié: (2025)
SinhalaMMLU: A Comprehensive Benchmark for Evaluating Multitask Language Understanding in Sinhala
par: Pramodya, Ashmari, et autres
Publié: (2025)
par: Pramodya, Ashmari, et autres
Publié: (2025)
Dictionaries to the Rescue: Cross-Lingual Vocabulary Transfer for Low-Resource Languages Using Bilingual Dictionaries
par: Sakajo, Haruki, et autres
Publié: (2025)
par: Sakajo, Haruki, et autres
Publié: (2025)
Unified Interpretation of Smoothing Methods for Negative Sampling Loss Functions in Knowledge Graph Embedding
par: Feng, Xincan, et autres
Publié: (2024)
par: Feng, Xincan, et autres
Publié: (2024)
Centroid-Based Efficient Minimum Bayes Risk Decoding
par: Deguchi, Hiroyuki, et autres
Publié: (2024)
par: Deguchi, Hiroyuki, et autres
Publié: (2024)
Are Data Augmentation Methods in Named Entity Recognition Applicable for Uncertainty Estimation?
par: Hashimoto, Wataru, et autres
Publié: (2024)
par: Hashimoto, Wataru, et autres
Publié: (2024)
Do LLMs Implicitly Determine the Suitable Text Difficulty for Users?
par: Gobara, Seiji, et autres
Publié: (2024)
par: Gobara, Seiji, et autres
Publié: (2024)
Attention Score is not All You Need for Token Importance Indicator in KV Cache Reduction: Value Also Matters
par: Guo, Zhiyu, et autres
Publié: (2024)
par: Guo, Zhiyu, et autres
Publié: (2024)
How to Make the Most of LLMs' Grammatical Knowledge for Acceptability Judgments
par: Ide, Yusuke, et autres
Publié: (2024)
par: Ide, Yusuke, et autres
Publié: (2024)
Enhancing Factuality through Consensus and Consistency in Summarization Using Minimum Bayes Risk Decoding
par: Soetedjo, Riza Setiawan, et autres
Publié: (2026)
par: Soetedjo, Riza Setiawan, et autres
Publié: (2026)
Introducing Syllable Tokenization for Low-resource Languages: A Case Study with Swahili
par: Atuhurra, Jesse, et autres
Publié: (2024)
par: Atuhurra, Jesse, et autres
Publié: (2024)
Context-Aware Machine Translation with Source Coreference Explanation
par: Vu, Huy Hien, et autres
Publié: (2024)
par: Vu, Huy Hien, et autres
Publié: (2024)
XQ-MEval: A Dataset with Cross-lingual Parallel Quality for Benchmarking Translation Metrics
par: Liu, Jingxuan, et autres
Publié: (2026)
par: Liu, Jingxuan, et autres
Publié: (2026)
Dependency-Aware Semi-Structured Sparsity of GLU Variants in Large Language Models
par: Guo, Zhiyu, et autres
Publié: (2024)
par: Guo, Zhiyu, et autres
Publié: (2024)
Toward Automatic Safe Driving Instruction: A Large-Scale Vision Language Model Approach
par: Sakajo, Haruki, et autres
Publié: (2025)
par: Sakajo, Haruki, et autres
Publié: (2025)
Constructing Multilingual Visual-Text Datasets Revealing Visual Multilingual Ability of Vision Language Models
par: Atuhurra, Jesse, et autres
Publié: (2024)
par: Atuhurra, Jesse, et autres
Publié: (2024)
Cross-lingual Contextualized Phrase Retrieval
par: Li, Huayang, et autres
Publié: (2024)
par: Li, Huayang, et autres
Publié: (2024)
IterKey: Iterative Keyword Generation with LLMs for Enhanced Retrieval Augmented Generation
par: Hayashi, Kazuki, et autres
Publié: (2025)
par: Hayashi, Kazuki, et autres
Publié: (2025)
Generating Diverse Translation with Perturbed kNN-MT
par: Nishida, Yuto, et autres
Publié: (2024)
par: Nishida, Yuto, et autres
Publié: (2024)
Model-based Subsampling for Knowledge Graph Completion
par: Feng, Xincan, et autres
Publié: (2023)
par: Feng, Xincan, et autres
Publié: (2023)
Documents similaires
-
Simultaneous Interpretation Corpus Construction by Large Language Models in Distant Language Pair
par: Sakai, Yusuke, et autres
Publié: (2024) -
Revisiting Compositional Generalization Capability of Large Language Models Considering Instruction Following Ability
par: Sakai, Yusuke, et autres
Publié: (2025) -
mCSQA: Multilingual Commonsense Reasoning Dataset with Unified Creation Strategy by Language Models and Humans
par: Sakai, Yusuke, et autres
Publié: (2024) -
Tonguescape: Exploring Language Models Understanding of Vowel Articulation
par: Sakajo, Haruki, et autres
Publié: (2025) -
HalluCitation Matters: Revealing the Impact of Hallucinated References with 300 Hallucinated Papers in ACL Conferences
par: Sakai, Yusuke, et autres
Publié: (2026)