CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Shudong, Jin, Yiqiao, Li, Cheng, Wong, Derek F., Wen, Qingsong, Sun, Lichao, Chen, Haipeng, Xie, Xing, Wang, Jindong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CulturePark: Boosting Cross-cultural Understanding in Large Language Models
di: Li, Cheng, et al.
Pubblicazione: (2024)
di: Li, Cheng, et al.
Pubblicazione: (2024)
CultureLLM: Incorporating Cultural Differences into Large Language Models
di: Li, Cheng, et al.
Pubblicazione: (2024)
di: Li, Cheng, et al.
Pubblicazione: (2024)
CompeteAI: Understanding the Competition Dynamics in Large Language Model-based Agents
di: Zhao, Qinlin, et al.
Pubblicazione: (2023)
di: Zhao, Qinlin, et al.
Pubblicazione: (2023)
Learning from "Silly" Questions Improves Large Language Models, But Only Slightly
di: Zhu, Tingyuan, et al.
Pubblicazione: (2024)
di: Zhu, Tingyuan, et al.
Pubblicazione: (2024)
Rice-VL: Evaluating Vision-Language Models for Cultural Understanding Across ASEAN Countries
di: Pranav, Tushar, et al.
Pubblicazione: (2025)
di: Pranav, Tushar, et al.
Pubblicazione: (2025)
Cultural Awareness in Vision-Language Models: A Cross-Country Exploration
di: Madasu, Avinash, et al.
Pubblicazione: (2025)
di: Madasu, Avinash, et al.
Pubblicazione: (2025)
Benchmarking Vision Language Models for Cultural Understanding
di: Nayak, Shravan, et al.
Pubblicazione: (2024)
di: Nayak, Shravan, et al.
Pubblicazione: (2024)
Time-VLM: Exploring Multimodal Vision-Language Models for Augmented Time Series Forecasting
di: Zhong, Siru, et al.
Pubblicazione: (2025)
di: Zhong, Siru, et al.
Pubblicazione: (2025)
MM-Soc: Benchmarking Multimodal Large Language Models in Social Media Platforms
di: Jin, Yiqiao, et al.
Pubblicazione: (2024)
di: Jin, Yiqiao, et al.
Pubblicazione: (2024)
CAReDiO: Cultural Alignment via Representativeness and Distinctiveness Guided Data Optimization
di: Yao, Jing, et al.
Pubblicazione: (2025)
di: Yao, Jing, et al.
Pubblicazione: (2025)
ARM2: Adaptive Reasoning Model with Vision Understanding and Executable Code
di: Xie, Jian, et al.
Pubblicazione: (2025)
di: Xie, Jian, et al.
Pubblicazione: (2025)
VULCA-Bench: A Multicultural Vision-Language Benchmark for Evaluating Cultural Understanding
di: Yu, Haorui, et al.
Pubblicazione: (2026)
di: Yu, Haorui, et al.
Pubblicazione: (2026)
BLEnD-Vis: Benchmarking Multimodal Cultural Understanding in Vision Language Models
di: Tan, Bryan Chen Zhengyu, et al.
Pubblicazione: (2025)
di: Tan, Bryan Chen Zhengyu, et al.
Pubblicazione: (2025)
SciEvo: A 2 Million, 30-Year Cross-disciplinary Dataset for Temporal Scientometric Analysis
di: Jin, Yiqiao, et al.
Pubblicazione: (2024)
di: Jin, Yiqiao, et al.
Pubblicazione: (2024)
HKCanto-Eval: A Benchmark for Evaluating Cantonese Language Understanding and Cultural Comprehension in LLMs
di: Cheng, Tsz Chung, et al.
Pubblicazione: (2025)
di: Cheng, Tsz Chung, et al.
Pubblicazione: (2025)
All Languages Matter: Evaluating LMMs on Culturally Diverse 100 Languages
di: Vayani, Ashmal, et al.
Pubblicazione: (2024)
di: Vayani, Ashmal, et al.
Pubblicazione: (2024)
Evaluation of Cultural Competence of Vision-Language Models
di: Yadav, Srishti, et al.
Pubblicazione: (2025)
di: Yadav, Srishti, et al.
Pubblicazione: (2025)
UniEDU: A Unified Language and Vision Assistant for Education Applications
di: Chu, Zhendong, et al.
Pubblicazione: (2025)
di: Chu, Zhendong, et al.
Pubblicazione: (2025)
Language over Content: Tracing Cultural Understanding in Multilingual Large Language Models
di: Cho, Seungho, et al.
Pubblicazione: (2025)
di: Cho, Seungho, et al.
Pubblicazione: (2025)
RoRA-VLM: Robust Retrieval-Augmented Vision Language Models
di: Qi, Jingyuan, et al.
Pubblicazione: (2024)
di: Qi, Jingyuan, et al.
Pubblicazione: (2024)
World in a Frame: Understanding Culture Mixing as a New Challenge for Vision-Language Models
di: Kim, Eunsu, et al.
Pubblicazione: (2025)
di: Kim, Eunsu, et al.
Pubblicazione: (2025)
Vision-Language Models under Cultural and Inclusive Considerations
di: Karamolegkou, Antonia, et al.
Pubblicazione: (2024)
di: Karamolegkou, Antonia, et al.
Pubblicazione: (2024)
Hanfu-Bench: A Multimodal Benchmark on Cross-Temporal Cultural Understanding and Transcreation
di: Zhou, Li, et al.
Pubblicazione: (2025)
di: Zhou, Li, et al.
Pubblicazione: (2025)
CDEval: A Benchmark for Measuring the Cultural Dimensions of Large Language Models
di: Wang, Yuhang, et al.
Pubblicazione: (2023)
di: Wang, Yuhang, et al.
Pubblicazione: (2023)
AlignVLM: Bridging Vision and Language Latent Spaces for Multimodal Document Understanding
di: Masry, Ahmed, et al.
Pubblicazione: (2025)
di: Masry, Ahmed, et al.
Pubblicazione: (2025)
Cross-Cultural Expert-Level Art Critique Evaluation with Vision-Language Models
di: Yu, Haorui, et al.
Pubblicazione: (2026)
di: Yu, Haorui, et al.
Pubblicazione: (2026)
Are Vision Language Models Cross-Cultural Theory of Mind Reasoners?
di: Nazi, Zabir Al, et al.
Pubblicazione: (2025)
di: Nazi, Zabir Al, et al.
Pubblicazione: (2025)
CultiVerse: Towards Cross-Cultural Understanding for Paintings with Large Language Model
di: Zhang, Wei, et al.
Pubblicazione: (2024)
di: Zhang, Wei, et al.
Pubblicazione: (2024)
How Culturally Aware are Vision-Language Models?
di: Burda-Lassen, Olena, et al.
Pubblicazione: (2024)
di: Burda-Lassen, Olena, et al.
Pubblicazione: (2024)
Evaluating Visual and Cultural Interpretation: The K-Viscuit Benchmark with Human-VLM Collaboration
di: Park, ChaeHun, et al.
Pubblicazione: (2024)
di: Park, ChaeHun, et al.
Pubblicazione: (2024)
OViP: Online Vision-Language Preference Learning for VLM Hallucination
di: Liu, Shujun, et al.
Pubblicazione: (2025)
di: Liu, Shujun, et al.
Pubblicazione: (2025)
Is Your Model Really A Good Math Reasoner? Evaluating Mathematical Reasoning with Checklist
di: Zhou, Zihao, et al.
Pubblicazione: (2024)
di: Zhou, Zihao, et al.
Pubblicazione: (2024)
LLM4Brain: Training a Large Language Model for Brain Video Understanding
di: Zheng, Ruizhe, et al.
Pubblicazione: (2024)
di: Zheng, Ruizhe, et al.
Pubblicazione: (2024)
On the Cultural Anachronism and Temporal Reasoning in Vision Language Models
di: Ranjan, Mukul, et al.
Pubblicazione: (2026)
di: Ranjan, Mukul, et al.
Pubblicazione: (2026)
Global PIQA: Evaluating Commonsense Reasoning Across 100+ Languages and Cultures
di: Chang, Tyler A., et al.
Pubblicazione: (2025)
di: Chang, Tyler A., et al.
Pubblicazione: (2025)
RelationVLM: Making Large Vision-Language Models Understand Visual Relations
di: Huang, Zhipeng, et al.
Pubblicazione: (2024)
di: Huang, Zhipeng, et al.
Pubblicazione: (2024)
ScVLM: Enhancing Vision-Language Model for Safety-Critical Event Understanding
di: Shi, Liang, et al.
Pubblicazione: (2024)
di: Shi, Liang, et al.
Pubblicazione: (2024)
Do LLMs Understand Wine Descriptors Across Cultures? A Benchmark for Cultural Adaptations of Wine Reviews
di: Zou, Chenye, et al.
Pubblicazione: (2025)
di: Zou, Chenye, et al.
Pubblicazione: (2025)
RE-VLM: Event-Augmented Vision-Language Model for Scene Understanding
di: Liu, Hanqing, et al.
Pubblicazione: (2026)
di: Liu, Hanqing, et al.
Pubblicazione: (2026)
Not All Countries Celebrate Thanksgiving: On the Cultural Dominance in Large Language Models
di: Wang, Wenxuan, et al.
Pubblicazione: (2023)
di: Wang, Wenxuan, et al.
Pubblicazione: (2023)
Documenti analoghi
-
CulturePark: Boosting Cross-cultural Understanding in Large Language Models
di: Li, Cheng, et al.
Pubblicazione: (2024) -
CultureLLM: Incorporating Cultural Differences into Large Language Models
di: Li, Cheng, et al.
Pubblicazione: (2024) -
CompeteAI: Understanding the Competition Dynamics in Large Language Model-based Agents
di: Zhao, Qinlin, et al.
Pubblicazione: (2023) -
Learning from "Silly" Questions Improves Large Language Models, But Only Slightly
di: Zhu, Tingyuan, et al.
Pubblicazione: (2024) -
Rice-VL: Evaluating Vision-Language Models for Cultural Understanding Across ASEAN Countries
di: Pranav, Tushar, et al.
Pubblicazione: (2025)