Evaluation of Cultural Competence of Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Yadav, Srishti, Tilton, Lauren, Antoniak, Maria, Arnold, Taylor, Li, Jiaang, Pawar, Siddhesh Milind, Karamolegkou, Antonia, Frank, Stella, An, Zhaochong, Rostamzadeh, Negar, Hershcovich, Daniel, Belongie, Serge, Shutova, Ekaterina |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond Words: Exploring Cultural Value Sensitivity in Multimodal Models
by: Yadav, Srishti, et al.
Published: (2025)
by: Yadav, Srishti, et al.
Published: (2025)
Exploring Visual Culture Awareness in GPT-4V: A Comprehensive Probing
by: Cao, Yong, et al.
Published: (2024)
by: Cao, Yong, et al.
Published: (2024)
Multi-Modal Framing Analysis of News
by: Arora, Arnav, et al.
Published: (2025)
by: Arora, Arnav, et al.
Published: (2025)
Cultural Compass: Predicting Transfer Learning Success in Offensive Language Detection with Cultural Features
by: Zhou, Li, et al.
Published: (2023)
by: Zhou, Li, et al.
Published: (2023)
Explainable Search and Discovery of Visual Cultural Heritage Collections with Multimodal Large Language Models
by: Arnold, Taylor, et al.
Published: (2024)
by: Arnold, Taylor, et al.
Published: (2024)
Cross-modal Information Flow in Multimodal Large Language Models
by: Zhang, Zhi, et al.
Published: (2024)
by: Zhang, Zhi, et al.
Published: (2024)
Automated Image Color Mapping for a Historic Photographic Collection
by: Arnold, Taylor, et al.
Published: (2024)
by: Arnold, Taylor, et al.
Published: (2024)
Vision-Language Models under Cultural and Inclusive Considerations
by: Karamolegkou, Antonia, et al.
Published: (2024)
by: Karamolegkou, Antonia, et al.
Published: (2024)
RAVENEA: A Benchmark for Multimodal Retrieval-Augmented Visual Culture Understanding
by: Li, Jiaang, et al.
Published: (2025)
by: Li, Jiaang, et al.
Published: (2025)
Revealing Fine-Grained Values and Opinions in Large Language Models
by: Wright, Dustin, et al.
Published: (2024)
by: Wright, Dustin, et al.
Published: (2024)
Self-Alignment: Improving Alignment of Cultural Values in LLMs via In-Context Learning
by: Choenni, Rochelle, et al.
Published: (2024)
by: Choenni, Rochelle, et al.
Published: (2024)
Survey of Cultural Awareness in Language Models: Text and Beyond
by: Pawar, Siddhesh, et al.
Published: (2024)
by: Pawar, Siddhesh, et al.
Published: (2024)
The Echoes of Multilinguality: Tracing Cultural Value Shifts during LM Fine-tuning
by: Choenni, Rochelle, et al.
Published: (2024)
by: Choenni, Rochelle, et al.
Published: (2024)
Thinking in Frames: How Visual Context and Test-Time Scaling Empower Video Reasoning
by: Li, Chengzu, et al.
Published: (2026)
by: Li, Chengzu, et al.
Published: (2026)
Not What, But How: A Communicative Audit of LLM Response Framing
by: Pawar, Siddhesh Milind, et al.
Published: (2026)
by: Pawar, Siddhesh Milind, et al.
Published: (2026)
EvalCards: A Framework for Standardized Evaluation Reporting
by: Dhar, Ruchira, et al.
Published: (2025)
by: Dhar, Ruchira, et al.
Published: (2025)
Generalized Few-shot 3D Point Cloud Segmentation with Vision-Language Model
by: An, Zhaochong, et al.
Published: (2025)
by: An, Zhaochong, et al.
Published: (2025)
Induction Heads as an Essential Mechanism for Pattern Matching in In-context Learning
by: Crosbie, Joy, et al.
Published: (2024)
by: Crosbie, Joy, et al.
Published: (2024)
Video Understanding: From Geometry and Semantics to Unified Models
by: An, Zhaochong, et al.
Published: (2026)
by: An, Zhaochong, et al.
Published: (2026)
Evaluating Multimodal Language Models as Visual Assistants for Visually Impaired Users
by: Karamolegkou, Antonia, et al.
Published: (2025)
by: Karamolegkou, Antonia, et al.
Published: (2025)
Revisiting the Perception-Distortion Trade-off with Spatial-Semantic Guided Super-Resolution
by: Wang, Dan, et al.
Published: (2026)
by: Wang, Dan, et al.
Published: (2026)
ChatMotion: A Multimodal Multi-Agent for Human Motion Analysis
by: Li, Lei, et al.
Published: (2025)
by: Li, Lei, et al.
Published: (2025)
Presumed Cultural Identity: How Names Shape LLM Responses
by: Pawar, Siddhesh, et al.
Published: (2025)
by: Pawar, Siddhesh, et al.
Published: (2025)
Multimodality Helps Few-shot 3D Point Cloud Semantic Segmentation
by: An, Zhaochong, et al.
Published: (2024)
by: An, Zhaochong, et al.
Published: (2024)
Rethinking Few-shot 3D Point Cloud Semantic Segmentation
by: An, Zhaochong, et al.
Published: (2024)
by: An, Zhaochong, et al.
Published: (2024)
A framework for annotating and modelling intentions behind metaphor use
by: Michelli, Gianluca, et al.
Published: (2024)
by: Michelli, Gianluca, et al.
Published: (2024)
How do languages influence each other? Studying cross-lingual data sharing during LM fine-tuning
by: Choenni, Rochelle, et al.
Published: (2023)
by: Choenni, Rochelle, et al.
Published: (2023)
Density Matrices for Metaphor Understanding
by: Owers, Jay, et al.
Published: (2024)
by: Owers, Jay, et al.
Published: (2024)
Position: Cracking the Code of Cascading Disparity Towards Marginalized Communities
by: Farnadi, Golnoosh, et al.
Published: (2024)
by: Farnadi, Golnoosh, et al.
Published: (2024)
Bridging Cultural Nuances in Dialogue Agents through Cultural Value Surveys
by: Cao, Yong, et al.
Published: (2024)
by: Cao, Yong, et al.
Published: (2024)
BiasGym: A Simple and Generalizable Framework for Analyzing and Removing Biases through Elicitation
by: Islam, Sekh Mainul, et al.
Published: (2025)
by: Islam, Sekh Mainul, et al.
Published: (2025)
VGGRPO: Towards World-Consistent Video Generation with 4D Latent Reward
by: An, Zhaochong, et al.
Published: (2026)
by: An, Zhaochong, et al.
Published: (2026)
Yesterday's News: Benchmarking Multi-Dimensional Out-of-Distribution Generalization of Misinformation Detection Models
by: Verhoeven, Ivo, et al.
Published: (2024)
by: Verhoeven, Ivo, et al.
Published: (2024)
Reading or Guessing? Visual Grounding Failures of Vision-Language Models for OCR in Ancient Greek Editions
by: Karamolegkou, Antonia, et al.
Published: (2026)
by: Karamolegkou, Antonia, et al.
Published: (2026)
Are LLMs classical or nonmonotonic reasoners? Lessons from generics
by: Leidinger, Alina, et al.
Published: (2024)
by: Leidinger, Alina, et al.
Published: (2024)
Learning New Tasks from a Few Examples with Soft-Label Prototypes
by: Singh, Avyav Kumar, et al.
Published: (2022)
by: Singh, Avyav Kumar, et al.
Published: (2022)
CK-Transformer: Commonsense Knowledge Enhanced Transformers for Referring Expression Comprehension
by: Zhang, Zhi, et al.
Published: (2023)
by: Zhang, Zhi, et al.
Published: (2023)
Provocations from the Humanities for Generative AI Research
by: Klein, Lauren, et al.
Published: (2025)
by: Klein, Lauren, et al.
Published: (2025)
Metaphor Understanding Challenge Dataset for LLMs
by: Tong, Xiaoyu, et al.
Published: (2024)
by: Tong, Xiaoyu, et al.
Published: (2024)
Research Borderlands: Analysing Writing Across Research Cultures
by: Bhatt, Shaily, et al.
Published: (2025)
by: Bhatt, Shaily, et al.
Published: (2025)
Similar Items
-
Beyond Words: Exploring Cultural Value Sensitivity in Multimodal Models
by: Yadav, Srishti, et al.
Published: (2025) -
Exploring Visual Culture Awareness in GPT-4V: A Comprehensive Probing
by: Cao, Yong, et al.
Published: (2024) -
Multi-Modal Framing Analysis of News
by: Arora, Arnav, et al.
Published: (2025) -
Cultural Compass: Predicting Transfer Learning Success in Offensive Language Detection with Cultural Features
by: Zhou, Li, et al.
Published: (2023) -
Explainable Search and Discovery of Visual Cultural Heritage Collections with Multimodal Large Language Models
by: Arnold, Taylor, et al.
Published: (2024)