Grounding Multilingual Multimodal LLMs With Cultural Knowledge
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nyandwi, Jean de Dieu, Song, Yueqi, Khanuja, Simran, Neubig, Graham |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
What Is Missing in Multilingual Visual Reasoning and How to Fix It
von: Song, Yueqi, et al.
Veröffentlicht: (2024)
von: Song, Yueqi, et al.
Veröffentlicht: (2024)
Pangea: A Fully Open Multilingual Multimodal LLM for 39 Languages
von: Yue, Xiang, et al.
Veröffentlicht: (2024)
von: Yue, Xiang, et al.
Veröffentlicht: (2024)
An image speaks a thousand words, but can everyone listen? On image transcreation for cultural relevance
von: Khanuja, Simran, et al.
Veröffentlicht: (2024)
von: Khanuja, Simran, et al.
Veröffentlicht: (2024)
CAIRe: Cultural Attribution of Images by Retrieval-Augmented Evaluation
von: Yayavaram, Arnav, et al.
Veröffentlicht: (2025)
von: Yayavaram, Arnav, et al.
Veröffentlicht: (2025)
NaturalBench: Evaluating Vision-Language Models on Natural Adversarial Samples
von: Li, Baiqi, et al.
Veröffentlicht: (2024)
von: Li, Baiqi, et al.
Veröffentlicht: (2024)
Towards Automatic Evaluation for Image Transcreation
von: Khanuja, Simran, et al.
Veröffentlicht: (2024)
von: Khanuja, Simran, et al.
Veröffentlicht: (2024)
MERLIN: A Testbed for Multilingual Multimodal Entity Recognition and Linking
von: Ramamoorthy, Sathyanarayanan, et al.
Veröffentlicht: (2025)
von: Ramamoorthy, Sathyanarayanan, et al.
Veröffentlicht: (2025)
Gained in Translation: Privileged Pairwise Judges Enhance Multilingual Reasoning
von: Sutawika, Lintang, et al.
Veröffentlicht: (2026)
von: Sutawika, Lintang, et al.
Veröffentlicht: (2026)
Everybody Prune Now: Structured Pruning of LLMs with only Forward Passes
von: Kolawole, Steven, et al.
Veröffentlicht: (2024)
von: Kolawole, Steven, et al.
Veröffentlicht: (2024)
Demystifying Long Chain-of-Thought Reasoning in LLMs
von: Yeo, Edward, et al.
Veröffentlicht: (2025)
von: Yeo, Edward, et al.
Veröffentlicht: (2025)
Do LLMs Understand Your Translations? Evaluating Paragraph-level MT with Question Answering
von: Fernandes, Patrick, et al.
Veröffentlicht: (2025)
von: Fernandes, Patrick, et al.
Veröffentlicht: (2025)
VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge
von: Song, Yueqi, et al.
Veröffentlicht: (2025)
von: Song, Yueqi, et al.
Veröffentlicht: (2025)
Synthetic Multimodal Question Generation
von: Wu, Ian, et al.
Veröffentlicht: (2024)
von: Wu, Ian, et al.
Veröffentlicht: (2024)
Steering LLMs for Culturally Localized Generation
von: Khanuja, Simran, et al.
Veröffentlicht: (2026)
von: Khanuja, Simran, et al.
Veröffentlicht: (2026)
Multilingual Amnesia: On the Transferability of Unlearning in Multilingual LLMs
von: Farashah, Alireza Dehghanpour, et al.
Veröffentlicht: (2026)
von: Farashah, Alireza Dehghanpour, et al.
Veröffentlicht: (2026)
Repetition Improves Language Model Embeddings
von: Springer, Jacob Mitchell, et al.
Veröffentlicht: (2024)
von: Springer, Jacob Mitchell, et al.
Veröffentlicht: (2024)
On the Calibration of Multilingual Question Answering LLMs
von: Yang, Yahan, et al.
Veröffentlicht: (2023)
von: Yang, Yahan, et al.
Veröffentlicht: (2023)
FoundationalASSIST: An Educational Dataset for Foundational Knowledge Tracing and Pedagogical Grounding of LLMs
von: Worden, Eamon, et al.
Veröffentlicht: (2026)
von: Worden, Eamon, et al.
Veröffentlicht: (2026)
Learn and Unlearn: Addressing Misinformation in Multilingual LLMs
von: Lu, Taiming, et al.
Veröffentlicht: (2024)
von: Lu, Taiming, et al.
Veröffentlicht: (2024)
How Does Quantization Affect Multilingual LLMs?
von: Marchisio, Kelly, et al.
Veröffentlicht: (2024)
von: Marchisio, Kelly, et al.
Veröffentlicht: (2024)
CultureGuard: Towards Culturally-Aware Dataset and Guard Model for Multilingual Safety Applications
von: Joshi, Raviraj, et al.
Veröffentlicht: (2025)
von: Joshi, Raviraj, et al.
Veröffentlicht: (2025)
Beyond Browsing: API-Based Web Agents
von: Song, Yueqi, et al.
Veröffentlicht: (2024)
von: Song, Yueqi, et al.
Veröffentlicht: (2024)
Instruction-tuned Language Models are Better Knowledge Learners
von: Jiang, Zhengbao, et al.
Veröffentlicht: (2024)
von: Jiang, Zhengbao, et al.
Veröffentlicht: (2024)
Crosslingual Capabilities and Knowledge Barriers in Multilingual Large Language Models
von: Chua, Lynn, et al.
Veröffentlicht: (2024)
von: Chua, Lynn, et al.
Veröffentlicht: (2024)
Affective Multimodal Agents with Proactive Knowledge Grounding for Emotionally Aligned Marketing Dialogue
von: Yu, Lin, et al.
Veröffentlicht: (2025)
von: Yu, Lin, et al.
Veröffentlicht: (2025)
Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis
von: Bhowmik, Shimanto, et al.
Veröffentlicht: (2025)
von: Bhowmik, Shimanto, et al.
Veröffentlicht: (2025)
RTP-LX: Can LLMs Evaluate Toxicity in Multilingual Scenarios?
von: de Wynter, Adrian, et al.
Veröffentlicht: (2024)
von: de Wynter, Adrian, et al.
Veröffentlicht: (2024)
Do Multilingual LLMs Think In English?
von: Schut, Lisa, et al.
Veröffentlicht: (2025)
von: Schut, Lisa, et al.
Veröffentlicht: (2025)
From Decoding to Meta-Generation: Inference-time Algorithms for Large Language Models
von: Welleck, Sean, et al.
Veröffentlicht: (2024)
von: Welleck, Sean, et al.
Veröffentlicht: (2024)
Merging Methods for Multilingual Knowledge Editing for Large Language Models: An Empirical Odyssey
von: Lee, Kunil, et al.
Veröffentlicht: (2026)
von: Lee, Kunil, et al.
Veröffentlicht: (2026)
URIEL+: Enhancing Linguistic Inclusion and Usability in a Typological and Multilingual Knowledge Base
von: Khan, Aditya, et al.
Veröffentlicht: (2024)
von: Khan, Aditya, et al.
Veröffentlicht: (2024)
GSR-BENCH: A Benchmark for Grounded Spatial Reasoning Evaluation via Multimodal LLMs
von: Rajabi, Navid, et al.
Veröffentlicht: (2024)
von: Rajabi, Navid, et al.
Veröffentlicht: (2024)
Building Multilingual Datasets for Predicting Mental Health Severity through LLMs: Prospects and Challenges
von: Skianis, Konstantinos, et al.
Veröffentlicht: (2024)
von: Skianis, Konstantinos, et al.
Veröffentlicht: (2024)
Toward Cultural Interpretability: A Linguistic Anthropological Framework for Describing and Evaluating Large Language Models (LLMs)
von: Jones, Graham M., et al.
Veröffentlicht: (2024)
von: Jones, Graham M., et al.
Veröffentlicht: (2024)
MUCAR: Benchmarking Multilingual Cross-Modal Ambiguity Resolution for Multimodal Large Language Models
von: Wang, Xiaolong, et al.
Veröffentlicht: (2025)
von: Wang, Xiaolong, et al.
Veröffentlicht: (2025)
Social Genome: Grounded Social Reasoning Abilities of Multimodal Models
von: Mathur, Leena, et al.
Veröffentlicht: (2025)
von: Mathur, Leena, et al.
Veröffentlicht: (2025)
MateInfoUB: A Real-World Benchmark for Testing LLMs in Competitive, Multilingual, and Multimodal Educational Tasks
von: Marius, Dumitran Adrian, et al.
Veröffentlicht: (2025)
von: Marius, Dumitran Adrian, et al.
Veröffentlicht: (2025)
VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks
von: Koh, Jing Yu, et al.
Veröffentlicht: (2024)
von: Koh, Jing Yu, et al.
Veröffentlicht: (2024)
Beyond Early-Token Bias: Model-Specific and Language-Specific Position Effects in Multilingual LLMs
von: Menschikov, Mikhail, et al.
Veröffentlicht: (2025)
von: Menschikov, Mikhail, et al.
Veröffentlicht: (2025)
Better To Ask in English? Evaluating Factual Accuracy of Multilingual LLMs in English and Low-Resource Languages
von: Rohera, Pritika, et al.
Veröffentlicht: (2025)
von: Rohera, Pritika, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
What Is Missing in Multilingual Visual Reasoning and How to Fix It
von: Song, Yueqi, et al.
Veröffentlicht: (2024) -
Pangea: A Fully Open Multilingual Multimodal LLM for 39 Languages
von: Yue, Xiang, et al.
Veröffentlicht: (2024) -
An image speaks a thousand words, but can everyone listen? On image transcreation for cultural relevance
von: Khanuja, Simran, et al.
Veröffentlicht: (2024) -
CAIRe: Cultural Attribution of Images by Retrieval-Augmented Evaluation
von: Yayavaram, Arnav, et al.
Veröffentlicht: (2025) -
NaturalBench: Evaluating Vision-Language Models on Natural Adversarial Samples
von: Li, Baiqi, et al.
Veröffentlicht: (2024)