What Is Missing in Multilingual Visual Reasoning and How to Fix It
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Song, Yueqi, Khanuja, Simran, Neubig, Graham |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Grounding Multilingual Multimodal LLMs With Cultural Knowledge
von: Nyandwi, Jean de Dieu, et al.
Veröffentlicht: (2025)
von: Nyandwi, Jean de Dieu, et al.
Veröffentlicht: (2025)
An image speaks a thousand words, but can everyone listen? On image transcreation for cultural relevance
von: Khanuja, Simran, et al.
Veröffentlicht: (2024)
von: Khanuja, Simran, et al.
Veröffentlicht: (2024)
Towards Automatic Evaluation for Image Transcreation
von: Khanuja, Simran, et al.
Veröffentlicht: (2024)
von: Khanuja, Simran, et al.
Veröffentlicht: (2024)
Pangea: A Fully Open Multilingual Multimodal LLM for 39 Languages
von: Yue, Xiang, et al.
Veröffentlicht: (2024)
von: Yue, Xiang, et al.
Veröffentlicht: (2024)
CAIRe: Cultural Attribution of Images by Retrieval-Augmented Evaluation
von: Yayavaram, Arnav, et al.
Veröffentlicht: (2025)
von: Yayavaram, Arnav, et al.
Veröffentlicht: (2025)
MERLIN: A Testbed for Multilingual Multimodal Entity Recognition and Linking
von: Ramamoorthy, Sathyanarayanan, et al.
Veröffentlicht: (2025)
von: Ramamoorthy, Sathyanarayanan, et al.
Veröffentlicht: (2025)
VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge
von: Song, Yueqi, et al.
Veröffentlicht: (2025)
von: Song, Yueqi, et al.
Veröffentlicht: (2025)
Beyond Browsing: API-Based Web Agents
von: Song, Yueqi, et al.
Veröffentlicht: (2024)
von: Song, Yueqi, et al.
Veröffentlicht: (2024)
Gained in Translation: Privileged Pairwise Judges Enhance Multilingual Reasoning
von: Sutawika, Lintang, et al.
Veröffentlicht: (2026)
von: Sutawika, Lintang, et al.
Veröffentlicht: (2026)
NaturalBench: Evaluating Vision-Language Models on Natural Adversarial Samples
von: Li, Baiqi, et al.
Veröffentlicht: (2024)
von: Li, Baiqi, et al.
Veröffentlicht: (2024)
On the Interplay of Pre-Training, Mid-Training, and RL on Reasoning Language Models
von: Zhang, Charlie, et al.
Veröffentlicht: (2025)
von: Zhang, Charlie, et al.
Veröffentlicht: (2025)
What Goes Into a LM Acceptability Judgment? Rethinking the Impact of Frequency and Length
von: Tjuatja, Lindia, et al.
Veröffentlicht: (2024)
von: Tjuatja, Lindia, et al.
Veröffentlicht: (2024)
Go-Browse: Training Web Agents with Structured Exploration
von: Gandhi, Apurva, et al.
Veröffentlicht: (2025)
von: Gandhi, Apurva, et al.
Veröffentlicht: (2025)
BehaviorBox: Automated Discovery of Fine-Grained Performance Differences Between Language Models
von: Tjuatja, Lindia, et al.
Veröffentlicht: (2025)
von: Tjuatja, Lindia, et al.
Veröffentlicht: (2025)
VisualWebBench: How Far Have Multimodal LLMs Evolved in Web Page Understanding and Grounding?
von: Liu, Junpeng, et al.
Veröffentlicht: (2024)
von: Liu, Junpeng, et al.
Veröffentlicht: (2024)
Massively Multilingual Joint Segmentation and Glossing
von: Ginn, Michael, et al.
Veröffentlicht: (2026)
von: Ginn, Michael, et al.
Veröffentlicht: (2026)
Steering LLMs for Culturally Localized Generation
von: Khanuja, Simran, et al.
Veröffentlicht: (2026)
von: Khanuja, Simran, et al.
Veröffentlicht: (2026)
Effective Strategies for Asynchronous Software Engineering Agents
von: Geng, Jiayi, et al.
Veröffentlicht: (2026)
von: Geng, Jiayi, et al.
Veröffentlicht: (2026)
GlossLM: A Massively Multilingual Corpus and Pretrained Model for Interlinear Glossed Text
von: Ginn, Michael, et al.
Veröffentlicht: (2024)
von: Ginn, Michael, et al.
Veröffentlicht: (2024)
Round-Trip Translation Reveals What Frontier Multilingual Benchmarks Miss
von: Skorobogat, Ronald, et al.
Veröffentlicht: (2026)
von: Skorobogat, Ronald, et al.
Veröffentlicht: (2026)
What Are Tools Anyway? A Survey from the Language Model Perspective
von: Wang, Zhiruo, et al.
Veröffentlicht: (2024)
von: Wang, Zhiruo, et al.
Veröffentlicht: (2024)
Demystifying Long Chain-of-Thought Reasoning in LLMs
von: Yeo, Edward, et al.
Veröffentlicht: (2025)
von: Yeo, Edward, et al.
Veröffentlicht: (2025)
An Incomplete Loop: Instruction Inference, Instruction Following, and In-context Learning in Language Models
von: Liu, Emmy, et al.
Veröffentlicht: (2024)
von: Liu, Emmy, et al.
Veröffentlicht: (2024)
Midtraining Bridges Pretraining and Posttraining Distributions
von: Liu, Emmy, et al.
Veröffentlicht: (2025)
von: Liu, Emmy, et al.
Veröffentlicht: (2025)
Solving NLP Problems through Human-System Collaboration: A Discussion-based Approach
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2023)
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2023)
Oolong: Evaluating Long Context Reasoning and Aggregation Capabilities
von: Bertsch, Amanda, et al.
Veröffentlicht: (2025)
von: Bertsch, Amanda, et al.
Veröffentlicht: (2025)
M-Prometheus: A Suite of Open Multilingual LLM Judges
von: Pombal, José, et al.
Veröffentlicht: (2025)
von: Pombal, José, et al.
Veröffentlicht: (2025)
What do Language Models Learn and When? The Implicit Curriculum Hypothesis
von: Liu, Emmy, et al.
Veröffentlicht: (2026)
von: Liu, Emmy, et al.
Veröffentlicht: (2026)
Reasoning Is Not All You Need: Examining LLMs for Multi-Turn Mental Health Conversations
von: Chandra, Mohit, et al.
Veröffentlicht: (2025)
von: Chandra, Mohit, et al.
Veröffentlicht: (2025)
RAGGED: Towards Informed Design of Scalable and Stable RAG Systems
von: Hsia, Jennifer, et al.
Veröffentlicht: (2024)
von: Hsia, Jennifer, et al.
Veröffentlicht: (2024)
ClusterFusion: Hybrid Clustering with Embedding Guidance and LLM Adaptation
von: Xu, Yiming, et al.
Veröffentlicht: (2025)
von: Xu, Yiming, et al.
Veröffentlicht: (2025)
What Makes Good Multilingual Reasoning? Disentangling Reasoning Traces with Measurable Features
von: Ki, Dayeon, et al.
Veröffentlicht: (2026)
von: Ki, Dayeon, et al.
Veröffentlicht: (2026)
Harnessing Webpage UIs for Text-Rich Visual Understanding
von: Liu, Junpeng, et al.
Veröffentlicht: (2024)
von: Liu, Junpeng, et al.
Veröffentlicht: (2024)
Multilingual Reasoning Gym: Multilingual Scaling of Procedural Reasoning Environments
von: Dobler, Konstantin, et al.
Veröffentlicht: (2026)
von: Dobler, Konstantin, et al.
Veröffentlicht: (2026)
Training Versatile Coding Agents in Synthetic Environments
von: Zhu, Yiqi, et al.
Veröffentlicht: (2025)
von: Zhu, Yiqi, et al.
Veröffentlicht: (2025)
Language Matters: How Do Multilingual Input and Reasoning Paths Affect Large Reasoning Models?
von: Tam, Zhi Rui, et al.
Veröffentlicht: (2025)
von: Tam, Zhi Rui, et al.
Veröffentlicht: (2025)
Agent Workflow Memory
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2024)
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2024)
Inducing Programmatic Skills for Agentic Tasks
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2025)
von: Wang, Zora Zhiruo, et al.
Veröffentlicht: (2025)
What Really Counts? Examining Step and Token Level Attribution in Multilingual CoT Reasoning
von: Ferrao, Jeremias, et al.
Veröffentlicht: (2025)
von: Ferrao, Jeremias, et al.
Veröffentlicht: (2025)
Can Large Language Models be Trusted for Evaluation? Scalable Meta-Evaluation of LLMs as Evaluators via Agent Debate
von: Chern, Steffi, et al.
Veröffentlicht: (2024)
von: Chern, Steffi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Grounding Multilingual Multimodal LLMs With Cultural Knowledge
von: Nyandwi, Jean de Dieu, et al.
Veröffentlicht: (2025) -
An image speaks a thousand words, but can everyone listen? On image transcreation for cultural relevance
von: Khanuja, Simran, et al.
Veröffentlicht: (2024) -
Towards Automatic Evaluation for Image Transcreation
von: Khanuja, Simran, et al.
Veröffentlicht: (2024) -
Pangea: A Fully Open Multilingual Multimodal LLM for 39 Languages
von: Yue, Xiang, et al.
Veröffentlicht: (2024) -
CAIRe: Cultural Attribution of Images by Retrieval-Augmented Evaluation
von: Yayavaram, Arnav, et al.
Veröffentlicht: (2025)