VMMU: A Vietnamese Multitask Multimodal Understanding and Reasoning Benchmark
Fuente:
arXiv
Saved in:
| Main Authors: | Dang, Vy Tuong, Vo, An, Villa-Cueva, Emilio, Tau, Quang, Dm, Duc, Solorio, Thamar, Kim, Daeyoung |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Vision Language Models are Biased
by: Vo, An, et al.
Published: (2025)
by: Vo, An, et al.
Published: (2025)
FIBER: A Differentially Private Optimizer with Filter-Aware Innovation Bias Correction
by: Dm, Duc, et al.
Published: (2026)
by: Dm, Duc, et al.
Published: (2026)
B-score: Detecting biases in large language models using response history
by: Vo, An, et al.
Published: (2025)
by: Vo, An, et al.
Published: (2025)
Adaptive Cross-lingual Text Classification through In-Context One-Shot Demonstrations
by: Villa-Cueva, Emilio, et al.
Published: (2024)
by: Villa-Cueva, Emilio, et al.
Published: (2024)
ROAST: Review-level Opinion Aspect Sentiment Target Joint Detection for ABSA
by: Chebolu, Siva Uday Sampreeth, et al.
Published: (2024)
by: Chebolu, Siva Uday Sampreeth, et al.
Published: (2024)
Vintern-1B: An Efficient Multimodal Large Language Model for Vietnamese
by: Doan, Khang T., et al.
Published: (2024)
by: Doan, Khang T., et al.
Published: (2024)
BERT-based model for Vietnamese Fact Verification Dataset
by: Tran, Bao, et al.
Published: (2025)
by: Tran, Bao, et al.
Published: (2025)
MOMENTS: A Comprehensive Multimodal Benchmark for Theory of Mind
by: Villa-Cueva, Emilio, et al.
Published: (2025)
by: Villa-Cueva, Emilio, et al.
Published: (2025)
Low-Dimensional Structure in the Space of Language Representations is Reflected in Brain Responses
by: Antonello, Richard, et al.
Published: (2021)
by: Antonello, Richard, et al.
Published: (2021)
Enhancing NER Performance in Low-Resource Pakistani Languages using Cross-Lingual Data Augmentation
by: Ehsan, Toqeer, et al.
Published: (2025)
by: Ehsan, Toqeer, et al.
Published: (2025)
MultiMed: Massively Multimodal and Multitask Medical Understanding
by: Mo, Shentong, et al.
Published: (2024)
by: Mo, Shentong, et al.
Published: (2024)
Benchmarking and Understanding Compositional Relational Reasoning of LLMs
by: Ni, Ruikang, et al.
Published: (2024)
by: Ni, Ruikang, et al.
Published: (2024)
DetoxBench: Benchmarking Large Language Models for Multitask Fraud & Abuse Detection
by: Chakraborty, Joymallya, et al.
Published: (2024)
by: Chakraborty, Joymallya, et al.
Published: (2024)
Reasoning Beyond Literal: Cross-style Multimodal Reasoning for Figurative Language Understanding
by: Cheshmi, Seyyed Saeid, et al.
Published: (2026)
by: Cheshmi, Seyyed Saeid, et al.
Published: (2026)
Do MLLMs Capture How Interfaces Guide User Behavior? A Benchmark for Multimodal UI/UX Design Understanding
by: Jeon, Jaehyun, et al.
Published: (2025)
by: Jeon, Jaehyun, et al.
Published: (2025)
Sentiment Reasoning for Healthcare
by: Nguyen, Khai-Nguyen, et al.
Published: (2024)
by: Nguyen, Khai-Nguyen, et al.
Published: (2024)
MUStReason: A Benchmark for Diagnosing Pragmatic Reasoning in Video-LMs for Multimodal Sarcasm Detection
by: Saha, Anisha, et al.
Published: (2025)
by: Saha, Anisha, et al.
Published: (2025)
RelUNet: Relative Channel Fusion U-Net for Multichannel Speech Enhancement
by: Aldarmaki, Ibrahim, et al.
Published: (2024)
by: Aldarmaki, Ibrahim, et al.
Published: (2024)
SUPERChem: A Multimodal Reasoning Benchmark in Chemistry
by: Zhao, Zehua, et al.
Published: (2025)
by: Zhao, Zehua, et al.
Published: (2025)
Context-aware Adversarial Attack on Named Entity Recognition
by: Chen, Shuguang, et al.
Published: (2023)
by: Chen, Shuguang, et al.
Published: (2023)
LaVy: Vietnamese Multimodal Large Language Model
by: Tran, Chi, et al.
Published: (2024)
by: Tran, Chi, et al.
Published: (2024)
LUME: LLM Unlearning with Multitask Evaluations
by: Ramakrishna, Anil, et al.
Published: (2025)
by: Ramakrishna, Anil, et al.
Published: (2025)
A Modular Multitask Reasoning Framework Integrating Spatio-temporal Models and LLMs
by: Hettige, Kethmi Hirushini, et al.
Published: (2025)
by: Hettige, Kethmi Hirushini, et al.
Published: (2025)
Dental-TriageBench: Benchmarking Multimodal Reasoning for Hierarchical Dental Triage
by: He, Ziyi, et al.
Published: (2026)
by: He, Ziyi, et al.
Published: (2026)
VietMix: A Naturally-Occurring Parallel Corpus and Augmentation Framework for Vietnamese-English Code-Mixed Machine Translation
by: Tran, Hieu, et al.
Published: (2025)
by: Tran, Hieu, et al.
Published: (2025)
IL-TUR: Benchmark for Indian Legal Text Understanding and Reasoning
by: Joshi, Abhinav, et al.
Published: (2024)
by: Joshi, Abhinav, et al.
Published: (2024)
MATEO: A Multimodal Benchmark for Temporal Reasoning and Planning in LVLMs
by: Roccabruna, Gabriel, et al.
Published: (2026)
by: Roccabruna, Gabriel, et al.
Published: (2026)
Beyond Understanding: Evaluating the Pragmatic Gap in LLMs' Cultural Processing of Figurative Language
by: Attia, Mena, et al.
Published: (2025)
by: Attia, Mena, et al.
Published: (2025)
OWLViz: An Open-World Benchmark for Visual Question Answering
by: Nguyen, Thuy, et al.
Published: (2025)
by: Nguyen, Thuy, et al.
Published: (2025)
Real-time Speech Summarization for Medical Conversations
by: Le-Duc, Khai, et al.
Published: (2024)
by: Le-Duc, Khai, et al.
Published: (2024)
Leveraging Large Language Models for Suicide Detection on Social Media with Limited Labels
by: Nguyen, Vy, et al.
Published: (2024)
by: Nguyen, Vy, et al.
Published: (2024)
Reinforcement Learning for Reasoning in Small LLMs: What Works and What Doesn't
by: Dang, Quy-Anh, et al.
Published: (2025)
by: Dang, Quy-Anh, et al.
Published: (2025)
Interpreting Themes from Educational Stories
by: Zhang, Yigeng, et al.
Published: (2024)
by: Zhang, Yigeng, et al.
Published: (2024)
Progressive Multi-granular Alignments for Grounded Reasoning in Large Vision-Language Models
by: Le, Quang-Hung, et al.
Published: (2024)
by: Le, Quang-Hung, et al.
Published: (2024)
Efficient Second-Order Neural Network Optimization via Adaptive Trust Region Methods
by: Vo, James
Published: (2024)
by: Vo, James
Published: (2024)
Benchmarking the Medical Understanding and Reasoning of Large Language Models in Arabic Healthcare Tasks
by: AlDahoul, Nouar, et al.
Published: (2025)
by: AlDahoul, Nouar, et al.
Published: (2025)
ECG-Reasoning-Benchmark: A Benchmark for Evaluating Clinical Reasoning Capabilities in ECG Interpretation
by: Oh, Jungwoo, et al.
Published: (2026)
by: Oh, Jungwoo, et al.
Published: (2026)
Adaptive Two-Phase Finetuning LLMs for Japanese Legal Text Retrieval
by: Trung, Quang Hoang, et al.
Published: (2024)
by: Trung, Quang Hoang, et al.
Published: (2024)
MMTU: A Massive Multi-Task Table Understanding and Reasoning Benchmark
by: Xing, Junjie, et al.
Published: (2025)
by: Xing, Junjie, et al.
Published: (2025)
Nested Named-Entity Recognition on Vietnamese COVID-19: Dataset and Experiments
by: Lê, Ngoc C., et al.
Published: (2025)
by: Lê, Ngoc C., et al.
Published: (2025)
Similar Items
-
Vision Language Models are Biased
by: Vo, An, et al.
Published: (2025) -
FIBER: A Differentially Private Optimizer with Filter-Aware Innovation Bias Correction
by: Dm, Duc, et al.
Published: (2026) -
B-score: Detecting biases in large language models using response history
by: Vo, An, et al.
Published: (2025) -
Adaptive Cross-lingual Text Classification through In-Context One-Shot Demonstrations
by: Villa-Cueva, Emilio, et al.
Published: (2024) -
ROAST: Review-level Opinion Aspect Sentiment Target Joint Detection for ABSA
by: Chebolu, Siva Uday Sampreeth, et al.
Published: (2024)