ViInfographicVQA: A Benchmark for Single and Multi-image Visual Question Answering on Vietnamese Infographics
Fuente:
arXiv
Saved in:
| Main Authors: | Van-Dinh, Tue-Thu, Tran, Hoang-Duy, Duong, Truong-Binh, Pham, Mai-Hanh, Le-Nguyen, Binh-Nam, Nguyen, Quoc-Thai |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ViCocktail: Automated Multi-Modal Data Collection for Vietnamese Audio-Visual Speech Recognition
by: Nguyen, Thai-Binh, et al.
Published: (2025)
by: Nguyen, Thai-Binh, et al.
Published: (2025)
AutoViVQA: A Large-Scale Automatically Constructed Dataset for Vietnamese Visual Question Answering
by: Tuong, Nguyen Anh, et al.
Published: (2026)
by: Tuong, Nguyen Anh, et al.
Published: (2026)
ViFactCheck: A New Benchmark Dataset and Methods for Multi-domain News Fact-Checking in Vietnamese
by: Hoa, Tran Thai, et al.
Published: (2024)
by: Hoa, Tran Thai, et al.
Published: (2024)
SemViQA: A Semantic Question Answering System for Vietnamese Information Fact-Checking
by: Tran, Dien X., et al.
Published: (2025)
by: Tran, Dien X., et al.
Published: (2025)
Human-Guided Reasoning with Large Language Models for Vietnamese Speech Emotion Recognition
by: Nguyen, Truc, et al.
Published: (2026)
by: Nguyen, Truc, et al.
Published: (2026)
ViOCRVQA: Novel Benchmark Dataset and Vision Reader for Visual Question Answering by Understanding Vietnamese Text in Images
by: Pham, Huy Quang, et al.
Published: (2024)
by: Pham, Huy Quang, et al.
Published: (2024)
ViTextVQA: A Large-Scale Visual Question Answering Dataset and a Novel Multimodal Feature Fusion Method for Vietnamese Text Comprehension in Images
by: Van Nguyen, Quan, et al.
Published: (2024)
by: Van Nguyen, Quan, et al.
Published: (2024)
Exploring the Application of Visual Question Answering (VQA) for Classroom Activity Monitoring
by: Vu, Sinh Trong, et al.
Published: (2025)
by: Vu, Sinh Trong, et al.
Published: (2025)
ViConBERT: Context-Gloss Aligned Vietnamese Word Embedding for Polysemous and Sense-Aware Representations
by: Huynh, Khang T., et al.
Published: (2025)
by: Huynh, Khang T., et al.
Published: (2025)
DSC2025 -- ViHallu Challenge: Detecting Hallucination in Vietnamese LLMs
by: Nguyen, Anh Thi-Hoang, et al.
Published: (2026)
by: Nguyen, Anh Thi-Hoang, et al.
Published: (2026)
Towards Signboard-Oriented Visual Question Answering: ViSignVQA Dataset, Method and Benchmark
by: Nguyen, Hieu Minh, et al.
Published: (2025)
by: Nguyen, Hieu Minh, et al.
Published: (2025)
ViX-Ray: A Vietnamese Chest X-Ray Dataset for Vision-Language Models
by: Nguyen, Duy Vu Minh, et al.
Published: (2026)
by: Nguyen, Duy Vu Minh, et al.
Published: (2026)
EFL teachers’ perceptions of professional development activities and their effects in a non-anglosphere context
by: Duy Binh Nguyen
Published: (2022)
by: Duy Binh Nguyen
Published: (2022)
ViGoEmotions: A Benchmark Dataset For Fine-grained Emotion Detection on Vietnamese Texts
by: Tran, Hung Quang, et al.
Published: (2026)
by: Tran, Hung Quang, et al.
Published: (2026)
Geomorphological characteristics of Nha Trang bay and adjacent area
by: Tran, Van Binh, et al.
Published: (2015)
by: Tran, Van Binh, et al.
Published: (2015)
Vietnamese Legal Information Retrieval in Question-Answering System
by: Ba, Thiem Nguyen, et al.
Published: (2024)
by: Ba, Thiem Nguyen, et al.
Published: (2024)
InfoChartQA: A Benchmark for Multimodal Question Answering on Infographic Charts
by: Xie, Tianchi, et al.
Published: (2025)
by: Xie, Tianchi, et al.
Published: (2025)
Enhancing Vietnamese VQA through Curriculum Learning on Raw and Augmented Text Representations
by: Nguyen, Khoi Anh, et al.
Published: (2025)
by: Nguyen, Khoi Anh, et al.
Published: (2025)
R2GQA: Retriever-Reader-Generator Question Answering System to Support Students Understanding Legal Regulations in Higher Education
by: Do, Phuc-Tinh Pham, et al.
Published: (2024)
by: Do, Phuc-Tinh Pham, et al.
Published: (2024)
SilVar: Speech Driven Multimodal Model for Reasoning Visual Question Answering and Object Localization
by: Pham, Tan-Hanh, et al.
Published: (2024)
by: Pham, Tan-Hanh, et al.
Published: (2024)
ViCLIP-OT: The First Foundation Vision-Language Model for Vietnamese Image-Text Retrieval with Optimal Transport
by: Tran, Quoc-Khang, et al.
Published: (2026)
by: Tran, Quoc-Khang, et al.
Published: (2026)
Structure- and Stability-Preserving Learning of Port-Hamiltonian Systems
by: Nguyen, Binh, et al.
Published: (2026)
by: Nguyen, Binh, et al.
Published: (2026)
Whisper based Cross-Lingual Phoneme Recognition between Vietnamese and English
by: Minh, Nguyen Huu Nhat, et al.
Published: (2025)
by: Minh, Nguyen Huu Nhat, et al.
Published: (2025)
AccurateRAG: A Framework for Building Accurate Retrieval-Augmented Question-Answering Applications
by: Nguyen, Linh The, et al.
Published: (2025)
by: Nguyen, Linh The, et al.
Published: (2025)
ViMQ: A Vietnamese Medical Question Dataset for Healthcare Dialogue System Development
by: Huy, Ta Duc, et al.
Published: (2023)
by: Huy, Ta Duc, et al.
Published: (2023)
ViMedCSS: A Vietnamese Medical Code-Switching Speech Dataset & Benchmark
by: Nguyen, Tung X., et al.
Published: (2026)
by: Nguyen, Tung X., et al.
Published: (2026)
Advancing Vietnamese Information Retrieval with Learning Objective and Benchmark
by: Nguyen, Phu-Vinh, et al.
Published: (2025)
by: Nguyen, Phu-Vinh, et al.
Published: (2025)
Bridging the divide between technology and pedagogy: What does a bibliometric analysis reveal about the future of inducing flow in e‐learning?
by: Nguyen Binh Phuong Duy
Published: (2026)
by: Nguyen Binh Phuong Duy
Published: (2026)
VLSP 2025 MLQA-TSR Challenge: Vietnamese Multimodal Legal Question Answering on Traffic Sign Regulation
by: Luu, Son T., et al.
Published: (2025)
by: Luu, Son T., et al.
Published: (2025)
Two New Benzoquinone Derivatives from Vietnamese Knema globularia Stems
by: Huy Truong Nguyen, et al.
Published: (2024)
by: Huy Truong Nguyen, et al.
Published: (2024)
EVJVQA Challenge: Multilingual Visual Question Answering
by: Nguyen, Ngan Luu-Thuy, et al.
Published: (2023)
by: Nguyen, Ngan Luu-Thuy, et al.
Published: (2023)
Describe Anything Model for Visual Question Answering on Text-rich Images
by: Vu, Yen-Linh, et al.
Published: (2025)
by: Vu, Yen-Linh, et al.
Published: (2025)
Evaluating the Symbol Binding Ability of Large Language Models for Multiple-Choice Questions in Vietnamese General Education
by: Nguyen, Duc-Vu, et al.
Published: (2023)
by: Nguyen, Duc-Vu, et al.
Published: (2023)
TSPC: A Two-Stage Phoneme-Centric Architecture for code-switching Vietnamese-English Speech Recognition
by: Anh, Tran Nguyen, et al.
Published: (2025)
by: Anh, Tran Nguyen, et al.
Published: (2025)
NEU-ESC: A Comprehensive Vietnamese dataset for Educational Sentiment analysis and topic Classification toward multitask learning
by: Mai, Phan Quoc Hung, et al.
Published: (2025)
by: Mai, Phan Quoc Hung, et al.
Published: (2025)
New Benchmark Dataset and Fine-Grained Cross-Modal Fusion Framework for Vietnamese Multimodal Aspect-Category Sentiment Analysis
by: Nguyen, Quy Hoang, et al.
Published: (2024)
by: Nguyen, Quy Hoang, et al.
Published: (2024)
Tác động của Ted Talks đến sự phát triển kỹ năng nói tiếng anh của sinh viên chuyên ngữ tại Trường Đại học Tây Nguyên
by: Phạm Văn Phước, et al.
Published: (2025)
by: Phạm Văn Phước, et al.
Published: (2025)
Spatiotemporal dynamics of suspended sediment in coastal Mekong Delta: a hydrodynamic modelling approach under tropical monsoon climate.
by: An, Nguyen Ngoc, et al.
Published: (2025)
by: An, Nguyen Ngoc, et al.
Published: (2025)
MSA-ASR: Efficient Multilingual Speaker Attribution with frozen ASR Models
by: Nguyen, Thai-Binh, et al.
Published: (2024)
by: Nguyen, Thai-Binh, et al.
Published: (2024)
Convoifilter: A case study of doing cocktail party speech recognition
by: Nguyen, Thai-Binh, et al.
Published: (2023)
by: Nguyen, Thai-Binh, et al.
Published: (2023)
Similar Items
-
ViCocktail: Automated Multi-Modal Data Collection for Vietnamese Audio-Visual Speech Recognition
by: Nguyen, Thai-Binh, et al.
Published: (2025) -
AutoViVQA: A Large-Scale Automatically Constructed Dataset for Vietnamese Visual Question Answering
by: Tuong, Nguyen Anh, et al.
Published: (2026) -
ViFactCheck: A New Benchmark Dataset and Methods for Multi-domain News Fact-Checking in Vietnamese
by: Hoa, Tran Thai, et al.
Published: (2024) -
SemViQA: A Semantic Question Answering System for Vietnamese Information Fact-Checking
by: Tran, Dien X., et al.
Published: (2025) -
Human-Guided Reasoning with Large Language Models for Vietnamese Speech Emotion Recognition
by: Nguyen, Truc, et al.
Published: (2026)