AutoViVQA: A Large-Scale Automatically Constructed Dataset for Vietnamese Visual Question Answering
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tuong, Nguyen Anh, Duc, Phan Ba, Quoc, Nguyen Trung, Thinh, Tran Dac, Lan, Dang Duy, Thinh, Nguyen Quoc, Le, Tung |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ViInfographicVQA: A Benchmark for Single and Multi-image Visual Question Answering on Vietnamese Infographics
von: Van-Dinh, Tue-Thu, et al.
Veröffentlicht: (2025)
von: Van-Dinh, Tue-Thu, et al.
Veröffentlicht: (2025)
BERT-based model for Vietnamese Fact Verification Dataset
von: Tran, Bao, et al.
Veröffentlicht: (2025)
von: Tran, Bao, et al.
Veröffentlicht: (2025)
ViMQ: A Vietnamese Medical Question Dataset for Healthcare Dialogue System Development
von: Huy, Ta Duc, et al.
Veröffentlicht: (2023)
von: Huy, Ta Duc, et al.
Veröffentlicht: (2023)
Evaluating the Symbol Binding Ability of Large Language Models for Multiple-Choice Questions in Vietnamese General Education
von: Nguyen, Duc-Vu, et al.
Veröffentlicht: (2023)
von: Nguyen, Duc-Vu, et al.
Veröffentlicht: (2023)
ViFactCheck: A New Benchmark Dataset and Methods for Multi-domain News Fact-Checking in Vietnamese
von: Hoa, Tran Thai, et al.
Veröffentlicht: (2024)
von: Hoa, Tran Thai, et al.
Veröffentlicht: (2024)
VietLyrics: A Large-Scale Dataset and Models for Vietnamese Automatic Lyrics Transcription
von: Nguyen, Quoc Anh, et al.
Veröffentlicht: (2025)
von: Nguyen, Quoc Anh, et al.
Veröffentlicht: (2025)
VLSP 2025 MLQA-TSR Challenge: Vietnamese Multimodal Legal Question Answering on Traffic Sign Regulation
von: Luu, Son T., et al.
Veröffentlicht: (2025)
von: Luu, Son T., et al.
Veröffentlicht: (2025)
Towards Signboard-Oriented Visual Question Answering: ViSignVQA Dataset, Method and Benchmark
von: Nguyen, Hieu Minh, et al.
Veröffentlicht: (2025)
von: Nguyen, Hieu Minh, et al.
Veröffentlicht: (2025)
Vietnamese Legal Information Retrieval in Question-Answering System
von: Ba, Thiem Nguyen, et al.
Veröffentlicht: (2024)
von: Ba, Thiem Nguyen, et al.
Veröffentlicht: (2024)
ViTextVQA: A Large-Scale Visual Question Answering Dataset and a Novel Multimodal Feature Fusion Method for Vietnamese Text Comprehension in Images
von: Van Nguyen, Quan, et al.
Veröffentlicht: (2024)
von: Van Nguyen, Quan, et al.
Veröffentlicht: (2024)
Vietnamese Automatic Speech Recognition: A Revisit
von: Vu, Thi, et al.
Veröffentlicht: (2026)
von: Vu, Thi, et al.
Veröffentlicht: (2026)
VlogQA: Task, Dataset, and Baseline Models for Vietnamese Spoken-Based Machine Reading Comprehension
von: Ngo, Thinh Phuoc, et al.
Veröffentlicht: (2024)
von: Ngo, Thinh Phuoc, et al.
Veröffentlicht: (2024)
ViCLIP-OT: The First Foundation Vision-Language Model for Vietnamese Image-Text Retrieval with Optimal Transport
von: Tran, Quoc-Khang, et al.
Veröffentlicht: (2026)
von: Tran, Quoc-Khang, et al.
Veröffentlicht: (2026)
DSC2025 -- ViHallu Challenge: Detecting Hallucination in Vietnamese LLMs
von: Nguyen, Anh Thi-Hoang, et al.
Veröffentlicht: (2026)
von: Nguyen, Anh Thi-Hoang, et al.
Veröffentlicht: (2026)
PhoWhisper: Automatic Speech Recognition for Vietnamese
von: Le, Thanh-Thien, et al.
Veröffentlicht: (2024)
von: Le, Thanh-Thien, et al.
Veröffentlicht: (2024)
ViOCRVQA: Novel Benchmark Dataset and Vision Reader for Visual Question Answering by Understanding Vietnamese Text in Images
von: Pham, Huy Quang, et al.
Veröffentlicht: (2024)
von: Pham, Huy Quang, et al.
Veröffentlicht: (2024)
VLQA: The First Comprehensive, Large, and High-Quality Vietnamese Dataset for Legal Question Answering
von: Nguyen, Tan-Minh, et al.
Veröffentlicht: (2025)
von: Nguyen, Tan-Minh, et al.
Veröffentlicht: (2025)
Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation
von: Nguyen, Huu Tien, et al.
Veröffentlicht: (2025)
von: Nguyen, Huu Tien, et al.
Veröffentlicht: (2025)
Bridging the Reasoning Gap in Vietnamese with Small Language Models via Test-Time Scaling
von: Trung, Bui The, et al.
Veröffentlicht: (2026)
von: Trung, Bui The, et al.
Veröffentlicht: (2026)
Advancing Vietnamese Visual Question Answering with Transformer and Convolutional Integration
von: Nguyen, Ngoc Son, et al.
Veröffentlicht: (2024)
von: Nguyen, Ngoc Son, et al.
Veröffentlicht: (2024)
LiGT: Layout-infused Generative Transformer for Visual Question Answering on Vietnamese Receipts
von: Le, Thanh-Phong, et al.
Veröffentlicht: (2025)
von: Le, Thanh-Phong, et al.
Veröffentlicht: (2025)
More Bias, Less Bias: BiasPrompting for Enhanced Multiple-Choice Question Answering
von: Vu, Duc Anh, et al.
Veröffentlicht: (2025)
von: Vu, Duc Anh, et al.
Veröffentlicht: (2025)
A Green Approach to Heavy Metal Removal Through Invasive Plants Combined with Superabsorbent Polymer
von: Toan Quoc Tran, et al.
Veröffentlicht: (2025)
von: Toan Quoc Tran, et al.
Veröffentlicht: (2025)
Efficiency of freeze‐ and spray‐dried microbial preparation as active dried starter culture in kombucha fermentation
von: Thach Phan Van, et al.
Veröffentlicht: (2024)
von: Thach Phan Van, et al.
Veröffentlicht: (2024)
From two simple problems to the connection of special points
von: Nguyen, Thinh
Veröffentlicht: (2024)
von: Nguyen, Thinh
Veröffentlicht: (2024)
OWLViz: An Open-World Benchmark for Visual Question Answering
von: Nguyen, Thuy, et al.
Veröffentlicht: (2025)
von: Nguyen, Thuy, et al.
Veröffentlicht: (2025)
An Attempt to Develop a Neural Parser based on Simplified Head-Driven Phrase Structure Grammar on Vietnamese
von: Nguyen, Duc-Vu, et al.
Veröffentlicht: (2024)
von: Nguyen, Duc-Vu, et al.
Veröffentlicht: (2024)
Synthesis of 1 H ‐pyrazole frameworks from chalcones using p ‐toluenesulfonic acid as an efficient catalyst
von: Nhat Minh Nguyen, et al.
Veröffentlicht: (2025)
von: Nhat Minh Nguyen, et al.
Veröffentlicht: (2025)
Application of Antioxidant‐ and Antimicrobial‐Rich Extracts From Hass Avocado Pulp in the Development of Chitosan/Gelatin‐Based Active Packaging Films for Raw Meat Preservation
von: Thi Tuong Vi Tran, et al.
Veröffentlicht: (2024)
von: Thi Tuong Vi Tran, et al.
Veröffentlicht: (2024)
The Comprehensive Checklist of the Earthworm Genus Metaphire (Oligochaeta: Megascolecidae) in Vietnam
von: Phan, Quoc T., et al.
Veröffentlicht: (2025)
von: Phan, Quoc T., et al.
Veröffentlicht: (2025)
MERVIN: A Unified Framework for Multimodal Event Retrieval in Vietnamese News Videos
von: Pham-Nguyen, Anh-Tai, et al.
Veröffentlicht: (2026)
von: Pham-Nguyen, Anh-Tai, et al.
Veröffentlicht: (2026)
Zero-Shot Text-to-Speech for Vietnamese
von: Vu, Thi, et al.
Veröffentlicht: (2025)
von: Vu, Thi, et al.
Veröffentlicht: (2025)
SemViQA: A Semantic Question Answering System for Vietnamese Information Fact-Checking
von: Tran, Dien X., et al.
Veröffentlicht: (2025)
von: Tran, Dien X., et al.
Veröffentlicht: (2025)
ViSpeechFormer: A Phonemic Approach for Vietnamese Automatic Speech Recognition
von: Nguyen, Khoa Anh, et al.
Veröffentlicht: (2026)
von: Nguyen, Khoa Anh, et al.
Veröffentlicht: (2026)
ViX-Ray: A Vietnamese Chest X-Ray Dataset for Vision-Language Models
von: Nguyen, Duy Vu Minh, et al.
Veröffentlicht: (2026)
von: Nguyen, Duy Vu Minh, et al.
Veröffentlicht: (2026)
Enhanced Multimodal Video Retrieval System: Integrating Query Expansion and Cross-modal Temporal Event Retrieval
von: Vo, Van-Thinh, et al.
Veröffentlicht: (2025)
von: Vo, Van-Thinh, et al.
Veröffentlicht: (2025)
MedXplain-VQA: Multi-Component Explainable Medical Visual Question Answering
von: Nguyen, Hai-Dang, et al.
Veröffentlicht: (2025)
von: Nguyen, Hai-Dang, et al.
Veröffentlicht: (2025)
Provably Data-driven Multiple Hyper-parameter Tuning with Structured Loss Function
von: Le, Tung Quoc, et al.
Veröffentlicht: (2026)
von: Le, Tung Quoc, et al.
Veröffentlicht: (2026)
Provably Data-driven Lagrangian Relaxation for Mixed Integer Linear Programming
von: Le, Tung Quoc, et al.
Veröffentlicht: (2026)
von: Le, Tung Quoc, et al.
Veröffentlicht: (2026)
EVJVQA Challenge: Multilingual Visual Question Answering
von: Nguyen, Ngan Luu-Thuy, et al.
Veröffentlicht: (2023)
von: Nguyen, Ngan Luu-Thuy, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
ViInfographicVQA: A Benchmark for Single and Multi-image Visual Question Answering on Vietnamese Infographics
von: Van-Dinh, Tue-Thu, et al.
Veröffentlicht: (2025) -
BERT-based model for Vietnamese Fact Verification Dataset
von: Tran, Bao, et al.
Veröffentlicht: (2025) -
ViMQ: A Vietnamese Medical Question Dataset for Healthcare Dialogue System Development
von: Huy, Ta Duc, et al.
Veröffentlicht: (2023) -
Evaluating the Symbol Binding Ability of Large Language Models for Multiple-Choice Questions in Vietnamese General Education
von: Nguyen, Duc-Vu, et al.
Veröffentlicht: (2023) -
ViFactCheck: A New Benchmark Dataset and Methods for Multi-domain News Fact-Checking in Vietnamese
von: Hoa, Tran Thai, et al.
Veröffentlicht: (2024)