Bridging the Reasoning Gap in Vietnamese with Small Language Models via Test-Time Scaling
Fuente:
arXiv
Saved in:
| Main Authors: | Trung, Bui The, Duc, Do Minh, Van Vinh, Nguyen, Trinh, Bui Nguyen Quoc |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Using Large Language Models for education managements in Vietnamese with low resources
by: Minh, Duc Do, et al.
Published: (2025)
by: Minh, Duc Do, et al.
Published: (2025)
ViBidirectionMT-Eval: Machine Translation for Vietnamese-Chinese and Vietnamese-Lao language pair
by: Tran, Hong-Viet, et al.
Published: (2025)
by: Tran, Hong-Viet, et al.
Published: (2025)
Auto-Prompting with Retrieval Guidance for Frame Detection in Logistics
by: Duc, Do Minh, et al.
Published: (2025)
by: Duc, Do Minh, et al.
Published: (2025)
A Vietnamese Dataset for Text Segmentation and Multiple Choices Reading Comprehension
by: Hai, Toan Nguyen, et al.
Published: (2025)
by: Hai, Toan Nguyen, et al.
Published: (2025)
Automated Web Application Testing: End-to-End Test Case Generation with Large Language Models and Screen Transition Graphs
by: Le, Nguyen-Khang, et al.
Published: (2025)
by: Le, Nguyen-Khang, et al.
Published: (2025)
VLSP 2025 MLQA-TSR Challenge: Vietnamese Multimodal Legal Question Answering on Traffic Sign Regulation
by: Luu, Son T., et al.
Published: (2025)
by: Luu, Son T., et al.
Published: (2025)
VLegal-Bench: Cognitively Grounded Benchmark for Vietnamese Legal Reasoning of Large Language Models
by: Dong, Nguyen Tien, et al.
Published: (2025)
by: Dong, Nguyen Tien, et al.
Published: (2025)
VietLyrics: A Large-Scale Dataset and Models for Vietnamese Automatic Lyrics Transcription
by: Nguyen, Quoc Anh, et al.
Published: (2025)
by: Nguyen, Quoc Anh, et al.
Published: (2025)
New Benchmark Dataset and Fine-Grained Cross-Modal Fusion Framework for Vietnamese Multimodal Aspect-Category Sentiment Analysis
by: Nguyen, Quy Hoang, et al.
Published: (2024)
by: Nguyen, Quy Hoang, et al.
Published: (2024)
VNJPTranslate: A comprehensive pipeline for Vietnamese-Japanese translation
by: Phan, Hoang Hai, et al.
Published: (2025)
by: Phan, Hoang Hai, et al.
Published: (2025)
DSC2025 -- ViHallu Challenge: Detecting Hallucination in Vietnamese LLMs
by: Nguyen, Anh Thi-Hoang, et al.
Published: (2026)
by: Nguyen, Anh Thi-Hoang, et al.
Published: (2026)
XMainframe: A Large Language Model for Mainframe Modernization
by: Dau, Anh T. V., et al.
Published: (2024)
by: Dau, Anh T. V., et al.
Published: (2024)
ClaimPKG: Enhancing Claim Verification via Pseudo-Subgraph Generation with Lightweight Specialized LLM
by: Pham, Hoang, et al.
Published: (2025)
by: Pham, Hoang, et al.
Published: (2025)
Vietnamese AI Generated Text Detection
by: Tran, Quang-Dan, et al.
Published: (2024)
by: Tran, Quang-Dan, et al.
Published: (2024)
Bridging LLMs and Symbolic Reasoning in Educational QA Systems: Insights from the XAI Challenge at IJCNN 2025
by: Nguyen, Long S. T., et al.
Published: (2025)
by: Nguyen, Long S. T., et al.
Published: (2025)
VLQA: The First Comprehensive, Large, and High-Quality Vietnamese Dataset for Legal Question Answering
by: Nguyen, Tan-Minh, et al.
Published: (2025)
by: Nguyen, Tan-Minh, et al.
Published: (2025)
ZeFaV: Boosting Large Language Models for Zero-shot Fact Verification
by: Luu, Son T., et al.
Published: (2024)
by: Luu, Son T., et al.
Published: (2024)
Formal Reasoning for Intelligent QA Systems: A Case Study in the Educational Domain
by: Bui, Tuan, et al.
Published: (2025)
by: Bui, Tuan, et al.
Published: (2025)
Reasoning Planning for Language Models
by: Nguyen, Bao, et al.
Published: (2025)
by: Nguyen, Bao, et al.
Published: (2025)
Whisper based Cross-Lingual Phoneme Recognition between Vietnamese and English
by: Minh, Nguyen Huu Nhat, et al.
Published: (2025)
by: Minh, Nguyen Huu Nhat, et al.
Published: (2025)
Enhancing Legal Document Retrieval: A Multi-Phase Approach with Large Language Models
by: Nguyen, Hai-Long, et al.
Published: (2024)
by: Nguyen, Hai-Long, et al.
Published: (2024)
PEEB: Part-based Image Classifiers with an Explainable and Editable Language Bottleneck
by: Pham, Thang M., et al.
Published: (2024)
by: Pham, Thang M., et al.
Published: (2024)
ViSoLex: An Open-Source Repository for Vietnamese Social Media Lexical Normalization
by: Nguyen, Anh Thi-Hoang, et al.
Published: (2025)
by: Nguyen, Anh Thi-Hoang, et al.
Published: (2025)
AutoViVQA: A Large-Scale Automatically Constructed Dataset for Vietnamese Visual Question Answering
by: Tuong, Nguyen Anh, et al.
Published: (2026)
by: Tuong, Nguyen Anh, et al.
Published: (2026)
Crossing Linguistic Horizons: Finetuning and Comprehensive Evaluation of Vietnamese Large Language Models
by: Truong, Sang T., et al.
Published: (2024)
by: Truong, Sang T., et al.
Published: (2024)
Exploiting LLMs' Reasoning Capability to Infer Implicit Concepts in Legal Information Retrieval
by: Nguyen, Hai-Long, et al.
Published: (2024)
by: Nguyen, Hai-Long, et al.
Published: (2024)
A Weakly Supervised Data Labeling Framework for Machine Lexical Normalization in Vietnamese Social Media
by: Nguyen, Dung Ha, et al.
Published: (2024)
by: Nguyen, Dung Ha, et al.
Published: (2024)
Bridging the Training-Deployment Gap: Gated Encoding and Multi-Scale Refinement for Efficient Quantization-Aware Image Enhancement
by: To-Thanh, Dat, et al.
Published: (2026)
by: To-Thanh, Dat, et al.
Published: (2026)
ViLLM-Eval: A Comprehensive Evaluation Suite for Vietnamese Large Language Models
by: Nguyen, Trong-Hieu, et al.
Published: (2024)
by: Nguyen, Trong-Hieu, et al.
Published: (2024)
Evaluating the Symbol Binding Ability of Large Language Models for Multiple-Choice Questions in Vietnamese General Education
by: Nguyen, Duc-Vu, et al.
Published: (2023)
by: Nguyen, Duc-Vu, et al.
Published: (2023)
NEU-ESC: A Comprehensive Vietnamese dataset for Educational Sentiment analysis and topic Classification toward multitask learning
by: Mai, Phan Quoc Hung, et al.
Published: (2025)
by: Mai, Phan Quoc Hung, et al.
Published: (2025)
Speaking in Words, Thinking in Logic: A Dual-Process Framework in QA Systems
by: Bui, Tuan, et al.
Published: (2025)
by: Bui, Tuan, et al.
Published: (2025)
TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching
by: Nguyen, Truong, et al.
Published: (2026)
by: Nguyen, Truong, et al.
Published: (2026)
VN-MTEB: Vietnamese Massive Text Embedding Benchmark
by: Pham, Loc, et al.
Published: (2025)
by: Pham, Loc, et al.
Published: (2025)
NOWJ@COLIEE 2025: A Multi-stage Framework Integrating Embedding Models and Large Language Models for Legal Retrieval and Entailment
by: Nguyen, Hoang-Trung, et al.
Published: (2025)
by: Nguyen, Hoang-Trung, et al.
Published: (2025)
Verify-in-the-Graph: Entity Disambiguation Enhancement for Complex Claim Verification with Interactive Graph Representation
by: Pham, Hoang, et al.
Published: (2025)
by: Pham, Hoang, et al.
Published: (2025)
A Hybrid Method for Low-Resource Named Entity Recognition
by: Duc, Do Minh, et al.
Published: (2026)
by: Duc, Do Minh, et al.
Published: (2026)
ViCLIP-OT: The First Foundation Vision-Language Model for Vietnamese Image-Text Retrieval with Optimal Transport
by: Tran, Quoc-Khang, et al.
Published: (2026)
by: Tran, Quoc-Khang, et al.
Published: (2026)
PhoGPT: Generative Pre-training for Vietnamese
by: Nguyen, Dat Quoc, et al.
Published: (2023)
by: Nguyen, Dat Quoc, et al.
Published: (2023)
ViMQ: A Vietnamese Medical Question Dataset for Healthcare Dialogue System Development
by: Huy, Ta Duc, et al.
Published: (2023)
by: Huy, Ta Duc, et al.
Published: (2023)
Similar Items
-
Using Large Language Models for education managements in Vietnamese with low resources
by: Minh, Duc Do, et al.
Published: (2025) -
ViBidirectionMT-Eval: Machine Translation for Vietnamese-Chinese and Vietnamese-Lao language pair
by: Tran, Hong-Viet, et al.
Published: (2025) -
Auto-Prompting with Retrieval Guidance for Frame Detection in Logistics
by: Duc, Do Minh, et al.
Published: (2025) -
A Vietnamese Dataset for Text Segmentation and Multiple Choices Reading Comprehension
by: Hai, Toan Nguyen, et al.
Published: (2025) -
Automated Web Application Testing: End-to-End Test Case Generation with Large Language Models and Screen Transition Graphs
by: Le, Nguyen-Khang, et al.
Published: (2025)