Saved in:
| Main Authors: | Nguyen, Van-Quang, Okatani, Takayuki |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2601.19193 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
360° Image Perception with MLLMs: A Comprehensive Benchmark and a Training-Free Method
by: Tran, Huyen T. T., et al.
Published: (2026)
by: Tran, Huyen T. T., et al.
Published: (2026)
LAMS-Edit: Latent and Attention Mixing with Schedulers for Improved Content Preservation in Diffusion-Based Image and Style Editing
by: Fu, Wingwa, et al.
Published: (2026)
by: Fu, Wingwa, et al.
Published: (2026)
Temporal Insight Enhancement: Mitigating Temporal Hallucination in Multimodal Large Language Models
by: Sun, Li, et al.
Published: (2024)
by: Sun, Li, et al.
Published: (2024)
Rethinking Unsupervised Domain Adaptation for Semantic Segmentation
by: Wang, Zhijie, et al.
Published: (2022)
by: Wang, Zhijie, et al.
Published: (2022)
Machine Intelligence that Understands Visual and Linguistic Information and Interacts with Humans and Environments
by: Nguyen, Van Quang
Published: (2026)
by: Nguyen, Van Quang
Published: (2026)
MS-DPPs: Multi-Source Determinantal Point Processes for Contextual Diversity Refinement of Composite Attributes in Text to Image Retrieval
by: Sogi, Naoya, et al.
Published: (2025)
by: Sogi, Naoya, et al.
Published: (2025)
Action-Agnostic Point-Level Supervision for Temporal Action Detection
by: Yoshida, Shuhei M., et al.
Published: (2024)
by: Yoshida, Shuhei M., et al.
Published: (2024)
CoRe-Code: Collaborative Reinforcement Learning for Code Generation
by: Dou, Zhihao, et al.
Published: (2026)
by: Dou, Zhihao, et al.
Published: (2026)
CoRe: Benchmarking LLMs Code Reasoning Capabilities through Static Analysis Tasks
by: Xie, Danning, et al.
Published: (2025)
by: Xie, Danning, et al.
Published: (2025)
CodeCoT: Tackling Code Syntax Errors in CoT Reasoning for Code Generation
by: Huang, Dong, et al.
Published: (2023)
by: Huang, Dong, et al.
Published: (2023)
TB-Bench: Training and Testing Multi-Modal AI for Understanding Spatio-Temporal Traffic Behaviors from Dashcam Images/Videos
by: Charoenpitaks, Korawat, et al.
Published: (2025)
by: Charoenpitaks, Korawat, et al.
Published: (2025)
Vietnamese AI Generated Text Detection
by: Tran, Quang-Dan, et al.
Published: (2024)
by: Tran, Quang-Dan, et al.
Published: (2024)
Temporal Reasoning Is Not the Bottleneck: A Probabilistic Inconsistency Framework for Neuro-Symbolic QA
by: Liem, Tran Quang
Published: (2026)
by: Liem, Tran Quang
Published: (2026)
Multimodal Table Understanding
by: Zheng, Mingyu, et al.
Published: (2024)
by: Zheng, Mingyu, et al.
Published: (2024)
FastCode: Fast and Cost-Efficient Code Understanding and Reasoning
by: Li, Zhonghang, et al.
Published: (2026)
by: Li, Zhonghang, et al.
Published: (2026)
KTVIC: A Vietnamese Image Captioning Dataset on the Life Domain
by: Pham, Anh-Cuong, et al.
Published: (2024)
by: Pham, Anh-Cuong, et al.
Published: (2024)
Point Cloud Compression with Bits-back Coding
by: Hieu, Nguyen Quang, et al.
Published: (2024)
by: Hieu, Nguyen Quang, et al.
Published: (2024)
CompeteSMoE -- Statistically Guaranteed Mixture of Experts Training via Competition
by: Nguyen, Nam V., et al.
Published: (2025)
by: Nguyen, Nam V., et al.
Published: (2025)
ReCode: Reinforcing Code Generation with Reasoning-Process Rewards
by: Fan, Lishui, et al.
Published: (2025)
by: Fan, Lishui, et al.
Published: (2025)
EquiCPI: SE(3)-Equivariant Geometric Deep Learning for Structure-Aware Prediction of Compound-Protein Interactions
by: Nguyen, Ngoc-Quang
Published: (2025)
by: Nguyen, Ngoc-Quang
Published: (2025)
Exploring the Potential of Multi-Modal AI for Driving Hazard Prediction
by: Charoenpitaks, Korawat, et al.
Published: (2023)
by: Charoenpitaks, Korawat, et al.
Published: (2023)
New Benchmark Dataset and Fine-Grained Cross-Modal Fusion Framework for Vietnamese Multimodal Aspect-Category Sentiment Analysis
by: Nguyen, Quy Hoang, et al.
Published: (2024)
by: Nguyen, Quy Hoang, et al.
Published: (2024)
VLegal-Bench: Cognitively Grounded Benchmark for Vietnamese Legal Reasoning of Large Language Models
by: Dong, Nguyen Tien, et al.
Published: (2025)
by: Dong, Nguyen Tien, et al.
Published: (2025)
TabDSR: Decompose, Sanitize, and Reason for Complex Numerical Reasoning in Tabular Data
by: Jiang, Changjiang, et al.
Published: (2025)
by: Jiang, Changjiang, et al.
Published: (2025)
CSD-VAR: Content-Style Decomposition in Visual Autoregressive Models
by: Nguyen, Quang-Binh, et al.
Published: (2025)
by: Nguyen, Quang-Binh, et al.
Published: (2025)
NormTab: Improving Symbolic Reasoning in LLMs Through Tabular Data Normalization
by: Nahid, Md Mahadi Hasan, et al.
Published: (2024)
by: Nahid, Md Mahadi Hasan, et al.
Published: (2024)
Chart-CoCa: Self-Improving Chart Understanding of Vision LMs via Code-Driven Synthesis and Candidate-Conditioned Answering
by: Jiang, Gongyao, et al.
Published: (2025)
by: Jiang, Gongyao, et al.
Published: (2025)
mSCoRe: a $M$ultilingual and Scalable Benchmark for $S$kill-based $Co$mmonsense $Re$asoning
by: Ngo, Nghia Trung, et al.
Published: (2025)
by: Ngo, Nghia Trung, et al.
Published: (2025)
LIBMoE: A Library for comprehensive benchmarking Mixture of Experts in Large Language Models
by: Nguyen, Nam V., et al.
Published: (2024)
by: Nguyen, Nam V., et al.
Published: (2024)
Secure and Efficient UAV-Based Face Detection via Homomorphic Encryption and Edge Computing
by: Van Duc, Nguyen, et al.
Published: (2025)
by: Van Duc, Nguyen, et al.
Published: (2025)
ReVEL: Multi-Turn Reflective LLM-Guided Heuristic Evolution via Structured Performance Feedback
by: Van Duc, Cuong, et al.
Published: (2026)
by: Van Duc, Cuong, et al.
Published: (2026)
ReCode: Improving LLM-based Code Repair with Fine-Grained Retrieval-Augmented Generation
by: Zhao, Yicong, et al.
Published: (2025)
by: Zhao, Yicong, et al.
Published: (2025)
Structural Reasoning Improves Molecular Understanding of LLM
by: Jang, Yunhui, et al.
Published: (2024)
by: Jang, Yunhui, et al.
Published: (2024)
TabTracer: Monte Carlo Tree Search for Complex Table Reasoning with Large Language Models
by: Luo, Zhizhao, et al.
Published: (2026)
by: Luo, Zhizhao, et al.
Published: (2026)
CRUXEval-X: A Benchmark for Multilingual Code Reasoning, Understanding and Execution
by: Xu, Ruiyang, et al.
Published: (2024)
by: Xu, Ruiyang, et al.
Published: (2024)
An Improved Method for Personalizing Diffusion Models
by: Zeng, Yan, et al.
Published: (2024)
by: Zeng, Yan, et al.
Published: (2024)
An Empirical Study on Capability of Large Language Models in Understanding Code Semantics
by: Nguyen, Thu-Trang, et al.
Published: (2024)
by: Nguyen, Thu-Trang, et al.
Published: (2024)
ReConcile: Round-Table Conference Improves Reasoning via Consensus among Diverse LLMs
by: Chen, Justin Chih-Yao, et al.
Published: (2023)
by: Chen, Justin Chih-Yao, et al.
Published: (2023)
Bridging the Version Gap: Multi-version Training Improves ICD Code Prediction, Especially for Rare Codes
by: Liu, Jinghui, et al.
Published: (2026)
by: Liu, Jinghui, et al.
Published: (2026)
ReFeR: Improving Evaluation and Reasoning through Hierarchy of Models
by: Narsupalli, Yaswanth, et al.
Published: (2024)
by: Narsupalli, Yaswanth, et al.
Published: (2024)
Similar Items
-
360° Image Perception with MLLMs: A Comprehensive Benchmark and a Training-Free Method
by: Tran, Huyen T. T., et al.
Published: (2026) -
LAMS-Edit: Latent and Attention Mixing with Schedulers for Improved Content Preservation in Diffusion-Based Image and Style Editing
by: Fu, Wingwa, et al.
Published: (2026) -
Temporal Insight Enhancement: Mitigating Temporal Hallucination in Multimodal Large Language Models
by: Sun, Li, et al.
Published: (2024) -
Rethinking Unsupervised Domain Adaptation for Semantic Segmentation
by: Wang, Zhijie, et al.
Published: (2022) -
Machine Intelligence that Understands Visual and Linguistic Information and Interacts with Humans and Environments
by: Nguyen, Van Quang
Published: (2026)