Multimedia Verification Through Multi-Agent Deep Research Multimodal Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Le, Huy Hoan, Nguyen, Van Sy Thinh, Dang, Thi Le Chi, Nguyen, Vo Thanh Khang, Nguyen, Truong Thanh Hung, Cao, Hung |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Two-Stage, Object-Centric Deep Learning Framework for Robust Exam Cheating Detection
by: Le, Van-Truong, et al.
Published: (2026)
by: Le, Van-Truong, et al.
Published: (2026)
Contestable Multi-Agent Debate with Arena-based Argumentative Computation for Multimedia Verification
by: Nguyen, Truong Thanh Hung, et al.
Published: (2026)
by: Nguyen, Truong Thanh Hung, et al.
Published: (2026)
A Two-stage Transformer Framework for Temporal Localization of Distracted Driver Behaviors
by: Doan, Gia-Bao, et al.
Published: (2026)
by: Doan, Gia-Bao, et al.
Published: (2026)
Examining Monitoring System: Detecting Abnormal Behavior In Online Examinations
by: Ngo, Dinh An, et al.
Published: (2024)
by: Ngo, Dinh An, et al.
Published: (2024)
From Benchmarking to Reasoning: A Dual-Aspect, Large-Scale Evaluation of LLMs on Vietnamese Legal Text
by: Le, Van-Truong
Published: (2026)
by: Le, Van-Truong
Published: (2026)
An Empirical Study for Representations of Videos in Video Question Answering via MLLMs
by: Li, Zhi, et al.
Published: (2025)
by: Li, Zhi, et al.
Published: (2025)
Variational Quantum Rainbow Deep Q-Network for Optimizing Resource Allocation Problem
by: Nguyen, Truong Thanh Hung, et al.
Published: (2025)
by: Nguyen, Truong Thanh Hung, et al.
Published: (2025)
Seeing Roads Through Words: A Language-Guided Framework for RGB-T Driving Scene Segmentation
by: Reddy, Ruturaj, et al.
Published: (2026)
by: Reddy, Ruturaj, et al.
Published: (2026)
ReCoVR: Closing the Loop in Interactive Composed Video Retrieval
by: Zhang, Bingqing, et al.
Published: (2026)
by: Zhang, Bingqing, et al.
Published: (2026)
Enhancing the Fairness and Performance of Edge Cameras with Explainable AI
by: Nguyen, Truong Thanh Hung, et al.
Published: (2024)
by: Nguyen, Truong Thanh Hung, et al.
Published: (2024)
Manipulating Recommender Systems: A Survey of Poisoning Attacks and Countermeasures
by: Nguyen, Thanh Toan, et al.
Published: (2024)
by: Nguyen, Thanh Toan, et al.
Published: (2024)
Enhanced Kalman with Adaptive Appearance Motion SORT for Grounded Generic Multiple Object Tracking
by: Anh, Duy Le Dinh, et al.
Published: (2024)
by: Anh, Duy Le Dinh, et al.
Published: (2024)
MIRA: Empowering One-Touch AI Services on Smartphones with MLLM-based Instruction Recommendation
by: Bian, Zhipeng, et al.
Published: (2025)
by: Bian, Zhipeng, et al.
Published: (2025)
Beyond Vision: Contextually Enriched Image Captioning with Multi-Modal Retrieval
by: Quy, Nguyen Lam Phu, et al.
Published: (2025)
by: Quy, Nguyen Lam Phu, et al.
Published: (2025)
Generalization performance of neural mapping schemes for the space-time interpolation of satellite-derived ocean colour datasets
by: Nguyen, Thi Thuy Nga, et al.
Published: (2025)
by: Nguyen, Thi Thuy Nga, et al.
Published: (2025)
Multi-modal Adaptive Mixture of Experts for Cold-start Recommendation
by: Nguyen, Van-Khang, et al.
Published: (2025)
by: Nguyen, Van-Khang, et al.
Published: (2025)
Efficient and Concise Explanations for Object Detection with Gaussian-Class Activation Mapping Explainer
by: Nguyen, Quoc Khanh, et al.
Published: (2024)
by: Nguyen, Quoc Khanh, et al.
Published: (2024)
TriAlignGR: Triangular Multitask Alignment with Multimodal Deep Interest Mining for Generative Recommendation
by: Zeng, Yangchen, et al.
Published: (2026)
by: Zeng, Yangchen, et al.
Published: (2026)
GIIM: Graph-based Learning of Inter- and Intra-view Dependencies for Multi-view Medical Image Diagnosis
by: Sam, Tran Bao, et al.
Published: (2026)
by: Sam, Tran Bao, et al.
Published: (2026)
Few TensoRF: Enhance the Few-shot on Tensorial Radiance Fields
by: Le, Thanh-Hai, et al.
Published: (2026)
by: Le, Thanh-Hai, et al.
Published: (2026)
Leveraging Lightweight Entity Extraction for Scalable Event-Based Image Retrieval
by: Minh, Dao Sy Duy, et al.
Published: (2025)
by: Minh, Dao Sy Duy, et al.
Published: (2025)
Observation-only learning of neural mapping schemes for gappy satellite-derived ocean colour parameters
by: Dorffer, Clément, et al.
Published: (2025)
by: Dorffer, Clément, et al.
Published: (2025)
Towards Accurate and Efficient Waste Image Classification: A Hybrid Deep Learning and Machine Learning Approach
by: Nguyen, Ngoc-Bao-Quang, et al.
Published: (2025)
by: Nguyen, Ngoc-Bao-Quang, et al.
Published: (2025)
Enhanced Multimodal Video Retrieval System: Integrating Query Expansion and Cross-modal Temporal Event Retrieval
by: Vo, Van-Thinh, et al.
Published: (2025)
by: Vo, Van-Thinh, et al.
Published: (2025)
Heterogeneous Hypergraph Embedding for Recommendation Systems
by: Sakong, Darnbi, et al.
Published: (2024)
by: Sakong, Darnbi, et al.
Published: (2024)
DocRetriever: A Plug-and-Play Framework for Multimodal Document Retrieval with Comprehensive Benchmark
by: Hu, Ruofan, et al.
Published: (2026)
by: Hu, Ruofan, et al.
Published: (2026)
From Top-1 to Top-K: A Reproducibility Study and Benchmarking of Counterfactual Explanations for Recommender Systems
by: Nguyen, Quang-Huy, et al.
Published: (2026)
by: Nguyen, Quang-Huy, et al.
Published: (2026)
Keyword-driven Retrieval-Augmented Large Language Models for Cold-start User Recommendations
by: Kieu, Hai-Dang, et al.
Published: (2024)
by: Kieu, Hai-Dang, et al.
Published: (2024)
FASH-iCNN: Making Editorial Fashion Identity Inspectable Through Multimodal CNN Probing
by: Adeyemi, Morayo Danielle, et al.
Published: (2026)
by: Adeyemi, Morayo Danielle, et al.
Published: (2026)
jina-clip-v2: Multilingual Multimodal Embeddings for Text and Images
by: Koukounas, Andreas, et al.
Published: (2024)
by: Koukounas, Andreas, et al.
Published: (2024)
Navigating Simply, Aligning Deeply: Winning Solutions for Mouse vs. AI 2025
by: Pham, Phu-Hoa, et al.
Published: (2026)
by: Pham, Phu-Hoa, et al.
Published: (2026)
Evaluating Perspectival Biases in Cross-Modal Retrieval
by: Saengsukhiran, Teerapol, et al.
Published: (2025)
by: Saengsukhiran, Teerapol, et al.
Published: (2025)
Bundle Recommendation with Item-level Causation-enhanced Multi-view Learning
by: Nguyen, Huy-Son, et al.
Published: (2024)
by: Nguyen, Huy-Son, et al.
Published: (2024)
Tricks and Plug-ins for Gradient Boosting in Image Classification
by: Fang, Biyi, et al.
Published: (2025)
by: Fang, Biyi, et al.
Published: (2025)
CAPTAIN at COLIEE 2023: Efficient Methods for Legal Information Retrieval and Entailment Tasks
by: Nguyen, Chau, et al.
Published: (2024)
by: Nguyen, Chau, et al.
Published: (2024)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
by: Raoufi, Behnam, et al.
Published: (2025)
by: Raoufi, Behnam, et al.
Published: (2025)
Visible Iris Area as a Quality Metric for Reliable Iris Recognition Under Pupil Dilation and Eyelid Occlusion
by: Pessaud, Jack, et al.
Published: (2025)
by: Pessaud, Jack, et al.
Published: (2025)
ReaderLM-v2: Small Language Model for HTML to Markdown and JSON
by: Wang, Feng, et al.
Published: (2025)
by: Wang, Feng, et al.
Published: (2025)
A Proposed Large Language Model-Based Smart Search for Archive System
by: Nguyen, Ha Dung, et al.
Published: (2025)
by: Nguyen, Ha Dung, et al.
Published: (2025)
Counterfactual Understanding via Retrieval-aware Multimodal Modeling for Time-to-Event Survival Prediction
by: Nguyen, Ha-Anh Hoang, et al.
Published: (2026)
by: Nguyen, Ha-Anh Hoang, et al.
Published: (2026)
Similar Items
-
A Two-Stage, Object-Centric Deep Learning Framework for Robust Exam Cheating Detection
by: Le, Van-Truong, et al.
Published: (2026) -
Contestable Multi-Agent Debate with Arena-based Argumentative Computation for Multimedia Verification
by: Nguyen, Truong Thanh Hung, et al.
Published: (2026) -
A Two-stage Transformer Framework for Temporal Localization of Distracted Driver Behaviors
by: Doan, Gia-Bao, et al.
Published: (2026) -
Examining Monitoring System: Detecting Abnormal Behavior In Online Examinations
by: Ngo, Dinh An, et al.
Published: (2024) -
From Benchmarking to Reasoning: A Dual-Aspect, Large-Scale Evaluation of LLMs on Vietnamese Legal Text
by: Le, Van-Truong
Published: (2026)