CMRAG: Co-modality-based visual document retrieval and question answering
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Wang, Yu, Wenhan, Qi, Guanqiang, Li, Weikang, Li, Yang, Sha, Lei, Xia, Deguo, Huang, Jizhou |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PAIRS: Parametric-Verified Adaptive Information Retrieval and Selection for Efficient RAG
by: Chen, Wang, et al.
Published: (2025)
by: Chen, Wang, et al.
Published: (2025)
Decide Then Retrieve: A Training-Free Framework with Uncertainty-Guided Triggering and Dual-Path Retrieval
by: Chen, Wang, et al.
Published: (2026)
by: Chen, Wang, et al.
Published: (2026)
SciEGQA: A Dataset for Scientific Evidence-Grounded Question Answering and Reasoning
by: Yu, Wenhan, et al.
Published: (2025)
by: Yu, Wenhan, et al.
Published: (2025)
Probabilistic Modeling of Intentions in Socially Intelligent LLM Agents
by: Xia, Feifan, et al.
Published: (2025)
by: Xia, Feifan, et al.
Published: (2025)
DuCCAE: A Hybrid Engine for Immersive Conversation via Collaboration, Augmentation, and Evolution
by: Shen, Xin, et al.
Published: (2026)
by: Shen, Xin, et al.
Published: (2026)
Enhancing textual textbook question answering with large language models and retrieval augmented generation
by: Alawwad, Hessa Abdulrahman, et al.
Published: (2024)
by: Alawwad, Hessa Abdulrahman, et al.
Published: (2024)
Student Guides Teacher: Weak-to-Strong Inference via Spectral Orthogonal Exploration
by: Wang, Dayu, et al.
Published: (2026)
by: Wang, Dayu, et al.
Published: (2026)
Multi-step retrieval and reasoning improves radiology question answering with large language models
by: Wind, Sebastian, et al.
Published: (2025)
by: Wind, Sebastian, et al.
Published: (2025)
keqing: knowledge-based question answering is a nature chain-of-thought mentor of LLM
by: Wang, Chaojie, et al.
Published: (2023)
by: Wang, Chaojie, et al.
Published: (2023)
Beyond End-to-End Video Models: An LLM-Based Multi-Agent System for Educational Video Generation
by: Yan, Lingyong, et al.
Published: (2026)
by: Yan, Lingyong, et al.
Published: (2026)
Unsupervised multiple choices question answering via universal corpus
by: Zhang, Qin, et al.
Published: (2024)
by: Zhang, Qin, et al.
Published: (2024)
Which questions should I answer? Salience Prediction of Inquisitive Questions
by: Wu, Yating, et al.
Published: (2024)
by: Wu, Yating, et al.
Published: (2024)
Can we repurpose multiple-choice question-answering models to rerank retrieved documents?
by: Catapang, Jasper Kyle
Published: (2025)
by: Catapang, Jasper Kyle
Published: (2025)
TANQ: An open domain dataset of table answered questions
by: Akhtar, Mubashara, et al.
Published: (2024)
by: Akhtar, Mubashara, et al.
Published: (2024)
Cross-LoRA: A Data-Free LoRA Transfer Framework across Heterogeneous LLMs
by: Xia, Feifan, et al.
Published: (2025)
by: Xia, Feifan, et al.
Published: (2025)
Agribot: agriculture-specific question answer system
by: Jain, Naman, et al.
Published: (2025)
by: Jain, Naman, et al.
Published: (2025)
FusionMind -- Improving question and answering with external context fusion
by: Verma, Shreyas, et al.
Published: (2023)
by: Verma, Shreyas, et al.
Published: (2023)
Query pipeline optimization for cancer patient question answering systems
by: He, Maolin, et al.
Published: (2024)
by: He, Maolin, et al.
Published: (2024)
Video-MSR: Benchmarking Multi-hop Spatial Reasoning Capabilities of MLLMs
by: Zhu, Rui, et al.
Published: (2026)
by: Zhu, Rui, et al.
Published: (2026)
ACL-Verbatim: hallucination-free question answering for research
by: Recski, Gábor, et al.
Published: (2026)
by: Recski, Gábor, et al.
Published: (2026)
Beyond decomposition: Hierarchical dependency management in multi‐document question answering
by: Xiaoyan Zheng, et al.
Published: (2024)
by: Xiaoyan Zheng, et al.
Published: (2024)
Learning the meanings of function words from grounded language using a visual question answering model
by: Portelance, Eva, et al.
Published: (2023)
by: Portelance, Eva, et al.
Published: (2023)
RealMedQA: A pilot biomedical question answering dataset containing realistic clinical questions
by: Kell, Gregory, et al.
Published: (2024)
by: Kell, Gregory, et al.
Published: (2024)
70B-parameter large language models in Japanese medical question-answering
by: Sukeda, Issey, et al.
Published: (2024)
by: Sukeda, Issey, et al.
Published: (2024)
Stay in Character, Stay Safe: Dual-Cycle Adversarial Self-Evolution for Safety Role-Playing Agents
by: Liao, Mingyang, et al.
Published: (2026)
by: Liao, Mingyang, et al.
Published: (2026)
ConSens: Assessing context grounding in open-book question answering
by: Vankov, Ivan, et al.
Published: (2025)
by: Vankov, Ivan, et al.
Published: (2025)
Spoken question answering for visual queries
by: Shabtay, Nimrod, et al.
Published: (2025)
by: Shabtay, Nimrod, et al.
Published: (2025)
Navigating the Knowledge Sea: Planet-scale answer retrieval using LLMs
by: Sarkar, Dipankar
Published: (2024)
by: Sarkar, Dipankar
Published: (2024)
ChatGPT for automated grading of short answer questions in mechanical ventilation
by: Jade, Tejas, et al.
Published: (2025)
by: Jade, Tejas, et al.
Published: (2025)
Unlocking Multi-View Insights in Knowledge-Dense Retrieval-Augmented Generation
by: Chen, Guanhua, et al.
Published: (2024)
by: Chen, Guanhua, et al.
Published: (2024)
Optimizing open-domain question answering with graph-based retrieval augmented generation
by: Cahoon, Joyce, et al.
Published: (2025)
by: Cahoon, Joyce, et al.
Published: (2025)
Large language models provide unsafe answers to patient-posed medical questions
by: Draelos, Rachel L., et al.
Published: (2025)
by: Draelos, Rachel L., et al.
Published: (2025)
Retrieval augmented text-to-SQL generation for epidemiological question answering using electronic health records
by: Ziletti, Angelo, et al.
Published: (2024)
by: Ziletti, Angelo, et al.
Published: (2024)
From text to multimodal: a survey of adversarial example generation in question answering systems
by: Yigit, Gulsum, et al.
Published: (2023)
by: Yigit, Gulsum, et al.
Published: (2023)
CaLMQA: Exploring culturally specific long-form question answering across 23 languages
by: Arora, Shane, et al.
Published: (2024)
by: Arora, Shane, et al.
Published: (2024)
KeyKnowledgeRAG (K^2RAG): An Enhanced RAG method for improved LLM question-answering capabilities
by: Markondapatnaikuni, Hruday, et al.
Published: (2025)
by: Markondapatnaikuni, Hruday, et al.
Published: (2025)
Few shot chain-of-thought driven reasoning to prompt LLMs for open ended medical question answering
by: Nachane, Saeel Sandeep, et al.
Published: (2024)
by: Nachane, Saeel Sandeep, et al.
Published: (2024)
DALK: Dynamic Co-Augmentation of LLMs and KG to answer Alzheimer's Disease Questions with Scientific Literature
by: Li, Dawei, et al.
Published: (2024)
by: Li, Dawei, et al.
Published: (2024)
What's the plan? Metrics for implicit planning in LLMs and their application to rhyme generation and question answering
by: Maar, Jim, et al.
Published: (2026)
by: Maar, Jim, et al.
Published: (2026)
Automatic question generation for propositional logical equivalences
by: Yang, Yicheng, et al.
Published: (2024)
by: Yang, Yicheng, et al.
Published: (2024)
Similar Items
-
PAIRS: Parametric-Verified Adaptive Information Retrieval and Selection for Efficient RAG
by: Chen, Wang, et al.
Published: (2025) -
Decide Then Retrieve: A Training-Free Framework with Uncertainty-Guided Triggering and Dual-Path Retrieval
by: Chen, Wang, et al.
Published: (2026) -
SciEGQA: A Dataset for Scientific Evidence-Grounded Question Answering and Reasoning
by: Yu, Wenhan, et al.
Published: (2025) -
Probabilistic Modeling of Intentions in Socially Intelligent LLM Agents
by: Xia, Feifan, et al.
Published: (2025) -
DuCCAE: A Hybrid Engine for Immersive Conversation via Collaboration, Augmentation, and Evolution
by: Shen, Xin, et al.
Published: (2026)