DocMMIR: A Framework for Document Multi-modal Information Retrieval
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Zirui, Wu, Siwei, Li, Yizhi, Wang, Xingyu, Zhou, Yi, Lin, Chenghua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SciMMIR: Benchmarking Scientific Multi-modal Information Retrieval
von: Wu, Siwei, et al.
Veröffentlicht: (2024)
von: Wu, Siwei, et al.
Veröffentlicht: (2024)
DocPruner: A Storage-Efficient Framework for Multi-Vector Visual Document Retrieval via Adaptive Patch-Level Embedding Pruning
von: Yan, Yibo, et al.
Veröffentlicht: (2025)
von: Yan, Yibo, et al.
Veröffentlicht: (2025)
DocReLM: Mastering Document Retrieval with Language Model
von: Wei, Gengchen, et al.
Veröffentlicht: (2024)
von: Wei, Gengchen, et al.
Veröffentlicht: (2024)
DocRetriever: A Plug-and-Play Framework for Multimodal Document Retrieval with Comprehensive Benchmark
von: Hu, Ruofan, et al.
Veröffentlicht: (2026)
von: Hu, Ruofan, et al.
Veröffentlicht: (2026)
DocGraphLM: Documental Graph Language Model for Information Extraction
von: Wang, Dongsheng, et al.
Veröffentlicht: (2024)
von: Wang, Dongsheng, et al.
Veröffentlicht: (2024)
Doc-Researcher: A Unified System for Multimodal Document Parsing and Deep Research
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
Improving Multi-modal Recommender Systems by Denoising and Aligning Multi-modal Content and User Feedback
von: Xv, Guipeng, et al.
Veröffentlicht: (2024)
von: Xv, Guipeng, et al.
Veröffentlicht: (2024)
LLM-Augmented Retrieval: Enhancing Retrieval Models Through Language Models and Doc-Level Embedding
von: Wu, Mingrui, et al.
Veröffentlicht: (2024)
von: Wu, Mingrui, et al.
Veröffentlicht: (2024)
MoLoRAG: Bootstrapping Document Understanding via Multi-modal Logic-aware Retrieval
von: Wu, Xixi, et al.
Veröffentlicht: (2025)
von: Wu, Xixi, et al.
Veröffentlicht: (2025)
Simple but Efficient: A Multi-Scenario Nearline Retrieval Framework for Recommendation on Taobao
von: Ma, Yingcai, et al.
Veröffentlicht: (2024)
von: Ma, Yingcai, et al.
Veröffentlicht: (2024)
M3DocDep: Multi-modal, Multi-page, Multi-document Dependency Chunking with Large Vision-Language Models
von: Shin, Joongmin, et al.
Veröffentlicht: (2026)
von: Shin, Joongmin, et al.
Veröffentlicht: (2026)
Unifying Multimodal Retrieval via Document Screenshot Embedding
von: Ma, Xueguang, et al.
Veröffentlicht: (2024)
von: Ma, Xueguang, et al.
Veröffentlicht: (2024)
PAMAS: Self-Adaptive Multi-Agent System with Perspective Aggregation for Misinformation Detection
von: Wang, Zongwei, et al.
Veröffentlicht: (2026)
von: Wang, Zongwei, et al.
Veröffentlicht: (2026)
Composed Multi-modal Retrieval: A Survey of Approaches and Applications
von: Zhang, Kun, et al.
Veröffentlicht: (2025)
von: Zhang, Kun, et al.
Veröffentlicht: (2025)
Beyond Chunk-Then-Embed: A Comprehensive Taxonomy and Evaluation of Document Chunking Strategies for Information Retrieval
von: Zhou, Yongjie, et al.
Veröffentlicht: (2026)
von: Zhou, Yongjie, et al.
Veröffentlicht: (2026)
Teach Me How to Denoise: A Universal Framework for Denoising Multi-modal Recommender Systems via Guided Calibration
von: Li, Hongji, et al.
Veröffentlicht: (2025)
von: Li, Hongji, et al.
Veröffentlicht: (2025)
An Interactive Multi-modal Query Answering System with Retrieval-Augmented Large Language Models
von: Wang, Mengzhao, et al.
Veröffentlicht: (2024)
von: Wang, Mengzhao, et al.
Veröffentlicht: (2024)
Synergizing Implicit and Explicit User Interests: A Multi-Embedding Retrieval Framework at Pinterest
von: Fan, Zhibo, et al.
Veröffentlicht: (2025)
von: Fan, Zhibo, et al.
Veröffentlicht: (2025)
+VeriRel: Verification Feedback to Enhance Document Retrieval for Scientific Fact Checking
von: Deng, Xingyu, et al.
Veröffentlicht: (2025)
von: Deng, Xingyu, et al.
Veröffentlicht: (2025)
A Unified Retrieval Framework with Document Ranking and EDU Filtering for Multi-document Summarization
von: Tan, Shiyin, et al.
Veröffentlicht: (2025)
von: Tan, Shiyin, et al.
Veröffentlicht: (2025)
Deep Class-guided Hashing for Multi-label Cross-modal Retrieval
von: Chen, Hao, et al.
Veröffentlicht: (2024)
von: Chen, Hao, et al.
Veröffentlicht: (2024)
Doc2Query++: Topic-Coverage based Document Expansion and its Application to Dense Retrieval via Dual-Index Fusion
von: Kuo, Tzu-Lin, et al.
Veröffentlicht: (2025)
von: Kuo, Tzu-Lin, et al.
Veröffentlicht: (2025)
MRSE: An Efficient Multi-modality Retrieval System for Large Scale E-commerce
von: Jiang, Hao, et al.
Veröffentlicht: (2024)
von: Jiang, Hao, et al.
Veröffentlicht: (2024)
A Survey of Long-Document Retrieval in the PLM and LLM Era
von: Li, Minghan, et al.
Veröffentlicht: (2025)
von: Li, Minghan, et al.
Veröffentlicht: (2025)
Retrieval-GRPO: A Multi-Objective Reinforcement Learning Framework for Dense Retrieval in Taobao Search
von: Liu, Xingxian, et al.
Veröffentlicht: (2025)
von: Liu, Xingxian, et al.
Veröffentlicht: (2025)
UniFAR: A Unified Facet-Aware Retrieval Framework for Scientific Documents
von: Dou, Zheng, et al.
Veröffentlicht: (2026)
von: Dou, Zheng, et al.
Veröffentlicht: (2026)
Harnessing the Power of Reinforcement Learning for Language-Model-Based Information Retriever via Query-Document Co-Augmentation
von: Liu, Jingming, et al.
Veröffentlicht: (2025)
von: Liu, Jingming, et al.
Veröffentlicht: (2025)
AnnoRetrieve: Efficient Structured Retrieval for Unstructured Document Analysis
von: Lin, Teng, et al.
Veröffentlicht: (2026)
von: Lin, Teng, et al.
Veröffentlicht: (2026)
Model Editing for New Document Integration in Generative Information Retrieval
von: Zhang, Zhen, et al.
Veröffentlicht: (2026)
von: Zhang, Zhen, et al.
Veröffentlicht: (2026)
T-Retrievability: A Topic-Focused Approach to Measure Fair Document Exposure in Information Retrieval
von: Chang, Xuejun, et al.
Veröffentlicht: (2025)
von: Chang, Xuejun, et al.
Veröffentlicht: (2025)
MultiRAG: A Knowledge-guided Framework for Mitigating Hallucination in Multi-source Retrieval Augmented Generation
von: Wu, Wenlong, et al.
Veröffentlicht: (2025)
von: Wu, Wenlong, et al.
Veröffentlicht: (2025)
GeoGR: A Generative Retrieval Framework for Spatio-Temporal Aware POI Recommendation
von: Wang, Fangye, et al.
Veröffentlicht: (2026)
von: Wang, Fangye, et al.
Veröffentlicht: (2026)
Towards Bridging the Cross-modal Semantic Gap for Multi-modal Recommendation
von: Wu, Xinglong, et al.
Veröffentlicht: (2024)
von: Wu, Xinglong, et al.
Veröffentlicht: (2024)
VeritasFi: An Adaptable, Multi-tiered RAG Framework for Multi-modal Financial Question Answering
von: Tai, Zhenghan, et al.
Veröffentlicht: (2025)
von: Tai, Zhenghan, et al.
Veröffentlicht: (2025)
Uni-Retrieval: A Multi-Style Retrieval Framework for STEM's Education
von: Jia, Yanhao, et al.
Veröffentlicht: (2025)
von: Jia, Yanhao, et al.
Veröffentlicht: (2025)
A Multi-Granularity Retrieval Framework for Visually-Rich Documents
von: Xu, Mingjun, et al.
Veröffentlicht: (2025)
von: Xu, Mingjun, et al.
Veröffentlicht: (2025)
SDR-CIR: Semantic Debias Retrieval Framework for Training-Free Zero-Shot Composed Image Retrieval
von: Sun, Yi, et al.
Veröffentlicht: (2026)
von: Sun, Yi, et al.
Veröffentlicht: (2026)
Multimodal Representation Alignment for Cross-modal Information Retrieval
von: Xu, Fan, et al.
Veröffentlicht: (2025)
von: Xu, Fan, et al.
Veröffentlicht: (2025)
Doc2SAR: A Synergistic Framework for High-Fidelity Extraction of Structure-Activity Relationships from Scientific Documents
von: Zhuang, Jiaxi, et al.
Veröffentlicht: (2025)
von: Zhuang, Jiaxi, et al.
Veröffentlicht: (2025)
MSCRS: Multi-modal Semantic Graph Prompt Learning Framework for Conversational Recommender Systems
von: Wei, Yibiao, et al.
Veröffentlicht: (2025)
von: Wei, Yibiao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SciMMIR: Benchmarking Scientific Multi-modal Information Retrieval
von: Wu, Siwei, et al.
Veröffentlicht: (2024) -
DocPruner: A Storage-Efficient Framework for Multi-Vector Visual Document Retrieval via Adaptive Patch-Level Embedding Pruning
von: Yan, Yibo, et al.
Veröffentlicht: (2025) -
DocReLM: Mastering Document Retrieval with Language Model
von: Wei, Gengchen, et al.
Veröffentlicht: (2024) -
DocRetriever: A Plug-and-Play Framework for Multimodal Document Retrieval with Comprehensive Benchmark
von: Hu, Ruofan, et al.
Veröffentlicht: (2026) -
DocGraphLM: Documental Graph Language Model for Information Extraction
von: Wang, Dongsheng, et al.
Veröffentlicht: (2024)