Towards Natural Language-Based Document Image Retrieval: New Dataset and Benchmark
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Guo, Hao, Qin, Xugong, Yang, Jun Jie Ou, Zhang, Peng, Zeng, Gangyan, Li, Yubo, Lin, Hailun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Training-Free Scene Text Editing
von: Li, Yubo, et al.
Veröffentlicht: (2026)
von: Li, Yubo, et al.
Veröffentlicht: (2026)
Anveshana: A New Benchmark Dataset for Cross-Lingual Information Retrieval On English Queries and Sanskrit Documents
von: Jagadeeshan, Manoj Balaji, et al.
Veröffentlicht: (2025)
von: Jagadeeshan, Manoj Balaji, et al.
Veröffentlicht: (2025)
Towards Storage-Efficient Visual Document Retrieval: An Empirical Study on Reducing Patch-Level Embeddings
von: Ma, Yubo, et al.
Veröffentlicht: (2025)
von: Ma, Yubo, et al.
Veröffentlicht: (2025)
Summarization-Based Document IDs for Generative Retrieval with Language Models
von: Li, Haoxin, et al.
Veröffentlicht: (2023)
von: Li, Haoxin, et al.
Veröffentlicht: (2023)
Cocktail: A Comprehensive Information Retrieval Benchmark with LLM-Generated Documents Integration
von: Dai, Sunhao, et al.
Veröffentlicht: (2024)
von: Dai, Sunhao, et al.
Veröffentlicht: (2024)
INQUIRE: A Natural World Text-to-Image Retrieval Benchmark
von: Vendrow, Edward, et al.
Veröffentlicht: (2024)
von: Vendrow, Edward, et al.
Veröffentlicht: (2024)
Focus, Distinguish, and Prompt: Unleashing CLIP for Efficient and Flexible Scene Text Retrieval
von: Zeng, Gangyan, et al.
Veröffentlicht: (2024)
von: Zeng, Gangyan, et al.
Veröffentlicht: (2024)
DAPR: A Benchmark on Document-Aware Passage Retrieval
von: Wang, Kexin, et al.
Veröffentlicht: (2023)
von: Wang, Kexin, et al.
Veröffentlicht: (2023)
RAG-RewardBench: Benchmarking Reward Models in Retrieval Augmented Generation for Preference Alignment
von: Jin, Zhuoran, et al.
Veröffentlicht: (2024)
von: Jin, Zhuoran, et al.
Veröffentlicht: (2024)
Towards Text-Image Interleaved Retrieval
von: Zhang, Xin, et al.
Veröffentlicht: (2025)
von: Zhang, Xin, et al.
Veröffentlicht: (2025)
GDLLM: A Global Distance-aware Modeling Approach Based on Large Language Models for Event Temporal Relation Extraction
von: Zhao, Jie, et al.
Veröffentlicht: (2025)
von: Zhao, Jie, et al.
Veröffentlicht: (2025)
Wikipedia-based Datasets in Russian Information Retrieval Benchmark RusBEIR
von: Kovalev, Grigory, et al.
Veröffentlicht: (2025)
von: Kovalev, Grigory, et al.
Veröffentlicht: (2025)
Towards Completeness-Oriented Tool Retrieval for Large Language Models
von: Qu, Changle, et al.
Veröffentlicht: (2024)
von: Qu, Changle, et al.
Veröffentlicht: (2024)
Sculpting the Vector Space: Towards Efficient Multi-Vector Visual Document Retrieval via Prune-then-Merge Framework
von: Yan, Yibo, et al.
Veröffentlicht: (2026)
von: Yan, Yibo, et al.
Veröffentlicht: (2026)
Toward Automatic Relevance Judgment using Vision--Language Models for Image--Text Retrieval Evaluation
von: Yang, Jheng-Hong, et al.
Veröffentlicht: (2024)
von: Yang, Jheng-Hong, et al.
Veröffentlicht: (2024)
InfoDeepSeek: Benchmarking Agentic Information Seeking for Retrieval-Augmented Generation
von: Xi, Yunjia, et al.
Veröffentlicht: (2025)
von: Xi, Yunjia, et al.
Veröffentlicht: (2025)
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories
von: Deng, Chenlong, et al.
Veröffentlicht: (2026)
von: Deng, Chenlong, et al.
Veröffentlicht: (2026)
FreshStack: Building Realistic Benchmarks for Evaluating Retrieval on Technical Documents
von: Thakur, Nandan, et al.
Veröffentlicht: (2025)
von: Thakur, Nandan, et al.
Veröffentlicht: (2025)
Generative Multi-Modal Knowledge Retrieval with Large Language Models
von: Long, Xinwei, et al.
Veröffentlicht: (2024)
von: Long, Xinwei, et al.
Veröffentlicht: (2024)
REGEN: A Dataset and Benchmarks with Natural Language Critiques and Narratives
von: Su, Kun, et al.
Veröffentlicht: (2025)
von: Su, Kun, et al.
Veröffentlicht: (2025)
Large Language Model Informed Patent Image Retrieval
von: Lo, Hao-Cheng, et al.
Veröffentlicht: (2024)
von: Lo, Hao-Cheng, et al.
Veröffentlicht: (2024)
Benchmarking Retrieval-Augmented Multimodal Generation for Document Question Answering
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
FinRetrieval: A Benchmark for Financial Data Retrieval by AI Agents
von: Kim, Eric Y., et al.
Veröffentlicht: (2026)
von: Kim, Eric Y., et al.
Veröffentlicht: (2026)
DBCopilot: Natural Language Querying over Massive Databases via Schema Routing
von: Wang, Tianshu, et al.
Veröffentlicht: (2023)
von: Wang, Tianshu, et al.
Veröffentlicht: (2023)
Entity Image and Mixed-Modal Image Retrieval Datasets
von: Blaga, Cristian-Ioan, et al.
Veröffentlicht: (2025)
von: Blaga, Cristian-Ioan, et al.
Veröffentlicht: (2025)
Rethinking Composed Image Retrieval Evaluation: A Fine-Grained Benchmark from Image Editing
von: Song, Tingyu, et al.
Veröffentlicht: (2026)
von: Song, Tingyu, et al.
Veröffentlicht: (2026)
Domain-Aware RAG: MoL-Enhanced RL for Efficient Training and Scalable Retrieval
von: Lin, Hao, et al.
Veröffentlicht: (2025)
von: Lin, Hao, et al.
Veröffentlicht: (2025)
Beyond the Grid: Layout-Informed Multi-Vector Retrieval with Parsed Visual Document Representations
von: Yan, Yibo, et al.
Veröffentlicht: (2026)
von: Yan, Yibo, et al.
Veröffentlicht: (2026)
Unlocking Multimodal Document Intelligence: From Current Triumphs to Future Frontiers of Visual Document Retrieval
von: Yan, Yibo, et al.
Veröffentlicht: (2026)
von: Yan, Yibo, et al.
Veröffentlicht: (2026)
PosIR: Position-Aware Heterogeneous Information Retrieval Benchmark
von: Zeng, Ziyang, et al.
Veröffentlicht: (2026)
von: Zeng, Ziyang, et al.
Veröffentlicht: (2026)
Tug-of-War Between Knowledge: Exploring and Resolving Knowledge Conflicts in Retrieval-Augmented Language Models
von: Jin, Zhuoran, et al.
Veröffentlicht: (2024)
von: Jin, Zhuoran, et al.
Veröffentlicht: (2024)
Event GDR: Event-Centric Generative Document Retrieval
von: Guan, Yong, et al.
Veröffentlicht: (2024)
von: Guan, Yong, et al.
Veröffentlicht: (2024)
AIR-Bench: Automated Heterogeneous Information Retrieval Benchmark
von: Chen, Jianlyu, et al.
Veröffentlicht: (2024)
von: Chen, Jianlyu, et al.
Veröffentlicht: (2024)
DocReLM: Mastering Document Retrieval with Language Model
von: Wei, Gengchen, et al.
Veröffentlicht: (2024)
von: Wei, Gengchen, et al.
Veröffentlicht: (2024)
MMDocIR: Benchmarking Multimodal Retrieval for Long Documents
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
DeepRAG: Thinking to Retrieve Step by Step for Large Language Models
von: Guan, Xinyan, et al.
Veröffentlicht: (2025)
von: Guan, Xinyan, et al.
Veröffentlicht: (2025)
Towards a Unified Paradigm: Integrating Recommendation Systems as a New Language in Large Models
von: Zheng, Kai, et al.
Veröffentlicht: (2024)
von: Zheng, Kai, et al.
Veröffentlicht: (2024)
ViFactCheck: A New Benchmark Dataset and Methods for Multi-domain News Fact-Checking in Vietnamese
von: Hoa, Tran Thai, et al.
Veröffentlicht: (2024)
von: Hoa, Tran Thai, et al.
Veröffentlicht: (2024)
FollowTable: A Benchmark for Instruction-Following Table Retrieval
von: Jin, Rihui, et al.
Veröffentlicht: (2026)
von: Jin, Rihui, et al.
Veröffentlicht: (2026)
Visual Late Chunking: An Empirical Study of Contextual Chunking for Efficient Visual Document Retrieval
von: Yan, Yibo, et al.
Veröffentlicht: (2026)
von: Yan, Yibo, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Towards Training-Free Scene Text Editing
von: Li, Yubo, et al.
Veröffentlicht: (2026) -
Anveshana: A New Benchmark Dataset for Cross-Lingual Information Retrieval On English Queries and Sanskrit Documents
von: Jagadeeshan, Manoj Balaji, et al.
Veröffentlicht: (2025) -
Towards Storage-Efficient Visual Document Retrieval: An Empirical Study on Reducing Patch-Level Embeddings
von: Ma, Yubo, et al.
Veröffentlicht: (2025) -
Summarization-Based Document IDs for Generative Retrieval with Language Models
von: Li, Haoxin, et al.
Veröffentlicht: (2023) -
Cocktail: A Comprehensive Information Retrieval Benchmark with LLM-Generated Documents Integration
von: Dai, Sunhao, et al.
Veröffentlicht: (2024)