Vision-Guided Chunking Is All You Need: Enhancing RAG with Multimodal Document Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Tripathi, Vishesh, Odapally, Tanmay, Das, Indraneel, Allu, Uday, Ahmed, Biddwan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Web Retrieval-Aware Chunking (W-RAC) for Efficient and Cost-Effective Retrieval-Augmented Generation Systems
by: Allu, Uday, et al.
Published: (2026)
by: Allu, Uday, et al.
Published: (2026)
The Instruction Gap: LLMs get lost in Following Instruction
by: Tripathi, Vishesh, et al.
Published: (2025)
by: Tripathi, Vishesh, et al.
Published: (2025)
Adaptive Chunking: Optimizing Chunking-Method Selection for RAG
by: Júnior, Paulo Roberto de Moura, et al.
Published: (2026)
by: Júnior, Paulo Roberto de Moura, et al.
Published: (2026)
Beyond Extraction: Contextualising Tabular Data for Efficient Summarisation by Language Models
by: Allu, Uday, et al.
Published: (2024)
by: Allu, Uday, et al.
Published: (2024)
SmartChunk Retrieval: Query-Aware Chunk Compression with Planning for Efficient Document RAG
by: Zhang, Xuechen, et al.
Published: (2025)
by: Zhang, Xuechen, et al.
Published: (2025)
Enhancing Technical Documents Retrieval for RAG
by: Lai, Songjiang, et al.
Published: (2025)
by: Lai, Songjiang, et al.
Published: (2025)
Action is All You Need: Dual-Flow Generative Ranking Network for Recommendation
by: Guo, Hao, et al.
Published: (2025)
by: Guo, Hao, et al.
Published: (2025)
Chunking, Retrieval, and Re-ranking: An Empirical Evaluation of RAG Architectures for Policy Document Question Answering
by: Maharjan, Anuj, et al.
Published: (2026)
by: Maharjan, Anuj, et al.
Published: (2026)
The Chronicles of RAG: The Retriever, the Chunk and the Generator
by: Finardi, Paulo, et al.
Published: (2024)
by: Finardi, Paulo, et al.
Published: (2024)
Do We Still Need GraphRAG? Benchmarking RAG and GraphRAG for Agentic Search Systems
by: Fan, Dongzhe, et al.
Published: (2026)
by: Fan, Dongzhe, et al.
Published: (2026)
Evaluating Chunking Strategies For Retrieval-Augmented Generation in Oil and Gas Enterprise Documents
by: Taiwo, Samuel, et al.
Published: (2026)
by: Taiwo, Samuel, et al.
Published: (2026)
Information Extraction from Visually Rich Documents using LLM-based Organization of Documents into Independent Textual Segments
by: Bhattacharyya, Aniket, et al.
Published: (2025)
by: Bhattacharyya, Aniket, et al.
Published: (2025)
Chunk Knowledge Generation Model for Enhanced Information Retrieval: A Multi-task Learning Approach
by: Kim, Jisu, et al.
Published: (2025)
by: Kim, Jisu, et al.
Published: (2025)
Cross-Document Topic-Aligned Chunking for Retrieval-Augmented Generation
by: Stankovic, Mile
Published: (2025)
by: Stankovic, Mile
Published: (2025)
Lexicalization Is All You Need: Examining the Impact of Lexical Knowledge in a Compositional QALD System
by: Schmidt, David Maria, et al.
Published: (2024)
by: Schmidt, David Maria, et al.
Published: (2024)
LFRAG: Layout-oriented Fine-grained Retrieval-Augmented Generation on Multimodal Document Understanding
by: Zhu, Yifan, et al.
Published: (2026)
by: Zhu, Yifan, et al.
Published: (2026)
M3DocDep: Multi-modal, Multi-page, Multi-document Dependency Chunking with Large Vision-Language Models
by: Shin, Joongmin, et al.
Published: (2026)
by: Shin, Joongmin, et al.
Published: (2026)
Knowledge Graph RAG: Agentic Crawling and Graph Construction in Enterprise Documents
by: Chakraborty, Koushik, et al.
Published: (2026)
by: Chakraborty, Koushik, et al.
Published: (2026)
Toward General Semantic Chunking: A Discriminative Framework for Ultra-Long Documents
by: Wu, Kaifeng, et al.
Published: (2025)
by: Wu, Kaifeng, et al.
Published: (2025)
Beyond Chunking: Discourse-Aware Hierarchical Retrieval for Long Document Question Answering
by: Chen, Huiyao, et al.
Published: (2025)
by: Chen, Huiyao, et al.
Published: (2025)
MG$^2$-RAG: Multi-Granularity Graph for Multimodal Retrieval-Augmented Generation
by: Dai, Sijun, et al.
Published: (2026)
by: Dai, Sijun, et al.
Published: (2026)
Graph-Aware Late Chunking for Retrieval-Augmented Generation in Biomedical Literature
by: Mortezaagha, Pouria, et al.
Published: (2026)
by: Mortezaagha, Pouria, et al.
Published: (2026)
LARAG: Link-Aware Retrieval Strategy for RAG Systems in Hyperlinked Technical Documentation
by: Bolognesi, Giorgia, et al.
Published: (2026)
by: Bolognesi, Giorgia, et al.
Published: (2026)
LERA: LLM-Enhanced RAG for Ad Auction in Generative Chatbots
by: Sun, Haoran, et al.
Published: (2026)
by: Sun, Haoran, et al.
Published: (2026)
MCERF: Advancing Multimodal LLM Evaluation of Engineering Documentation with Enhanced Retrieval
by: Khanghah, Kiarash Naghavi, et al.
Published: (2026)
by: Khanghah, Kiarash Naghavi, et al.
Published: (2026)
Beyond Patch Aggregation: 3-Pass Pyramid Indexing for Vision-Enhanced Document Retrieval
by: Roy, Anup, et al.
Published: (2025)
by: Roy, Anup, et al.
Published: (2025)
Structured Attention Matters to Multimodal LLMs in Document Understanding
by: Liu, Chang, et al.
Published: (2025)
by: Liu, Chang, et al.
Published: (2025)
ChunkNorris: A High-Performance and Low-Energy Approach to PDF Parsing and Chunking
by: Ciancone, Mathieu, et al.
Published: (2025)
by: Ciancone, Mathieu, et al.
Published: (2025)
Pseudo-Knowledge Graph: Meta-Path Guided Retrieval and In-Graph Text for RAG-Equipped LLM
by: Yang, Yuxin, et al.
Published: (2025)
by: Yang, Yuxin, et al.
Published: (2025)
DMQR-RAG: Diverse Multi-Query Rewriting for RAG
by: Li, Zhicong, et al.
Published: (2024)
by: Li, Zhicong, et al.
Published: (2024)
HugRAG: Hierarchical Causal Knowledge Graph Design for RAG
by: Wang, Nengbo, et al.
Published: (2026)
by: Wang, Nengbo, et al.
Published: (2026)
M-RAG: Making RAG Faster, Stronger, and More Efficient
by: Xu, Sun, et al.
Published: (2026)
by: Xu, Sun, et al.
Published: (2026)
HF-RAG: Hierarchical Fusion-based RAG with Multiple Sources and Rankers
by: Santra, Payel, et al.
Published: (2025)
by: Santra, Payel, et al.
Published: (2025)
BookRAG: A Hierarchical Structure-aware Index-based Approach for Retrieval-Augmented Generation on Complex Documents
by: Wang, Shu, et al.
Published: (2025)
by: Wang, Shu, et al.
Published: (2025)
PriHA: A RAG-Enhanced LLM Framework for Primary Healthcare Assistant in Hong Kong
by: Chan, Richard Wai Cheung, et al.
Published: (2026)
by: Chan, Richard Wai Cheung, et al.
Published: (2026)
Can't Remember Details in Long Documents? You Need Some R&R
by: Agrawal, Devanshu, et al.
Published: (2024)
by: Agrawal, Devanshu, et al.
Published: (2024)
RAGdb: A Zero-Dependency, Embeddable Architecture for Multimodal Retrieval-Augmented Generation on the Edge
by: Khalid, Ahmed Bin
Published: (2025)
by: Khalid, Ahmed Bin
Published: (2025)
EcphoryRAG: Re-Imagining Knowledge-Graph RAG via Human Associative Memory
by: Liao, Zirui
Published: (2025)
by: Liao, Zirui
Published: (2025)
Bahasa Harmony: A Comprehensive Dataset for Bahasa Text-to-Speech Synthesis with Discrete Codec Modeling of EnGen-TTS
by: Susladkar, Onkar Kishor, et al.
Published: (2024)
by: Susladkar, Onkar Kishor, et al.
Published: (2024)
Progressive Searching for Retrieval in RAG
by: Jeong, Taehee, et al.
Published: (2026)
by: Jeong, Taehee, et al.
Published: (2026)
Similar Items
-
Web Retrieval-Aware Chunking (W-RAC) for Efficient and Cost-Effective Retrieval-Augmented Generation Systems
by: Allu, Uday, et al.
Published: (2026) -
The Instruction Gap: LLMs get lost in Following Instruction
by: Tripathi, Vishesh, et al.
Published: (2025) -
Adaptive Chunking: Optimizing Chunking-Method Selection for RAG
by: Júnior, Paulo Roberto de Moura, et al.
Published: (2026) -
Beyond Extraction: Contextualising Tabular Data for Efficient Summarisation by Language Models
by: Allu, Uday, et al.
Published: (2024) -
SmartChunk Retrieval: Query-Aware Chunk Compression with Planning for Efficient Document RAG
by: Zhang, Xuechen, et al.
Published: (2025)