Efficient RAG with Intent-Aware Retrieval and Semantics-Preserving Chunking
Fuente:
arXiv
Saved in:
| Main Authors: | Puspitasari, Fachrina Dewi, Zhang, Chaoning, Zhang, Jiaquan, Wang, Zhicheng, Awan, Hafiz Shakeel Ahmad, Qureshi, Rizwan, Lee, Jewon, Kim, Tae-Ho, Yang, Yang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Small Language Model Helps Resolve Semantic Ambiguity of LLM Prompt
by: Huang, Zhenzhen, et al.
Published: (2026)
by: Huang, Zhenzhen, et al.
Published: (2026)
Fast SAM2 with Text-Driven Token Pruning
by: Mandal, Avilasha, et al.
Published: (2025)
by: Mandal, Avilasha, et al.
Published: (2025)
Geometric Neural Operators via Lie Group-Constrained Latent Dynamics
by: Zhang, Jiaquan, et al.
Published: (2026)
by: Zhang, Jiaquan, et al.
Published: (2026)
Exploring Kernel Transformations for Implicit Neural Representations
by: Zheng, Sheng, et al.
Published: (2025)
by: Zheng, Sheng, et al.
Published: (2025)
Agent-GWO: Collaborative Agents for Dynamic Prompt Optimization in Large Language Models
by: Wang, Xudong, et al.
Published: (2026)
by: Wang, Xudong, et al.
Published: (2026)
Text summarization via global structure awareness
by: Zhang, Jiaquan, et al.
Published: (2026)
by: Zhang, Jiaquan, et al.
Published: (2026)
SmartChunk Retrieval: Query-Aware Chunk Compression with Planning for Efficient Document RAG
by: Zhang, Xuechen, et al.
Published: (2025)
by: Zhang, Xuechen, et al.
Published: (2025)
Topology-Aware Layer Pruning for Large Vision-Language Models
by: Zheng, Pengcheng, et al.
Published: (2026)
by: Zheng, Pengcheng, et al.
Published: (2026)
Sora as a World Model? A Complete Survey on Text-to-Video Generation
by: Puspitasari, Fachrina Dewi, et al.
Published: (2024)
by: Puspitasari, Fachrina Dewi, et al.
Published: (2024)
The Chronicles of RAG: The Retriever, the Chunk and the Generator
by: Finardi, Paulo, et al.
Published: (2024)
by: Finardi, Paulo, et al.
Published: (2024)
A Survey on Segment Anything Model (SAM): Vision Foundation Model Meets Prompt Engineering
by: Zhang, Chaoning, et al.
Published: (2023)
by: Zhang, Chaoning, et al.
Published: (2023)
Weak-Link Optimization for Multi-Agent Reasoning and Collaboration
by: Bian, Haoyu, et al.
Published: (2026)
by: Bian, Haoyu, et al.
Published: (2026)
Beyond Chunking: Discourse-Aware Hierarchical Retrieval for Long Document Question Answering
by: Chen, Huiyao, et al.
Published: (2025)
by: Chen, Huiyao, et al.
Published: (2025)
Transforming External Knowledge into Triplets for Enhanced Retrieval in RAG of LLMs
by: Wang, Xudong, et al.
Published: (2026)
by: Wang, Xudong, et al.
Published: (2026)
Voxel-Aggregated Feature Synthesis: Efficient Dense Mapping for Simulated 3D Reasoning
by: Burns, Owen, et al.
Published: (2024)
by: Burns, Owen, et al.
Published: (2024)
FlashRAG: A Modular Toolkit for Efficient Retrieval-Augmented Generation Research
by: Jin, Jiajie, et al.
Published: (2024)
by: Jin, Jiajie, et al.
Published: (2024)
Efficient and Interpretable Multi-Agent LLM Routing via Ant Colony Optimization
by: Wang, Xudong, et al.
Published: (2026)
by: Wang, Xudong, et al.
Published: (2026)
EHRAG: Bridging Semantic Gaps in Lightweight GraphRAG via Hybrid Hypergraph Construction and Retrieval
by: Song, Yifan, et al.
Published: (2026)
by: Song, Yifan, et al.
Published: (2026)
vGamba: Attentive State Space Bottleneck for efficient Long-range Dependencies in Visual Recognition
by: Haruna, Yunusa, et al.
Published: (2025)
by: Haruna, Yunusa, et al.
Published: (2025)
ChunkRAG: Novel LLM-Chunk Filtering Method for RAG Systems
by: Singh, Ishneet Sukhvinder, et al.
Published: (2024)
by: Singh, Ishneet Sukhvinder, et al.
Published: (2024)
Lightweight LLM Agent Memory with Small Language Models
by: Zhang, Jiaquan, et al.
Published: (2026)
by: Zhang, Jiaquan, et al.
Published: (2026)
Beyond Chunk-Local Extraction: Cross-Chunk Graph Augmentation for GraphRAG
by: Zhang, Jiaming, et al.
Published: (2026)
by: Zhang, Jiaming, et al.
Published: (2026)
AgriLens: Semantic Retrieval in Agricultural Texts Using Topic Modeling and Language Models
by: Shakeel, Heba, et al.
Published: (2026)
by: Shakeel, Heba, et al.
Published: (2026)
Grounding Language Model with Chunking-Free In-Context Retrieval
by: Qian, Hongjin, et al.
Published: (2024)
by: Qian, Hongjin, et al.
Published: (2024)
From Similarity to Structure: Training-free LLM Context Compression with Hybrid Graph Priors
by: Zhou, Yitian, et al.
Published: (2026)
by: Zhou, Yitian, et al.
Published: (2026)
Adaptive Chunking: Optimizing Chunking-Method Selection for RAG
by: Júnior, Paulo Roberto de Moura, et al.
Published: (2026)
by: Júnior, Paulo Roberto de Moura, et al.
Published: (2026)
P$^2$RAG: Efficient Privacy-Preserving RAG Service Supporting Arbitrary Top-$k$ Retrieval
by: Ming, Yulong, et al.
Published: (2026)
by: Ming, Yulong, et al.
Published: (2026)
Adaptive Semantic Chunking for Enhanced Contextual Fidelity in Long-Document RAG
by: Revista, Zen, et al.
Published: (2025)
by: Revista, Zen, et al.
Published: (2025)
EfficientRAG: Efficient Retriever for Multi-Hop Question Answering
by: Zhuang, Ziyuan, et al.
Published: (2024)
by: Zhuang, Ziyuan, et al.
Published: (2024)
ChunkKV: Semantic-Preserving KV Cache Compression for Efficient Long-Context LLM Inference
by: Liu, Xiang, et al.
Published: (2025)
by: Liu, Xiang, et al.
Published: (2025)
Understanding Chain-of-Thought in Large Language Models via Topological Data Analysis
by: Li, Chenghao, et al.
Published: (2025)
by: Li, Chenghao, et al.
Published: (2025)
Rethinking Input Domains in Physics-Informed Neural Networks via Geometric Compactification Mappings
by: Huang, Zhenzhen, et al.
Published: (2026)
by: Huang, Zhenzhen, et al.
Published: (2026)
Web Retrieval-Aware Chunking (W-RAC) for Efficient and Cost-Effective Retrieval-Augmented Generation Systems
by: Allu, Uday, et al.
Published: (2026)
by: Allu, Uday, et al.
Published: (2026)
PACE: Phase-Aware Chunk Execution for Robot Policies with Action Chunking
by: Nie, Junnan, et al.
Published: (2026)
by: Nie, Junnan, et al.
Published: (2026)
REIC: RAG-Enhanced Intent Classification at Scale
by: Zhang, Ziji, et al.
Published: (2025)
by: Zhang, Ziji, et al.
Published: (2025)
Visual Late Chunking: An Empirical Study of Contextual Chunking for Efficient Visual Document Retrieval
by: Yan, Yibo, et al.
Published: (2026)
by: Yan, Yibo, et al.
Published: (2026)
Privacy-Aware RAG: Secure and Isolated Knowledge Retrieval
by: Zhou, Pengcheng, et al.
Published: (2025)
by: Zhou, Pengcheng, et al.
Published: (2025)
TurboRAG: Accelerating Retrieval-Augmented Generation with Precomputed KV Caches for Chunked Text
by: Lu, Songshuo, et al.
Published: (2024)
by: Lu, Songshuo, et al.
Published: (2024)
Breaking Semantic-Aware Watermarks via LLM-Guided Coherence-Preserving Semantic Injection
by: Gao, Zheng, et al.
Published: (2026)
by: Gao, Zheng, et al.
Published: (2026)
LiteSemRAG: Lightweight LLM-Free Semantic-Aware Graph Retrieval for Robust RAG
by: Yue, Xiao, et al.
Published: (2026)
by: Yue, Xiao, et al.
Published: (2026)
Similar Items
-
Small Language Model Helps Resolve Semantic Ambiguity of LLM Prompt
by: Huang, Zhenzhen, et al.
Published: (2026) -
Fast SAM2 with Text-Driven Token Pruning
by: Mandal, Avilasha, et al.
Published: (2025) -
Geometric Neural Operators via Lie Group-Constrained Latent Dynamics
by: Zhang, Jiaquan, et al.
Published: (2026) -
Exploring Kernel Transformations for Implicit Neural Representations
by: Zheng, Sheng, et al.
Published: (2025) -
Agent-GWO: Collaborative Agents for Dynamic Prompt Optimization in Large Language Models
by: Wang, Xudong, et al.
Published: (2026)