Is Semantic Chunking Worth the Computational Cost?
Fuente:
arXiv
Guardado en:
| Autores principales: | Qu, Renyi, Tu, Ruixuan, Bao, Forrest |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Knowing When Not to Answer: Lightweight KB-Aligned OOD Detection for Safe RAG
por: Triantafyllopoulos, Ilias, et al.
Publicado: (2025)
por: Triantafyllopoulos, Ilias, et al.
Publicado: (2025)
Query-Adaptive Semantic Chunking for Retrieval-Augmented Generation: A Dynamic Strategy with Contextual Window Expansion
por: Rastogi, Mudit
Publicado: (2026)
por: Rastogi, Mudit
Publicado: (2026)
Adaptive Chunking: Optimizing Chunking-Method Selection for RAG
por: Júnior, Paulo Roberto de Moura, et al.
Publicado: (2026)
por: Júnior, Paulo Roberto de Moura, et al.
Publicado: (2026)
Toward General Semantic Chunking: A Discriminative Framework for Ultra-Long Documents
por: Wu, Kaifeng, et al.
Publicado: (2025)
por: Wu, Kaifeng, et al.
Publicado: (2025)
Structure-Aware Chunking for Tabular Data in Retrieval-Augmented Generation
por: Guttal, Pooja, et al.
Publicado: (2026)
por: Guttal, Pooja, et al.
Publicado: (2026)
A Systematic Analysis of Chunking Strategies for Reliable Question Answering
por: Bennani, Sofia, et al.
Publicado: (2026)
por: Bennani, Sofia, et al.
Publicado: (2026)
Empirical Evaluation of PDF Parsing and Chunking for Financial Question Answering with RAG
por: Bachyr, Omar El, et al.
Publicado: (2026)
por: Bachyr, Omar El, et al.
Publicado: (2026)
S2 Chunking: A Hybrid Framework for Document Segmentation Through Integrated Spatial and Semantic Analysis
por: Verma, Prashant
Publicado: (2025)
por: Verma, Prashant
Publicado: (2025)
ChunkNorris: A High-Performance and Low-Energy Approach to PDF Parsing and Chunking
por: Ciancone, Mathieu, et al.
Publicado: (2025)
por: Ciancone, Mathieu, et al.
Publicado: (2025)
A Text is Worth Several Tokens: Text Embedding from LLMs Secretly Aligns Well with The Key Tokens
por: Nie, Zhijie, et al.
Publicado: (2024)
por: Nie, Zhijie, et al.
Publicado: (2024)
Visual Late Chunking: An Empirical Study of Contextual Chunking for Efficient Visual Document Retrieval
por: Yan, Yibo, et al.
Publicado: (2026)
por: Yan, Yibo, et al.
Publicado: (2026)
Grounding Language Model with Chunking-Free In-Context Retrieval
por: Qian, Hongjin, et al.
Publicado: (2024)
por: Qian, Hongjin, et al.
Publicado: (2024)
LLM-based Embeddings: Attention Values Encode Sentence Semantics Better Than Hidden States
por: Zhang, Yeqin, et al.
Publicado: (2026)
por: Zhang, Yeqin, et al.
Publicado: (2026)
Cross-Document Topic-Aligned Chunking for Retrieval-Augmented Generation
por: Stankovic, Mile
Publicado: (2025)
por: Stankovic, Mile
Publicado: (2025)
SmartChunk Retrieval: Query-Aware Chunk Compression with Planning for Efficient Document RAG
por: Zhang, Xuechen, et al.
Publicado: (2025)
por: Zhang, Xuechen, et al.
Publicado: (2025)
Reconstructing Context: Evaluating Advanced Chunking Strategies for Retrieval-Augmented Generation
por: Merola, Carlo, et al.
Publicado: (2025)
por: Merola, Carlo, et al.
Publicado: (2025)
Beyond Chunking: Discourse-Aware Hierarchical Retrieval for Long Document Question Answering
por: Chen, Huiyao, et al.
Publicado: (2025)
por: Chen, Huiyao, et al.
Publicado: (2025)
PJB: A Reasoning-Aware Benchmark for Person-Job Retrieval
por: Wang, Guangzhi, et al.
Publicado: (2026)
por: Wang, Guangzhi, et al.
Publicado: (2026)
Unveiling the Hidden: Movie Genre and User Bias in Spoiler Detection
por: Zhang, Haokai, et al.
Publicado: (2025)
por: Zhang, Haokai, et al.
Publicado: (2025)
The Chronicles of RAG: The Retriever, the Chunk and the Generator
por: Finardi, Paulo, et al.
Publicado: (2024)
por: Finardi, Paulo, et al.
Publicado: (2024)
Late Chunking: Contextual Chunk Embeddings Using Long-Context Embedding Models
por: Günther, Michael, et al.
Publicado: (2024)
por: Günther, Michael, et al.
Publicado: (2024)
Chunking, Retrieval, and Re-ranking: An Empirical Evaluation of RAG Architectures for Policy Document Question Answering
por: Maharjan, Anuj, et al.
Publicado: (2026)
por: Maharjan, Anuj, et al.
Publicado: (2026)
Semantic Search Evaluation
por: Zheng, Chujie, et al.
Publicado: (2024)
por: Zheng, Chujie, et al.
Publicado: (2024)
Hierarchical Semantic Retrieval with Cobweb
por: Gupta, Anant, et al.
Publicado: (2025)
por: Gupta, Anant, et al.
Publicado: (2025)
Cost-Aware Retrieval-Augmentation Reasoning Models with Adaptive Retrieval Depth
por: Hashemi, Helia, et al.
Publicado: (2025)
por: Hashemi, Helia, et al.
Publicado: (2025)
XProvence: Zero-Cost Multilingual Context Pruning for Retrieval-Augmented Generation
por: Mohamed, Youssef, et al.
Publicado: (2026)
por: Mohamed, Youssef, et al.
Publicado: (2026)
Unlocking Insights: Semantic Search in Jupyter Notebooks
por: Li, Lan, et al.
Publicado: (2024)
por: Li, Lan, et al.
Publicado: (2024)
Multi-Step Semantic Reasoning in Generative Retrieval
por: Dong, Steven, et al.
Publicado: (2026)
por: Dong, Steven, et al.
Publicado: (2026)
Ordered Semantically Diverse Sampling for Textual Data
por: Tiwari, Ashish, et al.
Publicado: (2025)
por: Tiwari, Ashish, et al.
Publicado: (2025)
Isotropy-Optimized Contrastive Learning for Semantic Course Recommendation
por: Khreis, Ali, et al.
Publicado: (2026)
por: Khreis, Ali, et al.
Publicado: (2026)
Improving Document Retrieval Coherence for Semantically Equivalent Queries
por: Campese, Stefano, et al.
Publicado: (2025)
por: Campese, Stefano, et al.
Publicado: (2025)
Don't Start Over: A Cost-Effective Framework for Migrating Personalized Prompts Between LLMs
por: Zhao, Ziyi, et al.
Publicado: (2026)
por: Zhao, Ziyi, et al.
Publicado: (2026)
Purely Semantic Indexing for LLM-based Generative Recommendation and Retrieval
por: Zhang, Ruohan, et al.
Publicado: (2025)
por: Zhang, Ruohan, et al.
Publicado: (2025)
Separating Semantic Competition from Context Length in RAG Reading
por: Repantis, Vyzantinos, et al.
Publicado: (2026)
por: Repantis, Vyzantinos, et al.
Publicado: (2026)
Crafting Personalized Agents through Retrieval-Augmented Generation on Editable Memory Graphs
por: Wang, Zheng, et al.
Publicado: (2024)
por: Wang, Zheng, et al.
Publicado: (2024)
PRE: A Peer Review Based Large Language Model Evaluator
por: Chu, Zhumin, et al.
Publicado: (2024)
por: Chu, Zhumin, et al.
Publicado: (2024)
RbFT: Robust Fine-tuning for Retrieval-Augmented Generation against Retrieval Defects
por: Tu, Yiteng, et al.
Publicado: (2025)
por: Tu, Yiteng, et al.
Publicado: (2025)
Knowledge in Triples for LLMs: Enhancing Table QA Accuracy with Semantic Extraction
por: Sholehrasa, Hossein, et al.
Publicado: (2024)
por: Sholehrasa, Hossein, et al.
Publicado: (2024)
Retrieval over Classification: Integrating Relation Semantics for Multimodal Relation Extraction
por: Hei, Lei, et al.
Publicado: (2025)
por: Hei, Lei, et al.
Publicado: (2025)
Breaking It Down: Domain-Aware Semantic Segmentation for Retrieval Augmented Generation
por: Allamraju, Aparajitha, et al.
Publicado: (2025)
por: Allamraju, Aparajitha, et al.
Publicado: (2025)
Ejemplares similares
-
Knowing When Not to Answer: Lightweight KB-Aligned OOD Detection for Safe RAG
por: Triantafyllopoulos, Ilias, et al.
Publicado: (2025) -
Query-Adaptive Semantic Chunking for Retrieval-Augmented Generation: A Dynamic Strategy with Contextual Window Expansion
por: Rastogi, Mudit
Publicado: (2026) -
Adaptive Chunking: Optimizing Chunking-Method Selection for RAG
por: Júnior, Paulo Roberto de Moura, et al.
Publicado: (2026) -
Toward General Semantic Chunking: A Discriminative Framework for Ultra-Long Documents
por: Wu, Kaifeng, et al.
Publicado: (2025) -
Structure-Aware Chunking for Tabular Data in Retrieval-Augmented Generation
por: Guttal, Pooja, et al.
Publicado: (2026)