Efficient Multivector Retrieval with Token-Aware Clustering and Hierarchical Indexing
Fuente:
arXiv
Saved in:
| Main Authors: | Martinico, Silvio, Nardini, Franco Maria, Rulli, Cosimo, Venturini, Rossano |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multivector Reranking in the Era of Strong First-Stage Retrievers
by: Martinico, Silvio, et al.
Published: (2026)
by: Martinico, Silvio, et al.
Published: (2026)
Efficient Multi-Vector Dense Retrieval Using Bit Vectors
by: Nardini, Franco Maria, et al.
Published: (2024)
by: Nardini, Franco Maria, et al.
Published: (2024)
Efficient Inverted Indexes for Approximate Retrieval over Learned Sparse Representations
by: Bruch, Sebastian, et al.
Published: (2024)
by: Bruch, Sebastian, et al.
Published: (2024)
Pairing Clustered Inverted Indexes with kNN Graphs for Fast Approximate Retrieval over Learned Sparse Representations
by: Bruch, Sebastian, et al.
Published: (2024)
by: Bruch, Sebastian, et al.
Published: (2024)
Efficient Sketching and Nearest Neighbor Search Algorithms for Sparse Vector Sets
by: Bruch, Sebastian, et al.
Published: (2025)
by: Bruch, Sebastian, et al.
Published: (2025)
kANNolo: Sweet and Smooth Approximate k-Nearest Neighbors Search
by: Delfino, Leonardo, et al.
Published: (2025)
by: Delfino, Leonardo, et al.
Published: (2025)
Forward Index Compression for Learned Sparse Retrieval
by: Bruch, Sebastian, et al.
Published: (2026)
by: Bruch, Sebastian, et al.
Published: (2026)
Sparton: Fast and Memory-Efficient Triton Kernel for Learned Sparse Retrieval
by: Nguyen, Thong, et al.
Published: (2026)
by: Nguyen, Thong, et al.
Published: (2026)
Effective Inference-Free Retrieval for Learned Sparse Representations
by: Nardini, Franco Maria, et al.
Published: (2025)
by: Nardini, Franco Maria, et al.
Published: (2025)
Investigating the Scalability of Approximate Sparse Retrieval Algorithms to Massive Datasets
by: Bruch, Sebastian, et al.
Published: (2025)
by: Bruch, Sebastian, et al.
Published: (2025)
Distilled Neural Networks for Efficient Learning to Rank
by: Nardini, F. M., et al.
Published: (2022)
by: Nardini, F. M., et al.
Published: (2022)
Optimistic Query Routing in Clustering-based Approximate Maximum Inner Product Search
by: Bruch, Sebastian, et al.
Published: (2024)
by: Bruch, Sebastian, et al.
Published: (2024)
Efficient Conversational Search via Topical Locality in Dense Retrieval
by: Muntean, Cristina Ioana, et al.
Published: (2025)
by: Muntean, Cristina Ioana, et al.
Published: (2025)
EHI: End-to-end Learning of Hierarchical Index for Efficient Dense Retrieval
by: Kumar, Ramnath, et al.
Published: (2023)
by: Kumar, Ramnath, et al.
Published: (2023)
Blending Learning to Rank and Dense Representations for Efficient and Effective Cascades
by: Nardini, Franco Maria, et al.
Published: (2025)
by: Nardini, Franco Maria, et al.
Published: (2025)
RAG over Tables: Hierarchical Memory Index, Multi-Stage Retrieval, and Benchmarking
by: Zou, Jiaru, et al.
Published: (2025)
by: Zou, Jiaru, et al.
Published: (2025)
Adaptive Retrieval and Scalable Indexing for k-NN Search with Cross-Encoders
by: Yadav, Nishant, et al.
Published: (2024)
by: Yadav, Nishant, et al.
Published: (2024)
Deep Uncertainty-Based Explore for Index Construction and Retrieval in Recommendation System
by: Jiang, Xin, et al.
Published: (2024)
by: Jiang, Xin, et al.
Published: (2024)
Hierarchical Uncertainty-Aware Graph Neural Network
by: Choi, Yoonhyuk, et al.
Published: (2025)
by: Choi, Yoonhyuk, et al.
Published: (2025)
Context Awareness Gate For Retrieval Augmented Generation
by: Heydari, Mohammad Hassan, et al.
Published: (2024)
by: Heydari, Mohammad Hassan, et al.
Published: (2024)
LLM-guided Hierarchical Search for End-to-end Reasoning Intensive Retrieval
by: Gupta, Nilesh, et al.
Published: (2025)
by: Gupta, Nilesh, et al.
Published: (2025)
PinRec: Outcome-Conditioned, Multi-Token Generative Retrieval for Industry-Scale Recommendation Systems
by: Agarwal, Prabhat, et al.
Published: (2025)
by: Agarwal, Prabhat, et al.
Published: (2025)
Resolution-Aware Retrieval Augmented Zero-Shot Forecasting
by: Deznabi, Iman, et al.
Published: (2025)
by: Deznabi, Iman, et al.
Published: (2025)
Bidding-Aware Retrieval for Multi-Stage Consistency in Online Advertising
by: Liu, Bin, et al.
Published: (2025)
by: Liu, Bin, et al.
Published: (2025)
Offline Reasoning for Efficient Recommendation: LLM-Empowered Persona-Profiled Item Indexing
by: Kim, Deogyong, et al.
Published: (2026)
by: Kim, Deogyong, et al.
Published: (2026)
SEAL: Structure and Element Aware Learning to Improve Long Structured Document Retrieval
by: Huang, Xinhao, et al.
Published: (2025)
by: Huang, Xinhao, et al.
Published: (2025)
Efficient Cold-Start Recommendation via BPE Token-Level Embedding Initialization with LLM
by: Zhao, Yushang, et al.
Published: (2025)
by: Zhao, Yushang, et al.
Published: (2025)
Multimodal RAG for Unstructured Data:Leveraging Modality-Aware Knowledge Graphs with Hybrid Retrieval
by: R, Rashmi, et al.
Published: (2025)
by: R, Rashmi, et al.
Published: (2025)
An Efficient Embedding Based Ad Retrieval with GPU-Powered Feature Interaction
by: Lei, Yifan, et al.
Published: (2025)
by: Lei, Yifan, et al.
Published: (2025)
MFBE: Leveraging Multi-Field Information of FAQs for Efficient Dense Retrieval
by: Banerjee, Debopriyo, et al.
Published: (2023)
by: Banerjee, Debopriyo, et al.
Published: (2023)
EraRAG: Efficient and Incremental Retrieval Augmented Generation for Growing Corpora
by: Zhang, Fangyuan, et al.
Published: (2025)
by: Zhang, Fangyuan, et al.
Published: (2025)
Efficient, Property-Aligned Fan-Out Retrieval via RL-Compiled Diffusion
by: Jiang, Pengcheng, et al.
Published: (2026)
by: Jiang, Pengcheng, et al.
Published: (2026)
Hierarchical Retrieval: The Geometry and a Pretrain-Finetune Recipe
by: You, Chong, et al.
Published: (2025)
by: You, Chong, et al.
Published: (2025)
RGL: A Graph-Centric, Modular Framework for Efficient Retrieval-Augmented Generation on Graphs
by: Li, Yuan, et al.
Published: (2025)
by: Li, Yuan, et al.
Published: (2025)
LightKG: Efficient Knowledge-Aware Recommendations with Simplified GNN Architecture
by: Li, Yanhui, et al.
Published: (2025)
by: Li, Yanhui, et al.
Published: (2025)
Towards Efficient Quantity Retrieval from Text:An Approach via Description Parsing and Weak Supervision
by: Cao, Yixuan, et al.
Published: (2025)
by: Cao, Yixuan, et al.
Published: (2025)
Hierarchical Abstract Tree for Cross-Document Retrieval-Augmented Generation
by: Zhao, Ziwen, et al.
Published: (2026)
by: Zhao, Ziwen, et al.
Published: (2026)
A Learning-to-Rank Formulation of Clustering-Based Approximate Nearest Neighbor Search
by: Vecchiato, Thomas, et al.
Published: (2024)
by: Vecchiato, Thomas, et al.
Published: (2024)
Efficient and Effective Query Context-Aware Learning-to-Rank Model for Sequential Recommendation
by: Dzhoha, Andrii, et al.
Published: (2025)
by: Dzhoha, Andrii, et al.
Published: (2025)
VIBE: Vector Index Benchmark for Embeddings
by: Jääsaari, Elias, et al.
Published: (2025)
by: Jääsaari, Elias, et al.
Published: (2025)
Similar Items
-
Multivector Reranking in the Era of Strong First-Stage Retrievers
by: Martinico, Silvio, et al.
Published: (2026) -
Efficient Multi-Vector Dense Retrieval Using Bit Vectors
by: Nardini, Franco Maria, et al.
Published: (2024) -
Efficient Inverted Indexes for Approximate Retrieval over Learned Sparse Representations
by: Bruch, Sebastian, et al.
Published: (2024) -
Pairing Clustered Inverted Indexes with kNN Graphs for Fast Approximate Retrieval over Learned Sparse Representations
by: Bruch, Sebastian, et al.
Published: (2024) -
Efficient Sketching and Nearest Neighbor Search Algorithms for Sparse Vector Sets
by: Bruch, Sebastian, et al.
Published: (2025)