GleanVec: Accelerating vector search with minimalist nonlinear dimensionality reduction
Fuente:
arXiv
Saved in:
| Main Authors: | Tepper, Mariano, Bhati, Ishwar Singh, Aguerrebere, Cecilia, Willke, Ted |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Locally-Adaptive Quantization for Streaming Vector Search
by: Aguerrebere, Cecilia, et al.
Published: (2024)
by: Aguerrebere, Cecilia, et al.
Published: (2024)
Individualized non-uniform quantization for vector search
by: Tepper, Mariano, et al.
Published: (2025)
by: Tepper, Mariano, et al.
Published: (2025)
LeanVec: Searching vectors faster by making them fit
by: Tepper, Mariano, et al.
Published: (2023)
by: Tepper, Mariano, et al.
Published: (2023)
The kernel of graph indices for vector search
by: Tepper, Mariano, et al.
Published: (2025)
by: Tepper, Mariano, et al.
Published: (2025)
Toward Optimal Search and Retrieval for RAG
by: Leto, Alexandria, et al.
Published: (2024)
by: Leto, Alexandria, et al.
Published: (2024)
Interact2Vec -- An efficient neural network-based model for simultaneously learning users and items embeddings in recommender systems
by: Pires, Pedro R., et al.
Published: (2025)
by: Pires, Pedro R., et al.
Published: (2025)
Multimodal semantic retrieval for product search
by: Liu, Dong, et al.
Published: (2025)
by: Liu, Dong, et al.
Published: (2025)
AI Guided Accelerator For Search Experience
by: Yetukuri, Jayanth, et al.
Published: (2025)
by: Yetukuri, Jayanth, et al.
Published: (2025)
Accelerating Matrix Factorization by Dynamic Pruning for Fast Recommendation
by: Wu, Yining, et al.
Published: (2024)
by: Wu, Yining, et al.
Published: (2024)
Accelerating Recommender Model Training by Dynamically Skipping Stale Embeddings
by: Maboud, Yassaman Ebrahimzadeh, et al.
Published: (2024)
by: Maboud, Yassaman Ebrahimzadeh, et al.
Published: (2024)
Building a privacy-preserving Federated Recommender system for mobile devices
by: Singh, Aasheesh
Published: (2026)
by: Singh, Aasheesh
Published: (2026)
MST-R: Multi-Stage Tuning for Retrieval Systems and Metric Evaluation
by: Malviya, Yash, et al.
Published: (2024)
by: Malviya, Yash, et al.
Published: (2024)
Guarding Digital Privacy: Exploring User Profiling and Security Enhancements
by: Kohli, Rishika, et al.
Published: (2025)
by: Kohli, Rishika, et al.
Published: (2025)
Learning Metrics that Maximise Power for Accelerated A/B-Tests
by: Jeunen, Olivier, et al.
Published: (2024)
by: Jeunen, Olivier, et al.
Published: (2024)
Retrieval Mechanisms Surpass Long-Context Scaling in Time Series Forecasting
by: Ahuja, Rishi, et al.
Published: (2026)
by: Ahuja, Rishi, et al.
Published: (2026)
AMES: Approximate Multi-modal Enterprise Search via Late Interaction Retrieval
by: Joseph, Tony, et al.
Published: (2026)
by: Joseph, Tony, et al.
Published: (2026)
Segment Discovery: Enhancing E-commerce Targeting
by: Li, Qiqi, et al.
Published: (2024)
by: Li, Qiqi, et al.
Published: (2024)
MS2MetGAN: Latent-space adversarial training for metabolite-spectrum matching in MS/MS database search
by: Tsai, Meng, et al.
Published: (2026)
by: Tsai, Meng, et al.
Published: (2026)
Accelerating Retrieval-Augmented Language Model Serving with Speculation
by: Zhang, Zhihao, et al.
Published: (2024)
by: Zhang, Zhihao, et al.
Published: (2024)
Better Generalization with Semantic IDs: A Case Study in Ranking for Recommendations
by: Singh, Anima, et al.
Published: (2023)
by: Singh, Anima, et al.
Published: (2023)
DV365: Extremely Long User History Modeling at Instagram
by: Lyu, Wenhan, et al.
Published: (2025)
by: Lyu, Wenhan, et al.
Published: (2025)
Embedding-based search in JetBrains IDEs
by: Abramov, Evgeny, et al.
Published: (2024)
by: Abramov, Evgeny, et al.
Published: (2024)
Graph Regularized Encoder Training for Extreme Classification
by: Mittal, Anshul, et al.
Published: (2024)
by: Mittal, Anshul, et al.
Published: (2024)
Deep Learning Model Acceleration and Optimization Strategies for Real-Time Recommendation Systems
by: Shao, Junli, et al.
Published: (2025)
by: Shao, Junli, et al.
Published: (2025)
Vectorizing the Trie: Efficient Constrained Decoding for LLM-based Generative Retrieval on Accelerators
by: Su, Zhengyang, et al.
Published: (2026)
by: Su, Zhengyang, et al.
Published: (2026)
Lookahead: An Inference Acceleration Framework for Large Language Model with Lossless Generation Accuracy
by: Zhao, Yao, et al.
Published: (2023)
by: Zhao, Yao, et al.
Published: (2023)
Sequential Recommendation via Adaptive Robust Attention with Multi-dimensional Embeddings
by: Pang, Linsey, et al.
Published: (2024)
by: Pang, Linsey, et al.
Published: (2024)
Intelligent Elastic Feature Fading: Enabling Model Retrain-Free Feature Efficiency Rollouts at Scale
by: Di, Jieming, et al.
Published: (2026)
by: Di, Jieming, et al.
Published: (2026)
A case study of Generative AI in MSX Sales Copilot: Improving seller productivity with a real-time question-answering system for content recommendation
by: Singh, Manpreet, et al.
Published: (2024)
by: Singh, Manpreet, et al.
Published: (2024)
Query Attribute Modeling: Improving search relevance with Semantic Search and Meta Data Filtering
by: Menon, Karthik, et al.
Published: (2025)
by: Menon, Karthik, et al.
Published: (2025)
Fine-grained large-scale content recommendations for MSX sellers
by: Singh, Manpreet, et al.
Published: (2024)
by: Singh, Manpreet, et al.
Published: (2024)
How Many Tools Should an LLM Agent See? A Chance-Corrected Answer
by: Repantis, Vyzantinos, et al.
Published: (2026)
by: Repantis, Vyzantinos, et al.
Published: (2026)
Improving Few-Shot Cross-Domain Named Entity Recognition by Instruction Tuning a Word-Embedding based Retrieval Augmented Large Language Model
by: Nandi, Subhadip, et al.
Published: (2024)
by: Nandi, Subhadip, et al.
Published: (2024)
Autoregressive Generation Strategies for Top-K Sequential Recommendations
by: Volodkevich, Anna, et al.
Published: (2024)
by: Volodkevich, Anna, et al.
Published: (2024)
Revisiting Bi-Encoder Neural Search: An Encoding--Searching Separation Perspective
by: Tran, Hung-Nghiep, et al.
Published: (2024)
by: Tran, Hung-Nghiep, et al.
Published: (2024)
RDSA: A Robust Deep Graph Clustering Framework via Dual Soft Assignment
by: Xiang, Yang, et al.
Published: (2024)
by: Xiang, Yang, et al.
Published: (2024)
Against Filter Bubbles: Diversified Music Recommendation via Weighted Hypergraph Embedding Learning
by: Luo, Chaoguang, et al.
Published: (2024)
by: Luo, Chaoguang, et al.
Published: (2024)
Let's Get It Started: Fostering the Discoverability of New Releases on Deezer
by: Briand, Léa, et al.
Published: (2024)
by: Briand, Léa, et al.
Published: (2024)
Model-Free Approximate Bayesian Learning for Large-Scale Conversion Funnel Optimization
by: Iyengar, Garud, et al.
Published: (2024)
by: Iyengar, Garud, et al.
Published: (2024)
Domain Adaptation of Multilingual Semantic Search -- Literature Review
by: Bringmann, Anna, et al.
Published: (2024)
by: Bringmann, Anna, et al.
Published: (2024)
Similar Items
-
Locally-Adaptive Quantization for Streaming Vector Search
by: Aguerrebere, Cecilia, et al.
Published: (2024) -
Individualized non-uniform quantization for vector search
by: Tepper, Mariano, et al.
Published: (2025) -
LeanVec: Searching vectors faster by making them fit
by: Tepper, Mariano, et al.
Published: (2023) -
The kernel of graph indices for vector search
by: Tepper, Mariano, et al.
Published: (2025) -
Toward Optimal Search and Retrieval for RAG
by: Leto, Alexandria, et al.
Published: (2024)