Efficient Document Ranking with Learnable Late Interactions
Fuente:
arXiv
Saved in:
| Main Authors: | Ji, Ziwei, Jain, Himanshu, Veit, Andreas, Reddi, Sashank J., Jayasumana, Sadeep, Rawat, Ankit Singh, Menon, Aditya Krishna, Yu, Felix, Kumar, Sanjiv |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bipartite Ranking From Multiple Labels: On Loss Versus Label Aggregation
by: Lukasik, Michal, et al.
Published: (2025)
by: Lukasik, Michal, et al.
Published: (2025)
Rethinking FID: Towards a Better Evaluation Metric for Image Generation
by: Jayasumana, Sadeep, et al.
Published: (2023)
by: Jayasumana, Sadeep, et al.
Published: (2023)
Think before you speak: Training Language Models With Pause Tokens
by: Goyal, Sachin, et al.
Published: (2023)
by: Goyal, Sachin, et al.
Published: (2023)
Unified Learning-to-Rank for Multi-Channel Retrieval in Large-Scale E-Commerce Search
by: Gaydhani, Aditya, et al.
Published: (2026)
by: Gaydhani, Aditya, et al.
Published: (2026)
Scalable In-context Ranking with Generative Models
by: Gupta, Nilesh, et al.
Published: (2025)
by: Gupta, Nilesh, et al.
Published: (2025)
FastLane: Efficient Routed Systems for Late-Interaction Retrieval
by: Kumar, Ramnath, et al.
Published: (2026)
by: Kumar, Ramnath, et al.
Published: (2026)
LatentCRF: Continuous CRF for Efficient Latent Diffusion
by: Ranasinghe, Kanchana, et al.
Published: (2024)
by: Ranasinghe, Kanchana, et al.
Published: (2024)
A Learnable Fully Interacted Two-Tower Model for Pre-Ranking System
by: Xiong, Chao, et al.
Published: (2025)
by: Xiong, Chao, et al.
Published: (2025)
When Does Confidence-Based Cascade Deferral Suffice?
by: Jitkrittum, Wittawat, et al.
Published: (2023)
by: Jitkrittum, Wittawat, et al.
Published: (2023)
Language Model Cascades: Token-level uncertainty and beyond
by: Gupta, Neha, et al.
Published: (2024)
by: Gupta, Neha, et al.
Published: (2024)
Reproducibility, Replicability, and Insights into Visual Document Retrieval with Late Interaction
by: Qiao, Jingfen, et al.
Published: (2025)
by: Qiao, Jingfen, et al.
Published: (2025)
Autoregressive Ranking: Bridging the Gap Between Dual and Cross Encoders
by: Rozonoyer, Benjamin, et al.
Published: (2026)
by: Rozonoyer, Benjamin, et al.
Published: (2026)
Faster Cascades via Speculative Decoding
by: Narasimhan, Harikrishna, et al.
Published: (2024)
by: Narasimhan, Harikrishna, et al.
Published: (2024)
Revisiting Document-Level Relation Extraction with Context-Guided Link Prediction
by: Jain, Monika, et al.
Published: (2024)
by: Jain, Monika, et al.
Published: (2024)
Analysis of Plan-based Retrieval for Grounded Text Generation
by: Godbole, Ameya, et al.
Published: (2024)
by: Godbole, Ameya, et al.
Published: (2024)
InteractRank: Personalized Web-Scale Search Pre-Ranking with Cross Interaction Features
by: Khandagale, Sujay, et al.
Published: (2025)
by: Khandagale, Sujay, et al.
Published: (2025)
SPLATE: Sparse Late Interaction Retrieval
by: Formal, Thibault, et al.
Published: (2024)
by: Formal, Thibault, et al.
Published: (2024)
Nemotron ColEmbed V2: Top-Performing Late Interaction Embedding Models for Visual Document Retrieval
by: Moreira, Gabriel de Souza P., et al.
Published: (2026)
by: Moreira, Gabriel de Souza P., et al.
Published: (2026)
Knowledge-Driven Cross-Document Relation Extraction
by: Jain, Monika, et al.
Published: (2024)
by: Jain, Monika, et al.
Published: (2024)
AMES: Approximate Multi-modal Enterprise Search via Late Interaction Retrieval
by: Joseph, Tony, et al.
Published: (2026)
by: Joseph, Tony, et al.
Published: (2026)
Rank4Gen: RAG-Preference-Aligned Document Set Selection and Ranking
by: Fan, Yongqi, et al.
Published: (2026)
by: Fan, Yongqi, et al.
Published: (2026)
Ranking Heterogeneous Search Result Pages using the Interactive Probability Ranking Principle
by: Pathak, Kanaad, et al.
Published: (2024)
by: Pathak, Kanaad, et al.
Published: (2024)
Cross-Domain Recommendation Meets Large Language Models
by: Vajjala, Ajay Krishna, et al.
Published: (2024)
by: Vajjala, Ajay Krishna, et al.
Published: (2024)
Structured Preconditioners in Adaptive Optimization: A Unified Analysis
by: Xie, Shuo, et al.
Published: (2025)
by: Xie, Shuo, et al.
Published: (2025)
Query Exposure Prediction for Groups of Documents in Rankings
by: Jaenich, Thomas, et al.
Published: (2024)
by: Jaenich, Thomas, et al.
Published: (2024)
PyLate: Flexible Training and Retrieval for Late Interaction Models
by: Chaffin, Antoine, et al.
Published: (2025)
by: Chaffin, Antoine, et al.
Published: (2025)
The Ranking Blind Spot: Decision Hijacking in LLM-based Text Ranking
by: Qian, Yaoyao, et al.
Published: (2025)
by: Qian, Yaoyao, et al.
Published: (2025)
Learnable Item Tokenization for Generative Recommendation
by: Wang, Wenjie, et al.
Published: (2024)
by: Wang, Wenjie, et al.
Published: (2024)
BracketRank: Large Language Model Document Ranking via Reasoning-based Competitive Elimination
by: Abdallah, Abdelrahman, et al.
Published: (2026)
by: Abdallah, Abdelrahman, et al.
Published: (2026)
DiffuRank: Effective Document Reranking with Diffusion Language Models
by: Liu, Qi, et al.
Published: (2026)
by: Liu, Qi, et al.
Published: (2026)
Detecting Privileged Documents by Ranking Connected Network Entities
by: Zhang, Jianping, et al.
Published: (2025)
by: Zhang, Jianping, et al.
Published: (2025)
Weighted KL-Divergence for Document Ranking Model Refinement
by: Yang, Yingrui, et al.
Published: (2024)
by: Yang, Yingrui, et al.
Published: (2024)
Hierarchical Retrieval: The Geometry and a Pretrain-Finetune Recipe
by: You, Chong, et al.
Published: (2025)
by: You, Chong, et al.
Published: (2025)
Spike Hijacking in Late-Interaction Retrieval
by: Suresh, Karthik, et al.
Published: (2026)
by: Suresh, Karthik, et al.
Published: (2026)
REFINE on Scarce Data: Retrieval Enhancement through Fine-Tuning via Model Fusion of Embedding Models
by: Gupta, Ambuje, et al.
Published: (2024)
by: Gupta, Ambuje, et al.
Published: (2024)
Document Similarity Enhanced IPS Estimation for Unbiased Learning to Rank
by: Liang, Zeyan, et al.
Published: (2025)
by: Liang, Zeyan, et al.
Published: (2025)
RankMamba: Benchmarking Mamba's Document Ranking Performance in the Era of Transformers
by: Xu, Zhichao
Published: (2024)
by: Xu, Zhichao
Published: (2024)
ColBERT-Att: Late-Interaction Meets Attention for Enhanced Retrieval
by: Patel, Raj Nath, et al.
Published: (2026)
by: Patel, Raj Nath, et al.
Published: (2026)
FLASH-MAXSIM: IO-Aware Fused Kernels for Late-Interaction Scoring
by: Pony, Roi, et al.
Published: (2026)
by: Pony, Roi, et al.
Published: (2026)
Working Notes on Late Interaction Dynamics: Analyzing Targeted Behaviors of Late Interaction Models
by: Edy, Antoine, et al.
Published: (2026)
by: Edy, Antoine, et al.
Published: (2026)
Similar Items
-
Bipartite Ranking From Multiple Labels: On Loss Versus Label Aggregation
by: Lukasik, Michal, et al.
Published: (2025) -
Rethinking FID: Towards a Better Evaluation Metric for Image Generation
by: Jayasumana, Sadeep, et al.
Published: (2023) -
Think before you speak: Training Language Models With Pause Tokens
by: Goyal, Sachin, et al.
Published: (2023) -
Unified Learning-to-Rank for Multi-Channel Retrieval in Large-Scale E-Commerce Search
by: Gaydhani, Aditya, et al.
Published: (2026) -
Scalable In-context Ranking with Generative Models
by: Gupta, Nilesh, et al.
Published: (2025)