Drowning in Documents: Consequences of Scaling Reranker Inference
Fuente:
arXiv
Saved in:
| Main Authors: | Jacob, Mathew, Lindgren, Erik, Zaharia, Matei, Carbin, Michael, Khattab, Omar, Drozdov, Andrew |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FreshStack: Building Realistic Benchmarks for Evaluating Retrieval on Technical Documents
by: Thakur, Nandan, et al.
Published: (2025)
by: Thakur, Nandan, et al.
Published: (2025)
Can QPP Choose the Right Query Variant? Evaluating Query Variant Selection for RAG Pipelines
by: Arabzadeh, Negar, et al.
Published: (2026)
by: Arabzadeh, Negar, et al.
Published: (2026)
ARES: An Automated Evaluation Framework for Retrieval-Augmented Generation Systems
by: Saad-Falcon, Jon, et al.
Published: (2023)
by: Saad-Falcon, Jon, et al.
Published: (2023)
LangProBe: a Language Programs Benchmark
by: Tan, Shangyin, et al.
Published: (2025)
by: Tan, Shangyin, et al.
Published: (2025)
Relevance Isn't All You Need: Scaling RAG Systems With Inference-Time Compute Via Multi-Criteria Reranking
by: LeVine, Will, et al.
Published: (2025)
by: LeVine, Will, et al.
Published: (2025)
Beyond Sequential Reranking: Reranker-Guided Search Improves Reasoning Intensive Retrieval
by: Xu, Haike, et al.
Published: (2025)
by: Xu, Haike, et al.
Published: (2025)
Retrieval-Enhanced Machine Learning: Synthesis and Opportunities
by: Kim, To Eun, et al.
Published: (2024)
by: Kim, To Eun, et al.
Published: (2024)
Rank1: Test-Time Compute for Reranking in Information Retrieval
by: Weller, Orion, et al.
Published: (2025)
by: Weller, Orion, et al.
Published: (2025)
Diagnosing LLM Reranker Behavior Under Fixed Evidence Pools
by: Arat, Baris, et al.
Published: (2026)
by: Arat, Baris, et al.
Published: (2026)
Long Context RAG Performance of Large Language Models
by: Leng, Quinn, et al.
Published: (2024)
by: Leng, Quinn, et al.
Published: (2024)
Beyond Retrieval: Ensembling Cross-Encoders and GPT Rerankers with LLMs for Biomedical QA
by: Verma, Shashank, et al.
Published: (2025)
by: Verma, Shashank, et al.
Published: (2025)
GraphER: An Efficient Graph-Based Enrichment and Reranking Method for Retrieval-Augmented Generation
by: Miao, Ruizhong, et al.
Published: (2026)
by: Miao, Ruizhong, et al.
Published: (2026)
Efficient Rationale-based Retrieval: On-policy Distillation from Generative Rerankers based on JEPA
by: Chen, Teng, et al.
Published: (2026)
by: Chen, Teng, et al.
Published: (2026)
DS@GT at TREC TOT 2025: Bridging Vague Recollection with Fusion Retrieval and Learned Reranking
by: Zhou, Wenxin, et al.
Published: (2026)
by: Zhou, Wenxin, et al.
Published: (2026)
Enhancing Q&A Text Retrieval with Ranking Models: Benchmarking, fine-tuning and deploying Rerankers for RAG
by: Moreira, Gabriel de Souza P., et al.
Published: (2024)
by: Moreira, Gabriel de Souza P., et al.
Published: (2024)
Prompts as Auto-Optimized Training Hyperparameters: Training Best-in-Class IR Models from Scratch with 10 Gold Labels
by: Xian, Jasper, et al.
Published: (2024)
by: Xian, Jasper, et al.
Published: (2024)
1-800-SHARED-TASKS at RegNLP: Lexical Reranking of Semantic Retrieval (LeSeR) for Regulatory Question Answering
by: Purbey, Jebish, et al.
Published: (2024)
by: Purbey, Jebish, et al.
Published: (2024)
Don't "Overthink" Passage Reranking: Is Reasoning Truly Necessary?
by: Jedidi, Nour, et al.
Published: (2025)
by: Jedidi, Nour, et al.
Published: (2025)
Improving Bilingual Lexicon Induction with Cross-Encoder Reranking
by: Li, Yaoyiran, et al.
Published: (2022)
by: Li, Yaoyiran, et al.
Published: (2022)
Efficient Title Reranker for Fast and Improved Knowledge-Intense NLP
by: Chen, Ziyi, et al.
Published: (2023)
by: Chen, Ziyi, et al.
Published: (2023)
AcuRank: Uncertainty-Aware Adaptive Computation for Listwise Reranking
by: Yoon, Soyoung, et al.
Published: (2025)
by: Yoon, Soyoung, et al.
Published: (2025)
Retrieval Capabilities of Large Language Models Scale with Pretraining FLOPs
by: Portes, Jacob, et al.
Published: (2025)
by: Portes, Jacob, et al.
Published: (2025)
WARP: An Efficient Engine for Multi-Vector Retrieval
by: Scheerer, Jan Luca, et al.
Published: (2025)
by: Scheerer, Jan Luca, et al.
Published: (2025)
RAG over Thinking Traces Can Improve Reasoning Tasks
by: Arabzadeh, Negar, et al.
Published: (2026)
by: Arabzadeh, Negar, et al.
Published: (2026)
Solving the Content Gap in Roblox Game Recommendations: LLM-Based Profile Generation and Reranking
by: Wang, Chen, et al.
Published: (2025)
by: Wang, Chen, et al.
Published: (2025)
PaECTER: Patent-level Representation Learning using Citation-informed Transformers
by: Ghosh, Mainak, et al.
Published: (2024)
by: Ghosh, Mainak, et al.
Published: (2024)
Q-PEFT: Query-dependent Parameter Efficient Fine-tuning for Text Reranking with Large Language Models
by: Peng, Zhiyuan, et al.
Published: (2024)
by: Peng, Zhiyuan, et al.
Published: (2024)
ELMO: Efficiency via Low-precision and Peak Memory Optimization in Large Output Spaces
by: Zhang, Jinbin, et al.
Published: (2025)
by: Zhang, Jinbin, et al.
Published: (2025)
IncDSI: Incrementally Updatable Document Retrieval
by: Kishore, Varsha, et al.
Published: (2023)
by: Kishore, Varsha, et al.
Published: (2023)
Noise-Aware Named Entity Recognition for Historical VET Documents
by: Esser, Alexander M., et al.
Published: (2026)
by: Esser, Alexander M., et al.
Published: (2026)
VERDI: Single-Call Confidence Estimation for Verification-Based LLM Judges via Decomposed Inference
by: Qi, Jasmine, et al.
Published: (2026)
by: Qi, Jasmine, et al.
Published: (2026)
ReFIT: Relevance Feedback from a Reranker during Inference
by: Reddy, Revanth Gangi, et al.
Published: (2023)
by: Reddy, Revanth Gangi, et al.
Published: (2023)
TocBERT: Medical Document Structure Extraction Using Bidirectional Transformers
by: Saleh, Majd, et al.
Published: (2024)
by: Saleh, Majd, et al.
Published: (2024)
DiffuRank: Effective Document Reranking with Diffusion Language Models
by: Liu, Qi, et al.
Published: (2026)
by: Liu, Qi, et al.
Published: (2026)
ABCD-LINK: Annotation Bootstrapping for Cross-Document Fine-Grained Links
by: Basch, Serwar, et al.
Published: (2025)
by: Basch, Serwar, et al.
Published: (2025)
Scaling Up LLM Reviews for Google Ads Content Moderation
by: Qiao, Wei, et al.
Published: (2024)
by: Qiao, Wei, et al.
Published: (2024)
Scaling the Vocabulary of Non-autoregressive Models for Efficient Generative Retrieval
by: Valluri, Ravisri, et al.
Published: (2024)
by: Valluri, Ravisri, et al.
Published: (2024)
Information-Theoretic Generative Clustering of Documents
by: Du, Xin, et al.
Published: (2024)
by: Du, Xin, et al.
Published: (2024)
Gumbel Reranking: Differentiable End-to-End Reranker Optimization
by: Huang, Siyuan, et al.
Published: (2025)
by: Huang, Siyuan, et al.
Published: (2025)
AlpaPICO: Extraction of PICO Frames from Clinical Trial Documents Using LLMs
by: Ghosh, Madhusudan, et al.
Published: (2024)
by: Ghosh, Madhusudan, et al.
Published: (2024)
Similar Items
-
FreshStack: Building Realistic Benchmarks for Evaluating Retrieval on Technical Documents
by: Thakur, Nandan, et al.
Published: (2025) -
Can QPP Choose the Right Query Variant? Evaluating Query Variant Selection for RAG Pipelines
by: Arabzadeh, Negar, et al.
Published: (2026) -
ARES: An Automated Evaluation Framework for Retrieval-Augmented Generation Systems
by: Saad-Falcon, Jon, et al.
Published: (2023) -
LangProBe: a Language Programs Benchmark
by: Tan, Shangyin, et al.
Published: (2025) -
Relevance Isn't All You Need: Scaling RAG Systems With Inference-Time Compute Via Multi-Criteria Reranking
by: LeVine, Will, et al.
Published: (2025)