Inference Scaling for Long-Context Retrieval Augmented Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Yue, Zhenrui, Zhuang, Honglei, Bai, Aijun, Hui, Kai, Jagerman, Rolf, Zeng, Hansi, Qin, Zhen, Wang, Dong, Wang, Xuanhui, Bendersky, Michael |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Reliable Confidence Intervals for Information Retrieval Evaluation Using Generative A.I
por: Oosterhuis, Harrie, et al.
Publicado: (2024)
por: Oosterhuis, Harrie, et al.
Publicado: (2024)
Can Query Expansion Improve Generalization of Strong Cross-Encoder Rankers?
por: Li, Minghan, et al.
Publicado: (2023)
por: Li, Minghan, et al.
Publicado: (2023)
Consolidating Ranking and Relevance Predictions of Large Language Models through Post-Processing
por: Yan, Le, et al.
Publicado: (2024)
por: Yan, Le, et al.
Publicado: (2024)
Optimizing Compound Retrieval Systems
por: Oosterhuis, Harrie, et al.
Publicado: (2025)
por: Oosterhuis, Harrie, et al.
Publicado: (2025)
Can Pre-training Indicators Reliably Predict Fine-tuning Outcomes of LLMs?
por: Zeng, Hansi, et al.
Publicado: (2025)
por: Zeng, Hansi, et al.
Publicado: (2025)
Retrieval Augmented Conversational Recommendation with Reinforcement Learning
por: Yue, Zhenrui, et al.
Publicado: (2026)
por: Yue, Zhenrui, et al.
Publicado: (2026)
Beyond Yes and No: Improving Zero-Shot LLM Rankers via Scoring Fine-Grained Relevance Labels
por: Zhuang, Honglei, et al.
Publicado: (2023)
por: Zhuang, Honglei, et al.
Publicado: (2023)
Large Language Models are Effective Text Rankers with Pairwise Ranking Prompting
por: Qin, Zhen, et al.
Publicado: (2023)
por: Qin, Zhen, et al.
Publicado: (2023)
Federated Recommendation via Hybrid Retrieval Augmented Generation
por: Zeng, Huimin, et al.
Publicado: (2024)
por: Zeng, Huimin, et al.
Publicado: (2024)
Integrating Planning into Single-Turn Long-Form Text Generation
por: Liang, Yi, et al.
Publicado: (2024)
por: Liang, Yi, et al.
Publicado: (2024)
Evidence-Driven Retrieval Augmented Response Generation for Online Misinformation
por: Yue, Zhenrui, et al.
Publicado: (2024)
por: Yue, Zhenrui, et al.
Publicado: (2024)
Retrieval Augmented Fact Verification by Synthesizing Contrastive Arguments
por: Yue, Zhenrui, et al.
Publicado: (2024)
por: Yue, Zhenrui, et al.
Publicado: (2024)
Retrieval Augmented Generation or Long-Context LLMs? A Comprehensive Study and Hybrid Approach
por: Li, Zhuowan, et al.
Publicado: (2024)
por: Li, Zhuowan, et al.
Publicado: (2024)
Hybrid Latent Reasoning via Reinforcement Learning
por: Yue, Zhenrui, et al.
Publicado: (2025)
por: Yue, Zhenrui, et al.
Publicado: (2025)
Open-Vocabulary Federated Learning with Multimodal Prototyping
por: Zeng, Huimin, et al.
Publicado: (2024)
por: Zeng, Huimin, et al.
Publicado: (2024)
Searching Personal Collections
por: Bendersky, Michael, et al.
Publicado: (2024)
por: Bendersky, Michael, et al.
Publicado: (2024)
Scaling Sparse and Dense Retrieval in Decoder-Only LLMs
por: Zeng, Hansi, et al.
Publicado: (2025)
por: Zeng, Hansi, et al.
Publicado: (2025)
Stochastic RAG: End-to-End Retrieval-Augmented Generation through Expected Utility Maximization
por: Zamani, Hamed, et al.
Publicado: (2024)
por: Zamani, Hamed, et al.
Publicado: (2024)
Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
por: Jin, Bowen, et al.
Publicado: (2025)
por: Jin, Bowen, et al.
Publicado: (2025)
Planning Ahead in Generative Retrieval: Guiding Autoregressive Generation through Simultaneous Decoding
por: Zeng, Hansi, et al.
Publicado: (2024)
por: Zeng, Hansi, et al.
Publicado: (2024)
Multimodal Reranking for Knowledge-Intensive Visual Question Answering
por: Wen, Haoyang, et al.
Publicado: (2024)
por: Wen, Haoyang, et al.
Publicado: (2024)
Harnessing Pairwise Ranking Prompting Through Sample-Efficient Ranking Distillation
por: Wu, Junru, et al.
Publicado: (2025)
por: Wu, Junru, et al.
Publicado: (2025)
RAPID: Long-Context Inference with Retrieval-Augmented Speculative Decoding
por: Chen, Guanzheng, et al.
Publicado: (2025)
por: Chen, Guanzheng, et al.
Publicado: (2025)
Hypencoder: Hypernetworks for Information Retrieval
por: Killingback, Julian, et al.
Publicado: (2025)
por: Killingback, Julian, et al.
Publicado: (2025)
LongRAG: A Dual-Perspective Retrieval-Augmented Generation Paradigm for Long-Context Question Answering
por: Zhao, Qingfei, et al.
Publicado: (2024)
por: Zhao, Qingfei, et al.
Publicado: (2024)
RAG-HAR: Retrieval Augmented Generation-based Human Activity Recognition
por: Sivaroopan, Nirhoshan, et al.
Publicado: (2025)
por: Sivaroopan, Nirhoshan, et al.
Publicado: (2025)
Inference Scaling for Bridging Retrieval and Augmented Generation
por: Lee, Youngwon, et al.
Publicado: (2024)
por: Lee, Youngwon, et al.
Publicado: (2024)
VideoRAG: Retrieval-Augmented Generation with Extreme Long-Context Videos
por: Ren, Xubin, et al.
Publicado: (2025)
por: Ren, Xubin, et al.
Publicado: (2025)
FeB4RAG: Evaluating Federated Search in the Context of Retrieval Augmented Generation
por: Wang, Shuai, et al.
Publicado: (2024)
por: Wang, Shuai, et al.
Publicado: (2024)
Mindscape-Aware Retrieval Augmented Generation for Improved Long Context Understanding
por: Li, Yuqing, et al.
Publicado: (2025)
por: Li, Yuqing, et al.
Publicado: (2025)
Grounding Long-Context Reasoning with Contextual Normalization for Retrieval-Augmented Generation
por: Chen, Jiamin, et al.
Publicado: (2025)
por: Chen, Jiamin, et al.
Publicado: (2025)
Train Once, Deploy Anywhere: Matryoshka Representation Learning for Multimodal Recommendation
por: Wang, Yueqi, et al.
Publicado: (2024)
por: Wang, Yueqi, et al.
Publicado: (2024)
Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval
por: Yu, Jiwen, et al.
Publicado: (2025)
por: Yu, Jiwen, et al.
Publicado: (2025)
FlashBack:Efficient Retrieval-Augmented Language Modeling for Long Context Inference
por: Liu, Runheng, et al.
Publicado: (2024)
por: Liu, Runheng, et al.
Publicado: (2024)
Exploring Training and Inference Scaling Laws in Generative Retrieval
por: Cai, Hongru, et al.
Publicado: (2025)
por: Cai, Hongru, et al.
Publicado: (2025)
Long$^2$RAG: Evaluating Long-Context & Long-Form Retrieval-Augmented Generation with Key Point Recall
por: Qi, Zehan, et al.
Publicado: (2024)
por: Qi, Zehan, et al.
Publicado: (2024)
Transferable Sequential Recommendation via Vector Quantized Meta Learning
por: Yue, Zhenrui, et al.
Publicado: (2024)
por: Yue, Zhenrui, et al.
Publicado: (2024)
Accelerating Inference of Retrieval-Augmented Generation via Sparse Context Selection
por: Zhu, Yun, et al.
Publicado: (2024)
por: Zhu, Yun, et al.
Publicado: (2024)
RAT: Retrieval Augmented Thoughts Elicit Context-Aware Reasoning in Long-Horizon Generation
por: Wang, Zihao, et al.
Publicado: (2024)
por: Wang, Zihao, et al.
Publicado: (2024)
RetrievalAttention: Accelerating Long-Context LLM Inference via Vector Retrieval
por: Liu, Di, et al.
Publicado: (2024)
por: Liu, Di, et al.
Publicado: (2024)
Ejemplares similares
-
Reliable Confidence Intervals for Information Retrieval Evaluation Using Generative A.I
por: Oosterhuis, Harrie, et al.
Publicado: (2024) -
Can Query Expansion Improve Generalization of Strong Cross-Encoder Rankers?
por: Li, Minghan, et al.
Publicado: (2023) -
Consolidating Ranking and Relevance Predictions of Large Language Models through Post-Processing
por: Yan, Le, et al.
Publicado: (2024) -
Optimizing Compound Retrieval Systems
por: Oosterhuis, Harrie, et al.
Publicado: (2025) -
Can Pre-training Indicators Reliably Predict Fine-tuning Outcomes of LLMs?
por: Zeng, Hansi, et al.
Publicado: (2025)