Inference Scaling for Long-Context Retrieval Augmented Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Yue, Zhenrui, Zhuang, Honglei, Bai, Aijun, Hui, Kai, Jagerman, Rolf, Zeng, Hansi, Qin, Zhen, Wang, Dong, Wang, Xuanhui, Bendersky, Michael |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Reliable Confidence Intervals for Information Retrieval Evaluation Using Generative A.I
di: Oosterhuis, Harrie, et al.
Pubblicazione: (2024)
di: Oosterhuis, Harrie, et al.
Pubblicazione: (2024)
Can Query Expansion Improve Generalization of Strong Cross-Encoder Rankers?
di: Li, Minghan, et al.
Pubblicazione: (2023)
di: Li, Minghan, et al.
Pubblicazione: (2023)
Consolidating Ranking and Relevance Predictions of Large Language Models through Post-Processing
di: Yan, Le, et al.
Pubblicazione: (2024)
di: Yan, Le, et al.
Pubblicazione: (2024)
Optimizing Compound Retrieval Systems
di: Oosterhuis, Harrie, et al.
Pubblicazione: (2025)
di: Oosterhuis, Harrie, et al.
Pubblicazione: (2025)
Can Pre-training Indicators Reliably Predict Fine-tuning Outcomes of LLMs?
di: Zeng, Hansi, et al.
Pubblicazione: (2025)
di: Zeng, Hansi, et al.
Pubblicazione: (2025)
Retrieval Augmented Conversational Recommendation with Reinforcement Learning
di: Yue, Zhenrui, et al.
Pubblicazione: (2026)
di: Yue, Zhenrui, et al.
Pubblicazione: (2026)
Beyond Yes and No: Improving Zero-Shot LLM Rankers via Scoring Fine-Grained Relevance Labels
di: Zhuang, Honglei, et al.
Pubblicazione: (2023)
di: Zhuang, Honglei, et al.
Pubblicazione: (2023)
Large Language Models are Effective Text Rankers with Pairwise Ranking Prompting
di: Qin, Zhen, et al.
Pubblicazione: (2023)
di: Qin, Zhen, et al.
Pubblicazione: (2023)
Federated Recommendation via Hybrid Retrieval Augmented Generation
di: Zeng, Huimin, et al.
Pubblicazione: (2024)
di: Zeng, Huimin, et al.
Pubblicazione: (2024)
Integrating Planning into Single-Turn Long-Form Text Generation
di: Liang, Yi, et al.
Pubblicazione: (2024)
di: Liang, Yi, et al.
Pubblicazione: (2024)
Evidence-Driven Retrieval Augmented Response Generation for Online Misinformation
di: Yue, Zhenrui, et al.
Pubblicazione: (2024)
di: Yue, Zhenrui, et al.
Pubblicazione: (2024)
Retrieval Augmented Fact Verification by Synthesizing Contrastive Arguments
di: Yue, Zhenrui, et al.
Pubblicazione: (2024)
di: Yue, Zhenrui, et al.
Pubblicazione: (2024)
Retrieval Augmented Generation or Long-Context LLMs? A Comprehensive Study and Hybrid Approach
di: Li, Zhuowan, et al.
Pubblicazione: (2024)
di: Li, Zhuowan, et al.
Pubblicazione: (2024)
Hybrid Latent Reasoning via Reinforcement Learning
di: Yue, Zhenrui, et al.
Pubblicazione: (2025)
di: Yue, Zhenrui, et al.
Pubblicazione: (2025)
Open-Vocabulary Federated Learning with Multimodal Prototyping
di: Zeng, Huimin, et al.
Pubblicazione: (2024)
di: Zeng, Huimin, et al.
Pubblicazione: (2024)
Searching Personal Collections
di: Bendersky, Michael, et al.
Pubblicazione: (2024)
di: Bendersky, Michael, et al.
Pubblicazione: (2024)
Scaling Sparse and Dense Retrieval in Decoder-Only LLMs
di: Zeng, Hansi, et al.
Pubblicazione: (2025)
di: Zeng, Hansi, et al.
Pubblicazione: (2025)
Stochastic RAG: End-to-End Retrieval-Augmented Generation through Expected Utility Maximization
di: Zamani, Hamed, et al.
Pubblicazione: (2024)
di: Zamani, Hamed, et al.
Pubblicazione: (2024)
Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
di: Jin, Bowen, et al.
Pubblicazione: (2025)
di: Jin, Bowen, et al.
Pubblicazione: (2025)
Planning Ahead in Generative Retrieval: Guiding Autoregressive Generation through Simultaneous Decoding
di: Zeng, Hansi, et al.
Pubblicazione: (2024)
di: Zeng, Hansi, et al.
Pubblicazione: (2024)
Multimodal Reranking for Knowledge-Intensive Visual Question Answering
di: Wen, Haoyang, et al.
Pubblicazione: (2024)
di: Wen, Haoyang, et al.
Pubblicazione: (2024)
Harnessing Pairwise Ranking Prompting Through Sample-Efficient Ranking Distillation
di: Wu, Junru, et al.
Pubblicazione: (2025)
di: Wu, Junru, et al.
Pubblicazione: (2025)
RAPID: Long-Context Inference with Retrieval-Augmented Speculative Decoding
di: Chen, Guanzheng, et al.
Pubblicazione: (2025)
di: Chen, Guanzheng, et al.
Pubblicazione: (2025)
Hypencoder: Hypernetworks for Information Retrieval
di: Killingback, Julian, et al.
Pubblicazione: (2025)
di: Killingback, Julian, et al.
Pubblicazione: (2025)
LongRAG: A Dual-Perspective Retrieval-Augmented Generation Paradigm for Long-Context Question Answering
di: Zhao, Qingfei, et al.
Pubblicazione: (2024)
di: Zhao, Qingfei, et al.
Pubblicazione: (2024)
RAG-HAR: Retrieval Augmented Generation-based Human Activity Recognition
di: Sivaroopan, Nirhoshan, et al.
Pubblicazione: (2025)
di: Sivaroopan, Nirhoshan, et al.
Pubblicazione: (2025)
Inference Scaling for Bridging Retrieval and Augmented Generation
di: Lee, Youngwon, et al.
Pubblicazione: (2024)
di: Lee, Youngwon, et al.
Pubblicazione: (2024)
VideoRAG: Retrieval-Augmented Generation with Extreme Long-Context Videos
di: Ren, Xubin, et al.
Pubblicazione: (2025)
di: Ren, Xubin, et al.
Pubblicazione: (2025)
FeB4RAG: Evaluating Federated Search in the Context of Retrieval Augmented Generation
di: Wang, Shuai, et al.
Pubblicazione: (2024)
di: Wang, Shuai, et al.
Pubblicazione: (2024)
Mindscape-Aware Retrieval Augmented Generation for Improved Long Context Understanding
di: Li, Yuqing, et al.
Pubblicazione: (2025)
di: Li, Yuqing, et al.
Pubblicazione: (2025)
Grounding Long-Context Reasoning with Contextual Normalization for Retrieval-Augmented Generation
di: Chen, Jiamin, et al.
Pubblicazione: (2025)
di: Chen, Jiamin, et al.
Pubblicazione: (2025)
Train Once, Deploy Anywhere: Matryoshka Representation Learning for Multimodal Recommendation
di: Wang, Yueqi, et al.
Pubblicazione: (2024)
di: Wang, Yueqi, et al.
Pubblicazione: (2024)
Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval
di: Yu, Jiwen, et al.
Pubblicazione: (2025)
di: Yu, Jiwen, et al.
Pubblicazione: (2025)
FlashBack:Efficient Retrieval-Augmented Language Modeling for Long Context Inference
di: Liu, Runheng, et al.
Pubblicazione: (2024)
di: Liu, Runheng, et al.
Pubblicazione: (2024)
Exploring Training and Inference Scaling Laws in Generative Retrieval
di: Cai, Hongru, et al.
Pubblicazione: (2025)
di: Cai, Hongru, et al.
Pubblicazione: (2025)
Long$^2$RAG: Evaluating Long-Context & Long-Form Retrieval-Augmented Generation with Key Point Recall
di: Qi, Zehan, et al.
Pubblicazione: (2024)
di: Qi, Zehan, et al.
Pubblicazione: (2024)
Transferable Sequential Recommendation via Vector Quantized Meta Learning
di: Yue, Zhenrui, et al.
Pubblicazione: (2024)
di: Yue, Zhenrui, et al.
Pubblicazione: (2024)
Accelerating Inference of Retrieval-Augmented Generation via Sparse Context Selection
di: Zhu, Yun, et al.
Pubblicazione: (2024)
di: Zhu, Yun, et al.
Pubblicazione: (2024)
RAT: Retrieval Augmented Thoughts Elicit Context-Aware Reasoning in Long-Horizon Generation
di: Wang, Zihao, et al.
Pubblicazione: (2024)
di: Wang, Zihao, et al.
Pubblicazione: (2024)
RetrievalAttention: Accelerating Long-Context LLM Inference via Vector Retrieval
di: Liu, Di, et al.
Pubblicazione: (2024)
di: Liu, Di, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Reliable Confidence Intervals for Information Retrieval Evaluation Using Generative A.I
di: Oosterhuis, Harrie, et al.
Pubblicazione: (2024) -
Can Query Expansion Improve Generalization of Strong Cross-Encoder Rankers?
di: Li, Minghan, et al.
Pubblicazione: (2023) -
Consolidating Ranking and Relevance Predictions of Large Language Models through Post-Processing
di: Yan, Le, et al.
Pubblicazione: (2024) -
Optimizing Compound Retrieval Systems
di: Oosterhuis, Harrie, et al.
Pubblicazione: (2025) -
Can Pre-training Indicators Reliably Predict Fine-tuning Outcomes of LLMs?
di: Zeng, Hansi, et al.
Pubblicazione: (2025)