Long Context vs. RAG for LLMs: An Evaluation and Revisits
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Xinze, Cao, Yixin, Ma, Yubo, Sun, Aixin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Verifiable Generation: A Benchmark for Knowledge-aware Language Model Attribution
von: Li, Xinze, et al.
Veröffentlicht: (2023)
von: Li, Xinze, et al.
Veröffentlicht: (2023)
Skill-as-Pseudocode: Refactoring Skill Libraries to Pseudocode for LLM Agents
von: Li, Xinze, et al.
Veröffentlicht: (2026)
von: Li, Xinze, et al.
Veröffentlicht: (2026)
EMemBench: Interactive Benchmarking of Episodic Memory for VLM Agents
von: Li, Xinze, et al.
Veröffentlicht: (2026)
von: Li, Xinze, et al.
Veröffentlicht: (2026)
Large Language Model Is Not a Good Few-shot Information Extractor, but a Good Reranker for Hard Samples!
von: Ma, Yubo, et al.
Veröffentlicht: (2023)
von: Ma, Yubo, et al.
Veröffentlicht: (2023)
Route Before Retrieve: Activating Latent Routing Abilities of LLMs for RAG vs. Long-Context Selection
von: Chen, Yiwen, et al.
Veröffentlicht: (2026)
von: Chen, Yiwen, et al.
Veröffentlicht: (2026)
MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations
von: Ma, Yubo, et al.
Veröffentlicht: (2024)
von: Ma, Yubo, et al.
Veröffentlicht: (2024)
Evaluating LLM Adaptation to Sociodemographic Factors: User Profile vs. Dialogue History
von: Zhong, Qishuai, et al.
Veröffentlicht: (2025)
von: Zhong, Qishuai, et al.
Veröffentlicht: (2025)
EffiEval: Efficient and Generalizable Model Evaluation via Capability Coverage Maximization
von: Wang, Yaoning, et al.
Veröffentlicht: (2025)
von: Wang, Yaoning, et al.
Veröffentlicht: (2025)
Long-Context LLMs Meet RAG: Overcoming Challenges for Long Inputs in RAG
von: Jin, Bowen, et al.
Veröffentlicht: (2024)
von: Jin, Bowen, et al.
Veröffentlicht: (2024)
Agri-Query: A Case Study on RAG vs. Long-Context LLMs for Cross-Lingual Technical Question Answering
von: Gun, Julius, et al.
Veröffentlicht: (2025)
von: Gun, Julius, et al.
Veröffentlicht: (2025)
Cultural Value Differences of LLMs: Prompt, Language, and Model Size
von: Zhong, Qishuai, et al.
Veröffentlicht: (2024)
von: Zhong, Qishuai, et al.
Veröffentlicht: (2024)
Towards Storage-Efficient Visual Document Retrieval: An Empirical Study on Reducing Patch-Level Embeddings
von: Ma, Yubo, et al.
Veröffentlicht: (2025)
von: Ma, Yubo, et al.
Veröffentlicht: (2025)
Compressing Long Context for Enhancing RAG with AMR-based Concept Distillation
von: Shi, Kaize, et al.
Veröffentlicht: (2024)
von: Shi, Kaize, et al.
Veröffentlicht: (2024)
Summary of a Haystack: A Challenge to Long-Context LLMs and RAG Systems
von: Laban, Philippe, et al.
Veröffentlicht: (2024)
von: Laban, Philippe, et al.
Veröffentlicht: (2024)
LongRAG: Enhancing Retrieval-Augmented Generation with Long-context LLMs
von: Jiang, Ziyan, et al.
Veröffentlicht: (2024)
von: Jiang, Ziyan, et al.
Veröffentlicht: (2024)
On Context Utilization in Summarization with Large Language Models
von: Ravaut, Mathieu, et al.
Veröffentlicht: (2023)
von: Ravaut, Mathieu, et al.
Veröffentlicht: (2023)
Long$^2$RAG: Evaluating Long-Context & Long-Form Retrieval-Augmented Generation with Key Point Recall
von: Qi, Zehan, et al.
Veröffentlicht: (2024)
von: Qi, Zehan, et al.
Veröffentlicht: (2024)
SciAgent: Tool-augmented Language Models for Scientific Reasoning
von: Ma, Yubo, et al.
Veröffentlicht: (2024)
von: Ma, Yubo, et al.
Veröffentlicht: (2024)
On the Role of Discreteness in Diffusion LLMs
von: Jin, Ziqi, et al.
Veröffentlicht: (2025)
von: Jin, Ziqi, et al.
Veröffentlicht: (2025)
SimulRAG: Simulator-based RAG for Grounding LLMs in Long-form Scientific QA
von: Xu, Haozhou, et al.
Veröffentlicht: (2025)
von: Xu, Haozhou, et al.
Veröffentlicht: (2025)
Navigating the Nuances: A Fine-grained Evaluation of Vision-Language Navigation
von: Wang, Zehao, et al.
Veröffentlicht: (2024)
von: Wang, Zehao, et al.
Veröffentlicht: (2024)
U-NIAH: Unified RAG and LLM Evaluation for Long Context Needle-In-A-Haystack
von: Gao, Yunfan, et al.
Veröffentlicht: (2025)
von: Gao, Yunfan, et al.
Veröffentlicht: (2025)
In Defense of RAG in the Era of Long-Context Language Models
von: Yu, Tan, et al.
Veröffentlicht: (2024)
von: Yu, Tan, et al.
Veröffentlicht: (2024)
Revisiting In-Context Learning with Long Context Language Models
von: Baek, Jinheon, et al.
Veröffentlicht: (2024)
von: Baek, Jinheon, et al.
Veröffentlicht: (2024)
Analyzing Temporal Complex Events with Large Language Models? A Benchmark towards Temporal, Long Context Understanding
von: Zhang, Zhihan, et al.
Veröffentlicht: (2024)
von: Zhang, Zhihan, et al.
Veröffentlicht: (2024)
Toward Generalizable Evaluation in the LLM Era: A Survey Beyond Benchmarks
von: Cao, Yixin, et al.
Veröffentlicht: (2025)
von: Cao, Yixin, et al.
Veröffentlicht: (2025)
UncertaintyRAG: Span-Level Uncertainty Enhanced Long-Context Modeling for Retrieval-Augmented Generation
von: Li, Zixuan, et al.
Veröffentlicht: (2024)
von: Li, Zixuan, et al.
Veröffentlicht: (2024)
LaRA: Benchmarking Retrieval-Augmented Generation and Long-Context LLMs -- No Silver Bullet for LC or RAG Routing
von: Li, Kuan, et al.
Veröffentlicht: (2025)
von: Li, Kuan, et al.
Veröffentlicht: (2025)
Does RAG Really Perform Bad For Long-Context Processing?
von: Luo, Kun, et al.
Veröffentlicht: (2025)
von: Luo, Kun, et al.
Veröffentlicht: (2025)
Contradiction Detection in RAG Systems: Evaluating LLMs as Context Validators for Improved Information Consistency
von: Gokul, Vignesh, et al.
Veröffentlicht: (2025)
von: Gokul, Vignesh, et al.
Veröffentlicht: (2025)
ECoRAG: Evidentiality-guided Compression for Long Context RAG
von: Jeong, Yeonseok, et al.
Veröffentlicht: (2025)
von: Jeong, Yeonseok, et al.
Veröffentlicht: (2025)
Systematic Evaluation of Long-Context LLMs on Financial Concepts
von: Gupta, Lavanya, et al.
Veröffentlicht: (2024)
von: Gupta, Lavanya, et al.
Veröffentlicht: (2024)
UIO-LLMs: Unbiased Incremental Optimization for Long-Context LLMs
von: Li, Wenhao, et al.
Veröffentlicht: (2024)
von: Li, Wenhao, et al.
Veröffentlicht: (2024)
Beyond the Context Window: A Cost-Performance Analysis of Fact-Based Memory vs. Long-Context LLMs for Persistent Agents
von: Pollertlam, Natchanon, et al.
Veröffentlicht: (2026)
von: Pollertlam, Natchanon, et al.
Veröffentlicht: (2026)
The Mirage of Model Editing: Revisiting Evaluation in the Wild
von: Yang, Wanli, et al.
Veröffentlicht: (2025)
von: Yang, Wanli, et al.
Veröffentlicht: (2025)
Long Context RAG Performance of Large Language Models
von: Leng, Quinn, et al.
Veröffentlicht: (2024)
von: Leng, Quinn, et al.
Veröffentlicht: (2024)
LongFaith: Enhancing Long-Context Reasoning in LLMs with Faithful Synthetic Data
von: Yang, Cehao, et al.
Veröffentlicht: (2025)
von: Yang, Cehao, et al.
Veröffentlicht: (2025)
CURIE: Evaluating LLMs On Multitask Scientific Long Context Understanding and Reasoning
von: Cui, Hao, et al.
Veröffentlicht: (2025)
von: Cui, Hao, et al.
Veröffentlicht: (2025)
Hit-RAG: Learning to Reason with Long Contexts via Preference Alignment
von: Liu, Junming, et al.
Veröffentlicht: (2026)
von: Liu, Junming, et al.
Veröffentlicht: (2026)
Gavel: Agent Meets Checklist for Evaluating LLMs on Long-Context Legal Summarization
von: Dou, Yao, et al.
Veröffentlicht: (2026)
von: Dou, Yao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Towards Verifiable Generation: A Benchmark for Knowledge-aware Language Model Attribution
von: Li, Xinze, et al.
Veröffentlicht: (2023) -
Skill-as-Pseudocode: Refactoring Skill Libraries to Pseudocode for LLM Agents
von: Li, Xinze, et al.
Veröffentlicht: (2026) -
EMemBench: Interactive Benchmarking of Episodic Memory for VLM Agents
von: Li, Xinze, et al.
Veröffentlicht: (2026) -
Large Language Model Is Not a Good Few-shot Information Extractor, but a Good Reranker for Hard Samples!
von: Ma, Yubo, et al.
Veröffentlicht: (2023) -
Route Before Retrieve: Activating Latent Routing Abilities of LLMs for RAG vs. Long-Context Selection
von: Chen, Yiwen, et al.
Veröffentlicht: (2026)