Fast and Faithful: Real-Time Verification for Long-Document Retrieval-Augmented Generation Systems
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Xunzhuo, He, Bowei, Liu, Xue, Zhang, Haichen, Chen, Huamin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Knowledge Access Beats Model Size: Memory Augmented Routing for Persistent AI Agents
von: Liu, Xunzhuo, et al.
Veröffentlicht: (2026)
von: Liu, Xunzhuo, et al.
Veröffentlicht: (2026)
RAC: Retrieval-Augmented Clarification for Faithful Conversational Search
von: Kebir, Ahmed Rayane, et al.
Veröffentlicht: (2026)
von: Kebir, Ahmed Rayane, et al.
Veröffentlicht: (2026)
OkraLong: A Flexible Retrieval-Augmented Framework for Long-Text Query Processing
von: Hui, Yulong, et al.
Veröffentlicht: (2025)
von: Hui, Yulong, et al.
Veröffentlicht: (2025)
GeAR: Generation Augmented Retrieval
von: Liu, Haoyu, et al.
Veröffentlicht: (2025)
von: Liu, Haoyu, et al.
Veröffentlicht: (2025)
Dual-Pool Token-Budget Routing for Cost-Efficient and Reliable LLM Serving
von: Liu, Xunzhuo, et al.
Veröffentlicht: (2026)
von: Liu, Xunzhuo, et al.
Veröffentlicht: (2026)
98$\times$ Faster LLM Routing Without a Dedicated GPU: Flash Attention, Prompt Compression, and Near-Streaming for the vLLM Semantic Router
von: Liu, Xunzhuo, et al.
Veröffentlicht: (2026)
von: Liu, Xunzhuo, et al.
Veröffentlicht: (2026)
TableRAG: A Retrieval Augmented Generation Framework for Heterogeneous Document Reasoning
von: Yu, Xiaohan, et al.
Veröffentlicht: (2025)
von: Yu, Xiaohan, et al.
Veröffentlicht: (2025)
Dynamic and Parametric Retrieval-Augmented Generation
von: Su, Weihang, et al.
Veröffentlicht: (2025)
von: Su, Weihang, et al.
Veröffentlicht: (2025)
Parametric Retrieval Augmented Generation
von: Su, Weihang, et al.
Veröffentlicht: (2025)
von: Su, Weihang, et al.
Veröffentlicht: (2025)
Trustworthiness in Retrieval-Augmented Generation Systems: A Survey
von: Zhou, Yujia, et al.
Veröffentlicht: (2024)
von: Zhou, Yujia, et al.
Veröffentlicht: (2024)
Adaptive Retrieval-Augmented Generation for Conversational Systems
von: Wang, Xi, et al.
Veröffentlicht: (2024)
von: Wang, Xi, et al.
Veröffentlicht: (2024)
Adaptive Vision-Language Model Routing for Computer Use Agents
von: Liu, Xunzhuo, et al.
Veröffentlicht: (2026)
von: Liu, Xunzhuo, et al.
Veröffentlicht: (2026)
Visual Confused Deputy: Exploiting and Defending Perception Failures in Computer-Using Agents
von: Liu, Xunzhuo, et al.
Veröffentlicht: (2026)
von: Liu, Xunzhuo, et al.
Veröffentlicht: (2026)
Recall Them All: Retrieval-Augmented Language Models for Long Object List Extraction from Long Documents
von: Singhania, Sneha, et al.
Veröffentlicht: (2024)
von: Singhania, Sneha, et al.
Veröffentlicht: (2024)
Chain-of-Retrieval Augmented Generation
von: Wang, Liang, et al.
Veröffentlicht: (2025)
von: Wang, Liang, et al.
Veröffentlicht: (2025)
Generalizing Conversational Dense Retrieval via LLM-Cognition Data Augmentation
von: Chen, Haonan, et al.
Veröffentlicht: (2024)
von: Chen, Haonan, et al.
Veröffentlicht: (2024)
HopRAG: Multi-Hop Reasoning for Logic-Aware Retrieval-Augmented Generation
von: Liu, Hao, et al.
Veröffentlicht: (2025)
von: Liu, Hao, et al.
Veröffentlicht: (2025)
Investigating Retrieval-Augmented Generation Systems on Unanswerable, Uncheatable, Realistic, Multi-hop Queries
von: Liu, Gabrielle Kaili-May, et al.
Veröffentlicht: (2025)
von: Liu, Gabrielle Kaili-May, et al.
Veröffentlicht: (2025)
Thought-Retriever: Don't Just Retrieve Raw Data, Retrieve Thoughts for Memory-Augmented Agentic Systems
von: Feng, Tao, et al.
Veröffentlicht: (2026)
von: Feng, Tao, et al.
Veröffentlicht: (2026)
RbFT: Robust Fine-tuning for Retrieval-Augmented Generation against Retrieval Defects
von: Tu, Yiteng, et al.
Veröffentlicht: (2025)
von: Tu, Yiteng, et al.
Veröffentlicht: (2025)
InfoDeepSeek: Benchmarking Agentic Information Seeking for Retrieval-Augmented Generation
von: Xi, Yunjia, et al.
Veröffentlicht: (2025)
von: Xi, Yunjia, et al.
Veröffentlicht: (2025)
LexRAG: Benchmarking Retrieval-Augmented Generation in Multi-Turn Legal Consultation Conversation
von: Li, Haitao, et al.
Veröffentlicht: (2025)
von: Li, Haitao, et al.
Veröffentlicht: (2025)
Evaluating Retrieval Quality in Retrieval-Augmented Generation
von: Salemi, Alireza, et al.
Veröffentlicht: (2024)
von: Salemi, Alireza, et al.
Veröffentlicht: (2024)
A MapReduce Approach to Effectively Utilize Long Context Information in Retrieval Augmented Language Models
von: Zhang, Gongbo, et al.
Veröffentlicht: (2024)
von: Zhang, Gongbo, et al.
Veröffentlicht: (2024)
+VeriRel: Verification Feedback to Enhance Document Retrieval for Scientific Fact Checking
von: Deng, Xingyu, et al.
Veröffentlicht: (2025)
von: Deng, Xingyu, et al.
Veröffentlicht: (2025)
TA-Mem: Tool-Augmented Autonomous Memory Retrieval for LLM in Long-Term Conversational QA
von: Yuan, Mengwei, et al.
Veröffentlicht: (2026)
von: Yuan, Mengwei, et al.
Veröffentlicht: (2026)
ArbGraph: Conflict-Aware Evidence Arbitration for Reliable Long-Form Retrieval-Augmented Generation
von: Niu, Qingying, et al.
Veröffentlicht: (2026)
von: Niu, Qingying, et al.
Veröffentlicht: (2026)
Benchmarking Retrieval-Augmented Multimodal Generation for Document Question Answering
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
Retrieval Augmented Generation with Collaborative Filtering for Personalized Text Generation
von: Shi, Teng, et al.
Veröffentlicht: (2025)
von: Shi, Teng, et al.
Veröffentlicht: (2025)
SQuAI: Scientific Question-Answering with Multi-Agent Retrieval-Augmented Generation
von: Besrour, Ines, et al.
Veröffentlicht: (2025)
von: Besrour, Ines, et al.
Veröffentlicht: (2025)
MAO-ARAG: Multi-Agent Orchestration for Adaptive Retrieval-Augmented Generation
von: Chen, Yiqun, et al.
Veröffentlicht: (2025)
von: Chen, Yiqun, et al.
Veröffentlicht: (2025)
Generate-then-Ground in Retrieval-Augmented Generation for Multi-hop Question Answering
von: Shi, Zhengliang, et al.
Veröffentlicht: (2024)
von: Shi, Zhengliang, et al.
Veröffentlicht: (2024)
Beyond Semantic Relevance: Counterfactual Risk Minimization for Robust Retrieval-Augmented Generation
von: Liu, Peiyang, et al.
Veröffentlicht: (2026)
von: Liu, Peiyang, et al.
Veröffentlicht: (2026)
FlexRAG: A Flexible and Comprehensive Framework for Retrieval-Augmented Generation
von: Zhang, Zhuocheng, et al.
Veröffentlicht: (2025)
von: Zhang, Zhuocheng, et al.
Veröffentlicht: (2025)
A Large Language Model-based Framework for Semi-Structured Tender Document Retrieval-Augmented Generation
von: Zhao, Yilong, et al.
Veröffentlicht: (2024)
von: Zhao, Yilong, et al.
Veröffentlicht: (2024)
Improving Retrieval-Augmented Generation through Multi-Agent Reinforcement Learning
von: Chen, Yiqun, et al.
Veröffentlicht: (2025)
von: Chen, Yiqun, et al.
Veröffentlicht: (2025)
Optimizing User Profiles via Contextual Bandits for Retrieval-Augmented LLM Personalization
von: Du, Linfeng, et al.
Veröffentlicht: (2026)
von: Du, Linfeng, et al.
Veröffentlicht: (2026)
DRAGIN: Dynamic Retrieval Augmented Generation based on the Information Needs of Large Language Models
von: Su, Weihang, et al.
Veröffentlicht: (2024)
von: Su, Weihang, et al.
Veröffentlicht: (2024)
Utility-Focused LLM Annotation for Retrieval and Retrieval-Augmented Generation
von: Zhang, Hengran, et al.
Veröffentlicht: (2025)
von: Zhang, Hengran, et al.
Veröffentlicht: (2025)
Exploring Selective Retrieval-Augmentation for Long-Tail Legal Text Classification
von: Mao, Boheng
Veröffentlicht: (2025)
von: Mao, Boheng
Veröffentlicht: (2025)
Ähnliche Einträge
-
Knowledge Access Beats Model Size: Memory Augmented Routing for Persistent AI Agents
von: Liu, Xunzhuo, et al.
Veröffentlicht: (2026) -
RAC: Retrieval-Augmented Clarification for Faithful Conversational Search
von: Kebir, Ahmed Rayane, et al.
Veröffentlicht: (2026) -
OkraLong: A Flexible Retrieval-Augmented Framework for Long-Text Query Processing
von: Hui, Yulong, et al.
Veröffentlicht: (2025) -
GeAR: Generation Augmented Retrieval
von: Liu, Haoyu, et al.
Veröffentlicht: (2025) -
Dual-Pool Token-Budget Routing for Cost-Efficient and Reliable LLM Serving
von: Liu, Xunzhuo, et al.
Veröffentlicht: (2026)