DailyQA: A Benchmark to Evaluate Web Retrieval Augmented LLMs Based on Capturing Real-World Changes
Fuente:
arXiv
Saved in:
| Main Authors: | Cheng, Jiehan, Dou, Zhicheng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HierSearch: A Hierarchical Enterprise Deep Search Framework Integrating Local and Web Searches
by: Tan, Jiejun, et al.
Published: (2025)
by: Tan, Jiejun, et al.
Published: (2025)
WebQuest: A Benchmark for Multimodal QA on Web Page Sequences
by: Wang, Maria, et al.
Published: (2024)
by: Wang, Maria, et al.
Published: (2024)
ProRAG: Process-Supervised Reinforcement Learning for Retrieval-Augmented Generation
by: Wang, Zhao, et al.
Published: (2026)
by: Wang, Zhao, et al.
Published: (2026)
DomainRAG: A Chinese Benchmark for Evaluating Domain-specific Retrieval-Augmented Generation
by: Wang, Shuting, et al.
Published: (2024)
by: Wang, Shuting, et al.
Published: (2024)
Towards Mixed-Modal Retrieval for Universal Retrieval-Augmented Generation
by: Zhang, Chenghao, et al.
Published: (2025)
by: Zhang, Chenghao, et al.
Published: (2025)
UDA: A Benchmark Suite for Retrieval Augmented Generation in Real-world Document Analysis
by: Hui, Yulong, et al.
Published: (2024)
by: Hui, Yulong, et al.
Published: (2024)
S2G-RAG: Structured Sufficiency and Gap Judging for Iterative Retrieval-Augmented QA
by: Li, Minghan, et al.
Published: (2026)
by: Li, Minghan, et al.
Published: (2026)
MemSifter: Offloading LLM Memory Retrieval via Outcome-Driven Proxy Reasoning
by: Tan, Jiejun, et al.
Published: (2026)
by: Tan, Jiejun, et al.
Published: (2026)
IRSC: A Zero-shot Evaluation Benchmark for Information Retrieval through Semantic Comprehension in Retrieval-Augmented Generation Scenarios
by: Lin, Hai, et al.
Published: (2024)
by: Lin, Hai, et al.
Published: (2024)
Fine-Grained Guidance for Retrievers: Leveraging LLMs' Feedback in Retrieval-Augmented Generation
by: Liu, Yuhang, et al.
Published: (2024)
by: Liu, Yuhang, et al.
Published: (2024)
Comparative Analysis of Retrieval Systems in the Real World
by: Mozolevskyi, Dmytro, et al.
Published: (2024)
by: Mozolevskyi, Dmytro, et al.
Published: (2024)
Web Retrieval-Aware Chunking (W-RAC) for Efficient and Cost-Effective Retrieval-Augmented Generation Systems
by: Allu, Uday, et al.
Published: (2026)
by: Allu, Uday, et al.
Published: (2026)
Beyond Case Law: Evaluating Structure-Aware Retrieval and Safety in Statute-Centric Legal QA
by: Chae, Kyubyung, et al.
Published: (2026)
by: Chae, Kyubyung, et al.
Published: (2026)
RAGe: A Retrieval-Augmented Generation Evaluation Framework
by: Guder, Larissa, et al.
Published: (2026)
by: Guder, Larissa, et al.
Published: (2026)
Trustworthiness in Retrieval-Augmented Generation Systems: A Survey
by: Zhou, Yujia, et al.
Published: (2024)
by: Zhou, Yujia, et al.
Published: (2024)
Experience Retrieval-Augmentation with Electronic Health Records Enables Accurate Discharge QA
by: Ou, Justice, et al.
Published: (2025)
by: Ou, Justice, et al.
Published: (2025)
Web Retrieval Agents for Evidence-Based Misinformation Detection
by: Tian, Jacob-Junqi, et al.
Published: (2024)
by: Tian, Jacob-Junqi, et al.
Published: (2024)
Toward General Instruction-Following Alignment for Retrieval-Augmented Generation
by: Dong, Guanting, et al.
Published: (2024)
by: Dong, Guanting, et al.
Published: (2024)
WebFAQ 2.0: A Multilingual QA Dataset with Mined Hard Negatives for Dense Retrieval
by: Dinzinger, Michael, et al.
Published: (2026)
by: Dinzinger, Michael, et al.
Published: (2026)
StratRAG: A Multi-Hop Retrieval Evaluation Dataset for Retrieval-Augmented Generation Systems
by: Patodiya, Aryan
Published: (2026)
by: Patodiya, Aryan
Published: (2026)
Grounding Language Model with Chunking-Free In-Context Retrieval
by: Qian, Hongjin, et al.
Published: (2024)
by: Qian, Hongjin, et al.
Published: (2024)
Benchmarking Retrieval-Augmented Generation for Chemistry
by: Zhong, Xianrui, et al.
Published: (2025)
by: Zhong, Xianrui, et al.
Published: (2025)
Modeling Uncertainty and Using Post-fusion as Fallback Improves Retrieval Augmented Generation with LLMs
by: Liu, Ye, et al.
Published: (2023)
by: Liu, Ye, et al.
Published: (2023)
FRESCO: Benchmarking and Optimizing Re-rankers for Evolving Semantic Conflict in Retrieval-Augmented Generation
by: An, Sohyun, et al.
Published: (2026)
by: An, Sohyun, et al.
Published: (2026)
TrustRAG: An Information Assistant with Retrieval Augmented Generation
by: Fan, Yixing, et al.
Published: (2025)
by: Fan, Yixing, et al.
Published: (2025)
URAG: A Benchmark for Uncertainty Quantification in Retrieval-Augmented Large Language Models
by: Nguyen, Vinh, et al.
Published: (2026)
by: Nguyen, Vinh, et al.
Published: (2026)
A Multi-Task Embedder For Retrieval Augmented LLMs
by: Zhang, Peitian, et al.
Published: (2023)
by: Zhang, Peitian, et al.
Published: (2023)
Optimizing and Evaluating Enterprise Retrieval-Augmented Generation (RAG): A Content Design Perspective
by: Packowski, Sarah, et al.
Published: (2024)
by: Packowski, Sarah, et al.
Published: (2024)
Retrieval-Augmented Generation by Evidence Retroactivity in LLMs
by: Xiao, Liang, et al.
Published: (2025)
by: Xiao, Liang, et al.
Published: (2025)
WebNavigator: Global Web Navigation via Interaction Graph Retrieval
by: Zhang, Xuanwang, et al.
Published: (2026)
by: Zhang, Xuanwang, et al.
Published: (2026)
ItemRAG: Item-Based Retrieval-Augmented Generation for LLM-Based Recommendation
by: Kim, Sunwoo, et al.
Published: (2025)
by: Kim, Sunwoo, et al.
Published: (2025)
Using LLMs to Capture Users' Temporal Context for Recommendation
by: Sabouri, Milad, et al.
Published: (2025)
by: Sabouri, Milad, et al.
Published: (2025)
DICE: Discrete Interpretable Comparative Evaluation with Probabilistic Scoring for Retrieval-Augmented Generation
by: Liu, Shiyan, et al.
Published: (2025)
by: Liu, Shiyan, et al.
Published: (2025)
Evaluating Chunking Strategies For Retrieval-Augmented Generation in Oil and Gas Enterprise Documents
by: Taiwo, Samuel, et al.
Published: (2026)
by: Taiwo, Samuel, et al.
Published: (2026)
From Matching to Generation: A Survey on Generative Information Retrieval
by: Li, Xiaoxi, et al.
Published: (2024)
by: Li, Xiaoxi, et al.
Published: (2024)
OmniBench-RAG: A Multi-Domain Evaluation Platform for Retrieval-Augmented Generation Tools
by: Liang, Jiaxuan, et al.
Published: (2025)
by: Liang, Jiaxuan, et al.
Published: (2025)
AskNearby: An LLM-Based Application for Neighborhood Information Retrieval and Personalized Cognitive-Map Recommendations
by: Niu, Luyao, et al.
Published: (2025)
by: Niu, Luyao, et al.
Published: (2025)
Multi-Source Knowledge Pruning for Retrieval-Augmented Generation: A Benchmark and Empirical Study
by: Yu, Shuo, et al.
Published: (2024)
by: Yu, Shuo, et al.
Published: (2024)
Towards Adaptive Memory-Based Optimization for Enhanced Retrieval-Augmented Generation
by: Qin, Qitao, et al.
Published: (2025)
by: Qin, Qitao, et al.
Published: (2025)
HCT-QA: A Benchmark for Question Answering on Human-Centric Tables
by: Ahmad, Mohammad S., et al.
Published: (2025)
by: Ahmad, Mohammad S., et al.
Published: (2025)
Similar Items
-
HierSearch: A Hierarchical Enterprise Deep Search Framework Integrating Local and Web Searches
by: Tan, Jiejun, et al.
Published: (2025) -
WebQuest: A Benchmark for Multimodal QA on Web Page Sequences
by: Wang, Maria, et al.
Published: (2024) -
ProRAG: Process-Supervised Reinforcement Learning for Retrieval-Augmented Generation
by: Wang, Zhao, et al.
Published: (2026) -
DomainRAG: A Chinese Benchmark for Evaluating Domain-specific Retrieval-Augmented Generation
by: Wang, Shuting, et al.
Published: (2024) -
Towards Mixed-Modal Retrieval for Universal Retrieval-Augmented Generation
by: Zhang, Chenghao, et al.
Published: (2025)