PRGB Benchmark: A Robust Placeholder-Assisted Algorithm for Benchmarking Retrieval-Augmented Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Tan, Zhehao, Jiao, Yihan, Yang, Dan, Liu, Lei, Feng, Jie, Sun, Duolin, Shen, Yue, Wang, Jian, Wei, Peng, Gu, Jinjie |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
HANRAG: Heuristic Accurate Noise-resistant Retrieval-Augmented Generation for Multi-hop Question Answering
por: Sun, Duolin, et al.
Publicado: (2025)
por: Sun, Duolin, et al.
Publicado: (2025)
CTRL-RAG: Contrastive Likelihood Reward Based Reinforcement Learning for Context-Faithful RAG Models
por: Tan, Zhehao, et al.
Publicado: (2026)
por: Tan, Zhehao, et al.
Publicado: (2026)
WebClipper: Efficient Evolution of Web Agents with Graph-based Trajectory Pruning
por: Wang, Junjie, et al.
Publicado: (2026)
por: Wang, Junjie, et al.
Publicado: (2026)
GroupRank: A Groupwise Paradigm for Effective and Efficient Passage Reranking with LLMs
por: Long, Meixiu, et al.
Publicado: (2025)
por: Long, Meixiu, et al.
Publicado: (2025)
HIRAG: Hierarchical-Thought Instruction-Tuning Retrieval-Augmented Generation
por: Jiao, YiHan, et al.
Publicado: (2025)
por: Jiao, YiHan, et al.
Publicado: (2025)
DIVER: A Multi-Stage Approach for Reasoning-intensive Information Retrieval
por: Sun, Duolin, et al.
Publicado: (2025)
por: Sun, Duolin, et al.
Publicado: (2025)
Benchmarking Retrieval-Augmented Generation for Chemistry
por: Zhong, Xianrui, et al.
Publicado: (2025)
por: Zhong, Xianrui, et al.
Publicado: (2025)
LiveClin: A Live Clinical Benchmark without Leakage
por: Wang, Xidong, et al.
Publicado: (2026)
por: Wang, Xidong, et al.
Publicado: (2026)
Similarity is Not All You Need: Endowing Retrieval Augmented Generation with Multi Layered Thoughts
por: Gan, Chunjing, et al.
Publicado: (2024)
por: Gan, Chunjing, et al.
Publicado: (2024)
When Good OCR Is Not Enough: Benchmarking OCR Robustness for Retrieval-Augmented Generation
por: Sun, Lin, et al.
Publicado: (2026)
por: Sun, Lin, et al.
Publicado: (2026)
Bi'an: A Bilingual Benchmark and Model for Hallucination Detection in Retrieval-Augmented Generation
por: Jiang, Zhouyu, et al.
Publicado: (2025)
por: Jiang, Zhouyu, et al.
Publicado: (2025)
Benchmarking Retrieval-Augmented Generation for Medicine
por: Xiong, Guangzhi, et al.
Publicado: (2024)
por: Xiong, Guangzhi, et al.
Publicado: (2024)
QE-RAG: A Robust Retrieval-Augmented Generation Benchmark for Query Entry Errors
por: Zhang, Kepu, et al.
Publicado: (2025)
por: Zhang, Kepu, et al.
Publicado: (2025)
Benchmarking Retrieval-Augmented Generation in Multi-Modal Contexts
por: Liu, Zhenghao, et al.
Publicado: (2025)
por: Liu, Zhenghao, et al.
Publicado: (2025)
medinajorge/Marine-species-identification: Placeholder
por: Jorge Medina Hernández
Publicado: (2025)
por: Jorge Medina Hernández
Publicado: (2025)
FoRAG: Factuality-optimized Retrieval Augmented Generation for Web-enhanced Long-form Question Answering
por: Cai, Tianchi, et al.
Publicado: (2024)
por: Cai, Tianchi, et al.
Publicado: (2024)
Know Your Needs Better: Towards Structured Understanding of Marketer Demands with Analogical Reasoning Augmented LLMs
por: Wang, Junjie, et al.
Publicado: (2024)
por: Wang, Junjie, et al.
Publicado: (2024)
Benchmarking Knowledge-Extraction Attack and Defense on Retrieval-Augmented Generation
por: Qi, Zhisheng, et al.
Publicado: (2026)
por: Qi, Zhisheng, et al.
Publicado: (2026)
Multi-Source Knowledge Pruning for Retrieval-Augmented Generation: A Benchmark and Empirical Study
por: Yu, Shuo, et al.
Publicado: (2024)
por: Yu, Shuo, et al.
Publicado: (2024)
CFVBench: A Comprehensive Video Benchmark for Fine-grained Multimodal Retrieval-Augmented Generation
por: Wei, Kaiwen, et al.
Publicado: (2025)
por: Wei, Kaiwen, et al.
Publicado: (2025)
RAGPerf: An End-to-End Benchmarking Framework for Retrieval-Augmented Generation Systems
por: Li, Shaobo, et al.
Publicado: (2026)
por: Li, Shaobo, et al.
Publicado: (2026)
Language Model Evolutionary Algorithms for Recommender Systems: Benchmarks and Algorithm Comparisons
por: Liu, Jiao, et al.
Publicado: (2024)
por: Liu, Jiao, et al.
Publicado: (2024)
XRAG: eXamining the Core -- Benchmarking Foundational Components in Advanced Retrieval-Augmented Generation
por: Mao, Qianren, et al.
Publicado: (2024)
por: Mao, Qianren, et al.
Publicado: (2024)
MRAG: Benchmarking Retrieval-Augmented Generation for Bio-medicine
por: Li, Liz, et al.
Publicado: (2026)
por: Li, Liz, et al.
Publicado: (2026)
Benchmarking Poisoning Attacks against Retrieval-Augmented Generation
por: Zhang, Baolei, et al.
Publicado: (2025)
por: Zhang, Baolei, et al.
Publicado: (2025)
RAGBench: Explainable Benchmark for Retrieval-Augmented Generation Systems
por: Friel, Robert, et al.
Publicado: (2024)
por: Friel, Robert, et al.
Publicado: (2024)
BERGEN: A Benchmarking Library for Retrieval-Augmented Generation
por: Rau, David, et al.
Publicado: (2024)
por: Rau, David, et al.
Publicado: (2024)
Learning to Plan for Retrieval-Augmented Large Language Models from Knowledge Graphs
por: Wang, Junjie, et al.
Publicado: (2024)
por: Wang, Junjie, et al.
Publicado: (2024)
POLYRAG: Integrating Polyviews into Retrieval-Augmented Generation for Medical Applications
por: Gan, Chunjing, et al.
Publicado: (2025)
por: Gan, Chunjing, et al.
Publicado: (2025)
Towards Automated Smart Contract Generation: Evaluation, Benchmarking, and Retrieval-Augmented Repair
por: Chen, Zaoyu, et al.
Publicado: (2025)
por: Chen, Zaoyu, et al.
Publicado: (2025)
Multilingual Retrieval Augmented Generation for Culturally-Sensitive Tasks: A Benchmark for Cross-lingual Robustness
por: Li, Bryan, et al.
Publicado: (2024)
por: Li, Bryan, et al.
Publicado: (2024)
Pulse Shape Discrimination Algorithms: Survey and Benchmark
por: Liu, Haoran, et al.
Publicado: (2025)
por: Liu, Haoran, et al.
Publicado: (2025)
Legal-DC: Benchmarking Retrieval-Augmented Generation for Legal Documents
por: Li, Yaocong, et al.
Publicado: (2026)
por: Li, Yaocong, et al.
Publicado: (2026)
Benchmarking Retrieval-Augmented Multimodal Generation for Document Question Answering
por: Dong, Kuicai, et al.
Publicado: (2025)
por: Dong, Kuicai, et al.
Publicado: (2025)
Retrieval-Augmented Generation for Service Discovery: Chunking Strategies and Benchmarking
por: Pesl, Robin D., et al.
Publicado: (2025)
por: Pesl, Robin D., et al.
Publicado: (2025)
CORAL: Benchmarking Multi-turn Conversational Retrieval-Augmentation Generation
por: Cheng, Yiruo, et al.
Publicado: (2024)
por: Cheng, Yiruo, et al.
Publicado: (2024)
Retrieving Any Relevant Moments: Benchmark and Models for Generalized Moment Retrieval
por: Ding, Yiming, et al.
Publicado: (2026)
por: Ding, Yiming, et al.
Publicado: (2026)
HoH: A Dynamic Benchmark for Evaluating the Impact of Outdated Information on Retrieval-Augmented Generation
por: Ouyang, Jie, et al.
Publicado: (2025)
por: Ouyang, Jie, et al.
Publicado: (2025)
LibRec: Benchmarking Retrieval-Augmented LLMs for Library Migration Recommendations
por: Han, Junxiao, et al.
Publicado: (2025)
por: Han, Junxiao, et al.
Publicado: (2025)
Benchmarking Retrieval Strategies for Biomedical Retrieval-Augmented Generation: A Controlled Empirical Study
por: Bal, Devi Prasad, et al.
Publicado: (2026)
por: Bal, Devi Prasad, et al.
Publicado: (2026)
Ejemplares similares
-
HANRAG: Heuristic Accurate Noise-resistant Retrieval-Augmented Generation for Multi-hop Question Answering
por: Sun, Duolin, et al.
Publicado: (2025) -
CTRL-RAG: Contrastive Likelihood Reward Based Reinforcement Learning for Context-Faithful RAG Models
por: Tan, Zhehao, et al.
Publicado: (2026) -
WebClipper: Efficient Evolution of Web Agents with Graph-based Trajectory Pruning
por: Wang, Junjie, et al.
Publicado: (2026) -
GroupRank: A Groupwise Paradigm for Effective and Efficient Passage Reranking with LLMs
por: Long, Meixiu, et al.
Publicado: (2025) -
HIRAG: Hierarchical-Thought Instruction-Tuning Retrieval-Augmented Generation
por: Jiao, YiHan, et al.
Publicado: (2025)