IRB: Automated Generation of Robust Factuality Benchmarks
Fuente:
arXiv
Salvato in:
| Autori principali: | Do, Lam Thanh, Taleka, Bhagyashree, Bhutta, Hozaifa Ammar, Mailthody, Vikram Sharma, Chang, Kevin Chen-Chuan, Hwu, Wen-mei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CASPER: Concept-integrated Sparse Representation for Scientific Retrieval
di: Do, Lam Thanh, et al.
Pubblicazione: (2025)
di: Do, Lam Thanh, et al.
Pubblicazione: (2025)
AttentionRetriever: Attention Layers are Secretly Long Document Retrievers
di: Fu, David Jiahao, et al.
Pubblicazione: (2026)
di: Fu, David Jiahao, et al.
Pubblicazione: (2026)
Fact or Facsimile? Evaluating the Factual Robustness of Modern Retrievers
di: Wu, Haoyu, et al.
Pubblicazione: (2025)
di: Wu, Haoyu, et al.
Pubblicazione: (2025)
Evaluating Robustness of Generative Search Engine on Adversarial Factual Questions
di: Hu, Xuming, et al.
Pubblicazione: (2024)
di: Hu, Xuming, et al.
Pubblicazione: (2024)
Accelerating Sampling and Aggregation Operations in GNN Frameworks with GPU Initiated Direct Storage Accesses
di: Park, Jeongmin Brian, et al.
Pubblicazione: (2023)
di: Park, Jeongmin Brian, et al.
Pubblicazione: (2023)
Two-Stage Quranic QA via Ensemble Retrieval and Instruction-Tuned Answer Extraction
di: Basem, Mohamed, et al.
Pubblicazione: (2025)
di: Basem, Mohamed, et al.
Pubblicazione: (2025)
ConTReGen: Context-driven Tree-structured Retrieval for Open-domain Long-form Text Generation
di: Roy, Kashob Kumar, et al.
Pubblicazione: (2024)
di: Roy, Kashob Kumar, et al.
Pubblicazione: (2024)
BioPulse-QA: A Dynamic Biomedical Question-Answering Benchmark for Evaluating Factuality, Robustness, and Bias in Large Language Models
di: Bhattarai, Kriti, et al.
Pubblicazione: (2026)
di: Bhattarai, Kriti, et al.
Pubblicazione: (2026)
mmRAG: A Modular Benchmark for Retrieval-Augmented Generation over Text, Tables, and Knowledge Graphs
di: Xu, Chuan, et al.
Pubblicazione: (2025)
di: Xu, Chuan, et al.
Pubblicazione: (2025)
Can Users Detect Biases or Factual Errors in Generated Responses in Conversational Information-Seeking?
di: Łajewska, Weronika, et al.
Pubblicazione: (2024)
di: Łajewska, Weronika, et al.
Pubblicazione: (2024)
Enhancing Short-Text Topic Modeling with LLM-Driven Context Expansion and Prefix-Tuned VAEs
di: Akash, Pritom Saha, et al.
Pubblicazione: (2024)
di: Akash, Pritom Saha, et al.
Pubblicazione: (2024)
Structure Guided Retrieval-Augmented Generation for Factual Queries
di: Xie, Miao, et al.
Pubblicazione: (2026)
di: Xie, Miao, et al.
Pubblicazione: (2026)
On the Factual Consistency of Text-based Explainable Recommendation Models
di: Kabongo, Ben, et al.
Pubblicazione: (2025)
di: Kabongo, Ben, et al.
Pubblicazione: (2025)
QE-RAG: A Robust Retrieval-Augmented Generation Benchmark for Query Entry Errors
di: Zhang, Kepu, et al.
Pubblicazione: (2025)
di: Zhang, Kepu, et al.
Pubblicazione: (2025)
Retrieval-Augmented Generation: A Comprehensive Survey of Architectures, Enhancements, and Robustness Frontiers
di: Sharma, Chaitanya
Pubblicazione: (2025)
di: Sharma, Chaitanya
Pubblicazione: (2025)
Towards Dependable Retrieval-Augmented Generation Using Factual Confidence Prediction
di: Geissler, Florian, et al.
Pubblicazione: (2026)
di: Geissler, Florian, et al.
Pubblicazione: (2026)
Enhancing Health Information Retrieval with RAG by Prioritizing Topical Relevance and Factual Accuracy
di: Uapadhyay, Rishabh, et al.
Pubblicazione: (2025)
di: Uapadhyay, Rishabh, et al.
Pubblicazione: (2025)
RL-based Query Rewriting with Distilled LLM for online E-Commerce Systems
di: Nguyen, Duy A., et al.
Pubblicazione: (2025)
di: Nguyen, Duy A., et al.
Pubblicazione: (2025)
Investigating the Factual Knowledge Boundary of Large Language Models with Retrieval Augmentation
di: Ren, Ruiyang, et al.
Pubblicazione: (2023)
di: Ren, Ruiyang, et al.
Pubblicazione: (2023)
Context-Efficient Retrieval with Factual Decomposition
di: Li, Yanhong, et al.
Pubblicazione: (2025)
di: Li, Yanhong, et al.
Pubblicazione: (2025)
On Recommending Category: A Cascading Approach
di: Wang, Qihao, et al.
Pubblicazione: (2025)
di: Wang, Qihao, et al.
Pubblicazione: (2025)
Towards Reliable and Factual Response Generation: Detecting Unanswerable Questions in Information-Seeking Conversations
di: Łajewska, Weronika, et al.
Pubblicazione: (2024)
di: Łajewska, Weronika, et al.
Pubblicazione: (2024)
Automated Similarity Metric Generation for Recommendation
di: Qu, Liang, et al.
Pubblicazione: (2024)
di: Qu, Liang, et al.
Pubblicazione: (2024)
Efficient Item ID Generation for Large-Scale LLM-based Recommendation
di: Subbiah, Anushya, et al.
Pubblicazione: (2025)
di: Subbiah, Anushya, et al.
Pubblicazione: (2025)
Improved Estimation of Ranks for Learning Item Recommenders with Negative Sampling
di: Subbiah, Anushya, et al.
Pubblicazione: (2024)
di: Subbiah, Anushya, et al.
Pubblicazione: (2024)
Fix Before Search: Benchmarking Agentic Query Visual Pre-processing in Multimodal Retrieval-augmented Generation
di: Zhang, Jiankun, et al.
Pubblicazione: (2026)
di: Zhang, Jiankun, et al.
Pubblicazione: (2026)
Query Optimization for Parametric Knowledge Refinement in Retrieval-Augmented Large Language Models
di: Cong, Youan, et al.
Pubblicazione: (2024)
di: Cong, Youan, et al.
Pubblicazione: (2024)
RAGPerf: An End-to-End Benchmarking Framework for Retrieval-Augmented Generation Systems
di: Li, Shaobo, et al.
Pubblicazione: (2026)
di: Li, Shaobo, et al.
Pubblicazione: (2026)
Response Quality Assessment for Retrieval-Augmented Generation via Conditional Conformal Factuality
di: Feng, Naihe, et al.
Pubblicazione: (2025)
di: Feng, Naihe, et al.
Pubblicazione: (2025)
Automating Personalization: Prompt Optimization for Recommendation Reranking
di: Wang, Chen, et al.
Pubblicazione: (2025)
di: Wang, Chen, et al.
Pubblicazione: (2025)
Efficient Temporal-aware Matryoshka Adaptation for Temporal Information Retrieval
di: Huynh, Tuan-Luc, et al.
Pubblicazione: (2026)
di: Huynh, Tuan-Luc, et al.
Pubblicazione: (2026)
Token-Weighted Multi-Target Learning for Generative Recommenders with Curriculum Learning
di: Chiu, Wei-Ning, et al.
Pubblicazione: (2026)
di: Chiu, Wei-Ning, et al.
Pubblicazione: (2026)
On the Robustness of Generative Information Retrieval Models
di: Liu, Yu-An, et al.
Pubblicazione: (2024)
di: Liu, Yu-An, et al.
Pubblicazione: (2024)
Learning to Retrieve with Weakened Labels: Robust Training under Label Noise
di: Sharma, Arnab
Pubblicazione: (2025)
di: Sharma, Arnab
Pubblicazione: (2025)
SPAR: Session-based Pipeline for Adaptive Retrieval on Legacy File Systems
di: Nguyen, Duy A., et al.
Pubblicazione: (2025)
di: Nguyen, Duy A., et al.
Pubblicazione: (2025)
AMIR: Automated MisInformation Rebuttal -- A COVID-19 Vaccination Datasets based Recommendation System
di: Sharma, Shakshi, et al.
Pubblicazione: (2023)
di: Sharma, Shakshi, et al.
Pubblicazione: (2023)
Diagnosing Translated Benchmarks: An Automated Quality Assurance Study of the EU20 Benchmark Suite
di: Thellmann, Klaudia, et al.
Pubblicazione: (2026)
di: Thellmann, Klaudia, et al.
Pubblicazione: (2026)
Beyond Factual Grounding: The Case for Opinion-Aware Retrieval-Augmented Generation
di: Agrawal, Aditya, et al.
Pubblicazione: (2026)
di: Agrawal, Aditya, et al.
Pubblicazione: (2026)
AIR-Bench: Automated Heterogeneous Information Retrieval Benchmark
di: Chen, Jianlyu, et al.
Pubblicazione: (2024)
di: Chen, Jianlyu, et al.
Pubblicazione: (2024)
Adapting Standard Retrieval Benchmarks to Evaluate Generated Answers
di: Arabzadeh, Negar, et al.
Pubblicazione: (2024)
di: Arabzadeh, Negar, et al.
Pubblicazione: (2024)
Documenti analoghi
-
CASPER: Concept-integrated Sparse Representation for Scientific Retrieval
di: Do, Lam Thanh, et al.
Pubblicazione: (2025) -
AttentionRetriever: Attention Layers are Secretly Long Document Retrievers
di: Fu, David Jiahao, et al.
Pubblicazione: (2026) -
Fact or Facsimile? Evaluating the Factual Robustness of Modern Retrievers
di: Wu, Haoyu, et al.
Pubblicazione: (2025) -
Evaluating Robustness of Generative Search Engine on Adversarial Factual Questions
di: Hu, Xuming, et al.
Pubblicazione: (2024) -
Accelerating Sampling and Aggregation Operations in GNN Frameworks with GPU Initiated Direct Storage Accesses
di: Park, Jeongmin Brian, et al.
Pubblicazione: (2023)