MCiteBench: A Multimodal Benchmark for Generating Text with Citations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hu, Caiyu, Zhang, Yikai, Zhu, Tinghui, Ye, Yiwei, Xiao, Yanghua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GEM-Bench: A Benchmark for Ad-Injected Response Generation within Generative Engine Marketing
von: Hu, Silan, et al.
Veröffentlicht: (2025)
von: Hu, Silan, et al.
Veröffentlicht: (2025)
Towards Fine-Grained Citation Evaluation in Generated Text: A Comparative Analysis of Faithfulness Metrics
von: Zhang, Weijia, et al.
Veröffentlicht: (2024)
von: Zhang, Weijia, et al.
Veröffentlicht: (2024)
AIR-Bench: Automated Heterogeneous Information Retrieval Benchmark
von: Chen, Jianlyu, et al.
Veröffentlicht: (2024)
von: Chen, Jianlyu, et al.
Veröffentlicht: (2024)
DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents
von: Du, Mingxuan, et al.
Veröffentlicht: (2025)
von: Du, Mingxuan, et al.
Veröffentlicht: (2025)
exHarmony: Authorship and Citations for Benchmarking the Reviewer Assignment Problem
von: Ebrahimi, Sajad, et al.
Veröffentlicht: (2025)
von: Ebrahimi, Sajad, et al.
Veröffentlicht: (2025)
FAB-Bench: A Framework for Adaptive RAG Benchmarking in Semiconductor Manufacturing
von: Qian, Jingbin, et al.
Veröffentlicht: (2026)
von: Qian, Jingbin, et al.
Veröffentlicht: (2026)
Diagnosing and Repairing Citation Failures in Generative Engine Optimization
von: Tian, Zhihua, et al.
Veröffentlicht: (2026)
von: Tian, Zhihua, et al.
Veröffentlicht: (2026)
ClarifyMT-Bench: Benchmarking and Improving Multi-Turn Clarification for Conversational Large Language Models
von: Luo, Sichun, et al.
Veröffentlicht: (2025)
von: Luo, Sichun, et al.
Veröffentlicht: (2025)
CiteBART: Learning to Generate Citations for Local Citation Recommendation
von: Çelik, Ege Yiğit, et al.
Veröffentlicht: (2024)
von: Çelik, Ege Yiğit, et al.
Veröffentlicht: (2024)
A Comparative Analysis of Faithfulness Metrics and Humans in Citation Evaluation
von: Zhang, Weijia, et al.
Veröffentlicht: (2024)
von: Zhang, Weijia, et al.
Veröffentlicht: (2024)
Retrieval Augmented Generation with Collaborative Filtering for Personalized Text Generation
von: Shi, Teng, et al.
Veröffentlicht: (2025)
von: Shi, Teng, et al.
Veröffentlicht: (2025)
On the Capacity of Citation Generation by Large Language Models
von: Qian, Haosheng, et al.
Veröffentlicht: (2024)
von: Qian, Haosheng, et al.
Veröffentlicht: (2024)
LITE: LLM-Impelled efficient Taxonomy Evaluation
von: Zhang, Lin, et al.
Veröffentlicht: (2025)
von: Zhang, Lin, et al.
Veröffentlicht: (2025)
Exposing Citation Vulnerabilities in Generative Engines
von: Mochizuki, Riku, et al.
Veröffentlicht: (2025)
von: Mochizuki, Riku, et al.
Veröffentlicht: (2025)
SMARTFinRAG: Interactive Modularized Financial RAG Benchmark
von: Zha, Yiwei
Veröffentlicht: (2025)
von: Zha, Yiwei
Veröffentlicht: (2025)
PMG : Personalized Multimodal Generation with Large Language Models
von: Shen, Xiaoteng, et al.
Veröffentlicht: (2024)
von: Shen, Xiaoteng, et al.
Veröffentlicht: (2024)
InfoDeepSeek: Benchmarking Agentic Information Seeking for Retrieval-Augmented Generation
von: Xi, Yunjia, et al.
Veröffentlicht: (2025)
von: Xi, Yunjia, et al.
Veröffentlicht: (2025)
JUÁ -- A Benchmark for Information Retrieval in Brazilian Legal Text Collections
von: Pereira, Jayr, et al.
Veröffentlicht: (2026)
von: Pereira, Jayr, et al.
Veröffentlicht: (2026)
FinMTEB: Finance Massive Text Embedding Benchmark
von: Tang, Yixuan, et al.
Veröffentlicht: (2025)
von: Tang, Yixuan, et al.
Veröffentlicht: (2025)
JFinTEB: Japanese Financial Text Embedding Benchmark
von: Suzuki, Masahiro, et al.
Veröffentlicht: (2026)
von: Suzuki, Masahiro, et al.
Veröffentlicht: (2026)
mGTE: Generalized Long-Context Text Representation and Reranking Models for Multilingual Text Retrieval
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
ILCiteR: Evidence-grounded Interpretable Local Citation Recommendation
von: Roy, Sayar Ghosh, et al.
Veröffentlicht: (2024)
von: Roy, Sayar Ghosh, et al.
Veröffentlicht: (2024)
Exploiting Duality in Open Information Extraction with Predicate Prompt
von: Chen, Zhen, et al.
Veröffentlicht: (2024)
von: Chen, Zhen, et al.
Veröffentlicht: (2024)
RAG-RewardBench: Benchmarking Reward Models in Retrieval Augmented Generation for Preference Alignment
von: Jin, Zhuoran, et al.
Veröffentlicht: (2024)
von: Jin, Zhuoran, et al.
Veröffentlicht: (2024)
Structured Attention Matters to Multimodal LLMs in Document Understanding
von: Liu, Chang, et al.
Veröffentlicht: (2025)
von: Liu, Chang, et al.
Veröffentlicht: (2025)
AlzheimerRAG: Multimodal Retrieval Augmented Generation for Clinical Use Cases using PubMed articles
von: Lahiri, Aritra Kumar, et al.
Veröffentlicht: (2024)
von: Lahiri, Aritra Kumar, et al.
Veröffentlicht: (2024)
MuRAR: A Simple and Effective Multimodal Retrieval and Answer Refinement Framework for Multimodal Question Answering
von: Zhu, Zhengyuan, et al.
Veröffentlicht: (2024)
von: Zhu, Zhengyuan, et al.
Veröffentlicht: (2024)
DomainRAG: A Chinese Benchmark for Evaluating Domain-specific Retrieval-Augmented Generation
von: Wang, Shuting, et al.
Veröffentlicht: (2024)
von: Wang, Shuting, et al.
Veröffentlicht: (2024)
BiCA: Effective Biomedical Dense Retrieval with Citation-Aware Hard Negatives
von: Sinha, Aarush, et al.
Veröffentlicht: (2025)
von: Sinha, Aarush, et al.
Veröffentlicht: (2025)
CiteFix: Enhancing RAG Accuracy Through Post-Processing Citation Correction
von: Maheshwari, Harsh, et al.
Veröffentlicht: (2025)
von: Maheshwari, Harsh, et al.
Veröffentlicht: (2025)
What Makes an Ideal Quote? Recommending "Unexpected yet Rational" Quotations via Novelty
von: Zhang, Bowei, et al.
Veröffentlicht: (2025)
von: Zhang, Bowei, et al.
Veröffentlicht: (2025)
Structurally Refined Graph Transformer for Multimodal Recommendation
von: Shi, Ke, et al.
Veröffentlicht: (2025)
von: Shi, Ke, et al.
Veröffentlicht: (2025)
BERGEN: A Benchmarking Library for Retrieval-Augmented Generation
von: Rau, David, et al.
Veröffentlicht: (2024)
von: Rau, David, et al.
Veröffentlicht: (2024)
SAGE: Benchmarking and Improving Retrieval for Deep Research Agents
von: Hu, Tiansheng, et al.
Veröffentlicht: (2026)
von: Hu, Tiansheng, et al.
Veröffentlicht: (2026)
Citation Recommendation based on Argumentative Zoning of User Queries
von: Ma, Shutian, et al.
Veröffentlicht: (2025)
von: Ma, Shutian, et al.
Veröffentlicht: (2025)
LegalAgentBench: Evaluating LLM Agents in Legal Domain
von: Li, Haitao, et al.
Veröffentlicht: (2024)
von: Li, Haitao, et al.
Veröffentlicht: (2024)
RAR-b: Reasoning as Retrieval Benchmark
von: Xiao, Chenghao, et al.
Veröffentlicht: (2024)
von: Xiao, Chenghao, et al.
Veröffentlicht: (2024)
GME: Improving Universal Multimodal Retrieval by Multimodal LLMs
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
Doc-Researcher: A Unified System for Multimodal Document Parsing and Deep Research
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
LexRAG: Benchmarking Retrieval-Augmented Generation in Multi-Turn Legal Consultation Conversation
von: Li, Haitao, et al.
Veröffentlicht: (2025)
von: Li, Haitao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
GEM-Bench: A Benchmark for Ad-Injected Response Generation within Generative Engine Marketing
von: Hu, Silan, et al.
Veröffentlicht: (2025) -
Towards Fine-Grained Citation Evaluation in Generated Text: A Comparative Analysis of Faithfulness Metrics
von: Zhang, Weijia, et al.
Veröffentlicht: (2024) -
AIR-Bench: Automated Heterogeneous Information Retrieval Benchmark
von: Chen, Jianlyu, et al.
Veröffentlicht: (2024) -
DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents
von: Du, Mingxuan, et al.
Veröffentlicht: (2025) -
exHarmony: Authorship and Citations for Benchmarking the Reviewer Assignment Problem
von: Ebrahimi, Sajad, et al.
Veröffentlicht: (2025)