Cocktail: A Comprehensive Information Retrieval Benchmark with LLM-Generated Documents Integration
Fuente:
arXiv
Saved in:
| Main Authors: | Dai, Sunhao, Liu, Weihao, Zhou, Yuqi, Pang, Liang, Ruan, Rongju, Wang, Gang, Dong, Zhenhua, Xu, Jun, Wen, Ji-Rong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Exploring the Escalation of Source Bias in User, Data, and Recommender System Feedback Loop
by: Zhou, Yuqi, et al.
Published: (2024)
by: Zhou, Yuqi, et al.
Published: (2024)
Learning to Retrieve from Agent Trajectories
by: Zhou, Yuqi, et al.
Published: (2026)
by: Zhou, Yuqi, et al.
Published: (2026)
Bias and Unfairness in Information Retrieval Systems: New Challenges in the LLM Era
by: Dai, Sunhao, et al.
Published: (2024)
by: Dai, Sunhao, et al.
Published: (2024)
Perplexity Trap: PLM-Based Retrievers Overrate Low Perplexity Documents
by: Wang, Haoyu, et al.
Published: (2025)
by: Wang, Haoyu, et al.
Published: (2025)
Neural Retrievers are Biased Towards LLM-Generated Content
by: Dai, Sunhao, et al.
Published: (2023)
by: Dai, Sunhao, et al.
Published: (2023)
Towards Completeness-Oriented Tool Retrieval for Large Language Models
by: Qu, Changle, et al.
Published: (2024)
by: Qu, Changle, et al.
Published: (2024)
Length-Induced Embedding Collapse in PLM-based Models
by: Zhou, Yuqi, et al.
Published: (2024)
by: Zhou, Yuqi, et al.
Published: (2024)
NExT-Search: Rebuilding User Feedback Ecosystem for Generative AI Search
by: Dai, Sunhao, et al.
Published: (2025)
by: Dai, Sunhao, et al.
Published: (2025)
CreAgent: Towards Long-Term Evaluation of Recommender System under Platform-Creator Information Asymmetry
by: Ye, Xiaopeng, et al.
Published: (2025)
by: Ye, Xiaopeng, et al.
Published: (2025)
FairDiverse: A Comprehensive Toolkit for Fair and Diverse Information Retrieval Algorithms
by: Xu, Chen, et al.
Published: (2025)
by: Xu, Chen, et al.
Published: (2025)
Code-Switching Information Retrieval: Benchmarks, Analysis, and the Limits of Current Retrievers
by: Zeng, Qingcheng, et al.
Published: (2026)
by: Zeng, Qingcheng, et al.
Published: (2026)
CoIR: A Comprehensive Benchmark for Code Information Retrieval Models
by: Li, Xiangyang, et al.
Published: (2024)
by: Li, Xiangyang, et al.
Published: (2024)
Counteracting Duration Bias in Video Recommendation via Counterfactual Watch Time
by: Zhao, Haiyuan, et al.
Published: (2024)
by: Zhao, Haiyuan, et al.
Published: (2024)
Unbiased Top-k Learning to Rank with Causal Likelihood Decomposition
by: Zhao, Haiyuan, et al.
Published: (2022)
by: Zhao, Haiyuan, et al.
Published: (2022)
RecCocktail: A Generalizable and Efficient Framework for LLM-Based Recommendation
by: Hou, Min, et al.
Published: (2025)
by: Hou, Min, et al.
Published: (2025)
DisastIR: A Comprehensive Information Retrieval Benchmark for Disaster Management
by: Yin, Kai, et al.
Published: (2025)
by: Yin, Kai, et al.
Published: (2025)
Beyond Chunk-Then-Embed: A Comprehensive Taxonomy and Evaluation of Document Chunking Strategies for Information Retrieval
by: Zhou, Yongjie, et al.
Published: (2026)
by: Zhou, Yongjie, et al.
Published: (2026)
CORAL: Benchmarking Multi-turn Conversational Retrieval-Augmentation Generation
by: Cheng, Yiruo, et al.
Published: (2024)
by: Cheng, Yiruo, et al.
Published: (2024)
ReCODE: Modeling Repeat Consumption with Neural ODE
by: Dai, Sunhao, et al.
Published: (2024)
by: Dai, Sunhao, et al.
Published: (2024)
Model Editing for New Document Integration in Generative Information Retrieval
by: Zhang, Zhen, et al.
Published: (2026)
by: Zhang, Zhen, et al.
Published: (2026)
Inference Computation Scaling for Feature Augmentation in Recommendation Systems
by: Liu, Weihao, et al.
Published: (2025)
by: Liu, Weihao, et al.
Published: (2025)
CART: A Generative Cross-Modal Retrieval Framework with Coarse-To-Fine Semantic Modeling
by: Fang, Minghui, et al.
Published: (2024)
by: Fang, Minghui, et al.
Published: (2024)
MIRA: An LLM-Assisted Benchmark for Multi-Category Integrated Retrieval
by: Türkmen, Mehmet Deniz, et al.
Published: (2026)
by: Türkmen, Mehmet Deniz, et al.
Published: (2026)
MemSifter: Offloading LLM Memory Retrieval via Outcome-Driven Proxy Reasoning
by: Tan, Jiejun, et al.
Published: (2026)
by: Tan, Jiejun, et al.
Published: (2026)
A Survey of Long-Document Retrieval in the PLM and LLM Era
by: Li, Minghan, et al.
Published: (2025)
by: Li, Minghan, et al.
Published: (2025)
MIRB: Mathematical Information Retrieval Benchmark
by: Ju, Haocheng, et al.
Published: (2025)
by: Ju, Haocheng, et al.
Published: (2025)
UOEP: User-Oriented Exploration Policy for Enhancing Long-Term User Experiences in Recommender Systems
by: Zhang, Changshuo, et al.
Published: (2024)
by: Zhang, Changshuo, et al.
Published: (2024)
Effective Knowledge Transfer for Multi-Task Recommendation Models
by: Cai, Guohao, et al.
Published: (2026)
by: Cai, Guohao, et al.
Published: (2026)
A Contextual-Aware Position Encoding for Sequential Recommendation
by: Yuan, Jun, et al.
Published: (2025)
by: Yuan, Jun, et al.
Published: (2025)
EAGER-LLM: Enhancing Large Language Models as Recommenders through Exogenous Behavior-Semantic Integration
by: Hong, Minjie, et al.
Published: (2025)
by: Hong, Minjie, et al.
Published: (2025)
DocRetriever: A Plug-and-Play Framework for Multimodal Document Retrieval with Comprehensive Benchmark
by: Hu, Ruofan, et al.
Published: (2026)
by: Hu, Ruofan, et al.
Published: (2026)
Regret-aware Re-ranking for Guaranteeing Two-sided Fairness and Accuracy in Recommender Systems
by: Ye, Xiaopeng, et al.
Published: (2025)
by: Ye, Xiaopeng, et al.
Published: (2025)
Guaranteeing Accuracy and Fairness under Fluctuating User Traffic: A Bankruptcy-Inspired Re-ranking Approach
by: Ye, Xiaopeng, et al.
Published: (2024)
by: Ye, Xiaopeng, et al.
Published: (2024)
ComLQ: Benchmarking Complex Logical Queries in Information Retrieval
by: Xu, Ganlin, et al.
Published: (2025)
by: Xu, Ganlin, et al.
Published: (2025)
FollowTable: A Benchmark for Instruction-Following Table Retrieval
by: Jin, Rihui, et al.
Published: (2026)
by: Jin, Rihui, et al.
Published: (2026)
HtmlRAG: HTML is Better Than Plain Text for Modeling Retrieved Knowledge in RAG Systems
by: Tan, Jiejun, et al.
Published: (2024)
by: Tan, Jiejun, et al.
Published: (2024)
IFIR: A Comprehensive Benchmark for Evaluating Instruction-Following in Expert-Domain Information Retrieval
by: Song, Tingyu, et al.
Published: (2025)
by: Song, Tingyu, et al.
Published: (2025)
LLM-Oriented Information Retrieval: A Denoising-First Perspective
by: Dai, Lu, et al.
Published: (2026)
by: Dai, Lu, et al.
Published: (2026)
Explicitly Integrating Judgment Prediction with Legal Document Retrieval: A Law-Guided Generative Approach
by: Qin, Weicong, et al.
Published: (2023)
by: Qin, Weicong, et al.
Published: (2023)
IRSC: A Zero-shot Evaluation Benchmark for Information Retrieval through Semantic Comprehension in Retrieval-Augmented Generation Scenarios
by: Lin, Hai, et al.
Published: (2024)
by: Lin, Hai, et al.
Published: (2024)
Similar Items
-
Exploring the Escalation of Source Bias in User, Data, and Recommender System Feedback Loop
by: Zhou, Yuqi, et al.
Published: (2024) -
Learning to Retrieve from Agent Trajectories
by: Zhou, Yuqi, et al.
Published: (2026) -
Bias and Unfairness in Information Retrieval Systems: New Challenges in the LLM Era
by: Dai, Sunhao, et al.
Published: (2024) -
Perplexity Trap: PLM-Based Retrievers Overrate Low Perplexity Documents
by: Wang, Haoyu, et al.
Published: (2025) -
Neural Retrievers are Biased Towards LLM-Generated Content
by: Dai, Sunhao, et al.
Published: (2023)