CoIR: A Comprehensive Benchmark for Code Information Retrieval Models
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Xiangyang, Dong, Kuicai, Lee, Yi Quan, Xia, Wei, Zhang, Hao, Dai, Xinyi, Wang, Yasheng, Tang, Ruiming |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Benchmarking Retrieval-Augmented Multimodal Generation for Document Question Answering
by: Dong, Kuicai, et al.
Published: (2025)
by: Dong, Kuicai, et al.
Published: (2025)
MMDocIR: Benchmarking Multimodal Retrieval for Long Documents
by: Dong, Kuicai, et al.
Published: (2025)
by: Dong, Kuicai, et al.
Published: (2025)
CtrlA: Adaptive Retrieval-Augmented Generation via Inherent Control
by: Liu, Huanshuo, et al.
Published: (2024)
by: Liu, Huanshuo, et al.
Published: (2024)
InfoDeepSeek: Benchmarking Agentic Information Seeking for Retrieval-Augmented Generation
by: Xi, Yunjia, et al.
Published: (2025)
by: Xi, Yunjia, et al.
Published: (2025)
LLMTreeRec: Unleashing the Power of Large Language Models for Cold-Start Recommendations
by: Zhang, Wenlin, et al.
Published: (2024)
by: Zhang, Wenlin, et al.
Published: (2024)
Cocktail: A Comprehensive Information Retrieval Benchmark with LLM-Generated Documents Integration
by: Dai, Sunhao, et al.
Published: (2024)
by: Dai, Sunhao, et al.
Published: (2024)
Collaborative Cross-modal Fusion with Large Language Model for Recommendation
by: Liu, Zhongzhou, et al.
Published: (2024)
by: Liu, Zhongzhou, et al.
Published: (2024)
FollowTable: A Benchmark for Instruction-Following Table Retrieval
by: Jin, Rihui, et al.
Published: (2026)
by: Jin, Rihui, et al.
Published: (2026)
PosIR: Position-Aware Heterogeneous Information Retrieval Benchmark
by: Zeng, Ziyang, et al.
Published: (2026)
by: Zeng, Ziyang, et al.
Published: (2026)
MassTool: A Multi-Task Search-Based Tool Retrieval Framework for Large Language Models
by: Lin, Jianghao, et al.
Published: (2025)
by: Lin, Jianghao, et al.
Published: (2025)
NevIR: Negation in Neural Information Retrieval
by: Weller, Orion, et al.
Published: (2023)
by: Weller, Orion, et al.
Published: (2023)
Large Language Models Make Sample-Efficient Recommender Systems
by: Lin, Jianghao, et al.
Published: (2024)
by: Lin, Jianghao, et al.
Published: (2024)
A Unified Framework for Multi-Domain CTR Prediction via Large Language Models
by: Fu, Zichuan, et al.
Published: (2023)
by: Fu, Zichuan, et al.
Published: (2023)
IFIR: A Comprehensive Benchmark for Evaluating Instruction-Following in Expert-Domain Information Retrieval
by: Song, Tingyu, et al.
Published: (2025)
by: Song, Tingyu, et al.
Published: (2025)
LLM4Tag: Automatic Tagging System for Information Retrieval via Large Language Models
by: Tang, Ruiming, et al.
Published: (2025)
by: Tang, Ruiming, et al.
Published: (2025)
DisastIR: A Comprehensive Information Retrieval Benchmark for Disaster Management
by: Yin, Kai, et al.
Published: (2025)
by: Yin, Kai, et al.
Published: (2025)
MIRB: Mathematical Information Retrieval Benchmark
by: Ju, Haocheng, et al.
Published: (2025)
by: Ju, Haocheng, et al.
Published: (2025)
AIR-Bench: Automated Heterogeneous Information Retrieval Benchmark
by: Chen, Jianlyu, et al.
Published: (2024)
by: Chen, Jianlyu, et al.
Published: (2024)
SRR-Judge: Step-Level Rating and Refinement for Enhancing Search-Integrated Reasoning in Search Agents
by: Zhang, Chen, et al.
Published: (2026)
by: Zhang, Chen, et al.
Published: (2026)
UniRetriever: Multi-task Candidates Selection for Various Context-Adaptive Conversational Retrieval
by: Wang, Hongru, et al.
Published: (2024)
by: Wang, Hongru, et al.
Published: (2024)
mFollowIR: a Multilingual Benchmark for Instruction Following in Retrieval
by: Weller, Orion, et al.
Published: (2025)
by: Weller, Orion, et al.
Published: (2025)
Building Russian Benchmark for Evaluation of Information Retrieval Models
by: Kovalev, Grigory, et al.
Published: (2025)
by: Kovalev, Grigory, et al.
Published: (2025)
CoRNStack: High-Quality Contrastive Data for Better Code Retrieval and Reranking
by: Suresh, Tarun, et al.
Published: (2024)
by: Suresh, Tarun, et al.
Published: (2024)
Improve Dense Passage Retrieval with Entailment Tuning
by: Dai, Lu, et al.
Published: (2024)
by: Dai, Lu, et al.
Published: (2024)
MiLQ: Benchmarking IR Models for Bilingual Web Search with Mixed Language Queries
by: Kim, Jonghwi, et al.
Published: (2025)
by: Kim, Jonghwi, et al.
Published: (2025)
When Should Dense Retrievers Be Updated in Evolving Corpora? Detecting Out-of-Distribution Corpora Using GradNormIR
by: Ko, Dayoon, et al.
Published: (2025)
by: Ko, Dayoon, et al.
Published: (2025)
Rethinking the Role of Token Retrieval in Multi-Vector Retrieval
by: Lee, Jinhyuk, et al.
Published: (2023)
by: Lee, Jinhyuk, et al.
Published: (2023)
Evaluating the External and Parametric Knowledge Fusion of Large Language Models
by: Zhang, Hao, et al.
Published: (2024)
by: Zhang, Hao, et al.
Published: (2024)
RVR: Retrieve-Verify-Retrieve for Comprehensive Question Answering
by: Qian, Deniz, et al.
Published: (2026)
by: Qian, Deniz, et al.
Published: (2026)
IR2: Information Regularization for Information Retrieval
by: Wang, Jianyou, et al.
Published: (2024)
by: Wang, Jianyou, et al.
Published: (2024)
Uncovering the Bigger Picture: Comprehensive Event Understanding Via Diverse News Retrieval
by: Tang, Yixuan, et al.
Published: (2025)
by: Tang, Yixuan, et al.
Published: (2025)
Process vs. Outcome Reward: Which is Better for Agentic RAG Reinforcement Learning
by: Zhang, Wenlin, et al.
Published: (2025)
by: Zhang, Wenlin, et al.
Published: (2025)
Exploring Recommender System Evaluation: A Multi-Modal User Agent Framework for A/B Testing
by: Zhang, Wenlin, et al.
Published: (2026)
by: Zhang, Wenlin, et al.
Published: (2026)
Wikipedia-based Datasets in Russian Information Retrieval Benchmark RusBEIR
by: Kovalev, Grigory, et al.
Published: (2025)
by: Kovalev, Grigory, et al.
Published: (2025)
JUÁ -- A Benchmark for Information Retrieval in Brazilian Legal Text Collections
by: Pereira, Jayr, et al.
Published: (2026)
by: Pereira, Jayr, et al.
Published: (2026)
FinMTEB: Finance Massive Text Embedding Benchmark
by: Tang, Yixuan, et al.
Published: (2025)
by: Tang, Yixuan, et al.
Published: (2025)
FollowIR: Evaluating and Teaching Information Retrieval Models to Follow Instructions
by: Weller, Orion, et al.
Published: (2024)
by: Weller, Orion, et al.
Published: (2024)
SyNeg: LLM-Driven Synthetic Hard-Negatives for Dense Retrieval
by: Li, Xiaopeng, et al.
Published: (2024)
by: Li, Xiaopeng, et al.
Published: (2024)
BRIGHT: A Realistic and Challenging Benchmark for Reasoning-Intensive Retrieval
by: Su, Hongjin, et al.
Published: (2024)
by: Su, Hongjin, et al.
Published: (2024)
Multi-Step Semantic Reasoning in Generative Retrieval
by: Dong, Steven, et al.
Published: (2026)
by: Dong, Steven, et al.
Published: (2026)
Similar Items
-
Benchmarking Retrieval-Augmented Multimodal Generation for Document Question Answering
by: Dong, Kuicai, et al.
Published: (2025) -
MMDocIR: Benchmarking Multimodal Retrieval for Long Documents
by: Dong, Kuicai, et al.
Published: (2025) -
CtrlA: Adaptive Retrieval-Augmented Generation via Inherent Control
by: Liu, Huanshuo, et al.
Published: (2024) -
InfoDeepSeek: Benchmarking Agentic Information Seeking for Retrieval-Augmented Generation
by: Xi, Yunjia, et al.
Published: (2025) -
LLMTreeRec: Unleashing the Power of Large Language Models for Cold-Start Recommendations
by: Zhang, Wenlin, et al.
Published: (2024)