Lighting the Way for BRIGHT: Reproducible Baselines with Anserini, Pyserini, and RankLLM
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sharifymoghaddam, Sahel, Ge, Yijun, Lin, Jimmy |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RankLLM: A Python Package for Reranking with LLMs
von: Sharifymoghaddam, Sahel, et al.
Veröffentlicht: (2025)
von: Sharifymoghaddam, Sahel, et al.
Veröffentlicht: (2025)
Rerank Before You Reason: Analyzing Reranking Tradeoffs through Effective Token Cost in Deep Search Agents
von: Sharifymoghaddam, Sahel, et al.
Veröffentlicht: (2026)
von: Sharifymoghaddam, Sahel, et al.
Veröffentlicht: (2026)
UniRAG: Universal Retrieval Augmentation for Large Vision Language Models
von: Sharifymoghaddam, Sahel, et al.
Veröffentlicht: (2024)
von: Sharifymoghaddam, Sahel, et al.
Veröffentlicht: (2024)
Chatbot Arena Meets Nuggets: Towards Explanations and Diagnostics in the Evaluation of LLM Responses
von: Sharifymoghaddam, Sahel, et al.
Veröffentlicht: (2025)
von: Sharifymoghaddam, Sahel, et al.
Veröffentlicht: (2025)
Ragnarök: A Reusable RAG Framework and Baselines for TREC 2024 Retrieval-Augmented Generation Track
von: Pradeep, Ronak, et al.
Veröffentlicht: (2024)
von: Pradeep, Ronak, et al.
Veröffentlicht: (2024)
Spacerini: Plug-and-play Search Engines with Pyserini and Hugging Face
von: Akiki, Christopher, et al.
Veröffentlicht: (2023)
von: Akiki, Christopher, et al.
Veröffentlicht: (2023)
MM-BRIGHT: A Multi-Task Multimodal Benchmark for Reasoning-Intensive Retrieval
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2026)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2026)
Illusions of Relevance: Arbitrary Content Injection Attacks Deceive Retrievers, Rerankers, and LLM Judges
von: Tamber, Manveer Singh, et al.
Veröffentlicht: (2025)
von: Tamber, Manveer Singh, et al.
Veröffentlicht: (2025)
Rank-R1: Enhancing Reasoning in LLM-based Document Rerankers via Reinforcement Learning
von: Zhuang, Shengyao, et al.
Veröffentlicht: (2025)
von: Zhuang, Shengyao, et al.
Veröffentlicht: (2025)
Operational Advice for Dense and Sparse Retrievers: HNSW, Flat, or Inverted Indexes?
von: Lin, Jimmy
Veröffentlicht: (2024)
von: Lin, Jimmy
Veröffentlicht: (2024)
Teaching Dense Retrieval Models to Specialize with Listwise Distillation and LLM Data Augmentation
von: Tamber, Manveer Singh, et al.
Veröffentlicht: (2025)
von: Tamber, Manveer Singh, et al.
Veröffentlicht: (2025)
RankSteer: Activation Steering for Pointwise LLM Ranking
von: Wang, Yumeng, et al.
Veröffentlicht: (2026)
von: Wang, Yumeng, et al.
Veröffentlicht: (2026)
Revisiting Feedback Models for HyDE
von: Jedidi, Nour, et al.
Veröffentlicht: (2025)
von: Jedidi, Nour, et al.
Veröffentlicht: (2025)
Unlearning for Federated Online Learning to Rank: A Reproducibility Study
von: Tao, Yiling, et al.
Veröffentlicht: (2025)
von: Tao, Yiling, et al.
Veröffentlicht: (2025)
Paraphrase Brittleness in Production Retrieval-Augmented Commercial Recommendation: Reproducibility Below the Rerun-Stability Baseline
von: Jack, Will, et al.
Veröffentlicht: (2026)
von: Jack, Will, et al.
Veröffentlicht: (2026)
BRIGHT: A Realistic and Challenging Benchmark for Reasoning-Intensive Retrieval
von: Su, Hongjin, et al.
Veröffentlicht: (2024)
von: Su, Hongjin, et al.
Veröffentlicht: (2024)
TRUE: A Reproducible Framework for LLM-Driven Relevance Judgment in Information Retrieval
von: Dewan, Mouly, et al.
Veröffentlicht: (2025)
von: Dewan, Mouly, et al.
Veröffentlicht: (2025)
BrowseComp-Plus: A More Fair and Transparent Evaluation Benchmark of Deep-Research Agent
von: Chen, Zijian, et al.
Veröffentlicht: (2025)
von: Chen, Zijian, et al.
Veröffentlicht: (2025)
The LLM Effect on IR Benchmarks: A Meta-Analysis of Effectiveness, Baselines, and Contamination
von: Staudinger, Moritz, et al.
Veröffentlicht: (2026)
von: Staudinger, Moritz, et al.
Veröffentlicht: (2026)
LLMs Can Patch Up Missing Relevance Judgments in Evaluation
von: Upadhyay, Shivani, et al.
Veröffentlicht: (2024)
von: Upadhyay, Shivani, et al.
Veröffentlicht: (2024)
A Reproducibility Study of LLM-Based Query Reformulation
von: Bigdeli, Amin, et al.
Veröffentlicht: (2026)
von: Bigdeli, Amin, et al.
Veröffentlicht: (2026)
Rank-DistiLLM: Closing the Effectiveness Gap Between Cross-Encoders and LLMs for Passage Re-Ranking
von: Schlatt, Ferdinand, et al.
Veröffentlicht: (2024)
von: Schlatt, Ferdinand, et al.
Veröffentlicht: (2024)
Can't Hide Behind the API: Stealing Black-Box Commercial Embedding Models
von: Tamber, Manveer Singh, et al.
Veröffentlicht: (2024)
von: Tamber, Manveer Singh, et al.
Veröffentlicht: (2024)
Musings About the Future of Search: A Return to the Past?
von: Lin, Jimmy, et al.
Veröffentlicht: (2024)
von: Lin, Jimmy, et al.
Veröffentlicht: (2024)
Batched Self-Consistency Improves LLM Relevance Assessment and Ranking
von: Korikov, Anton, et al.
Veröffentlicht: (2025)
von: Korikov, Anton, et al.
Veröffentlicht: (2025)
Can LLM Annotations Replace User Clicks for Learning to Rank?
von: Yu, Lulu, et al.
Veröffentlicht: (2025)
von: Yu, Lulu, et al.
Veröffentlicht: (2025)
REALM: Recursive Relevance Modeling for LLM-based Document Re-Ranking
von: Wang, Pinhuan, et al.
Veröffentlicht: (2025)
von: Wang, Pinhuan, et al.
Veröffentlicht: (2025)
TFRank: Think-Free Reasoning Enables Practical Pointwise LLM Ranking
von: Fan, Yongqi, et al.
Veröffentlicht: (2025)
von: Fan, Yongqi, et al.
Veröffentlicht: (2025)
A Systematic Study of Pseudo-Relevance Feedback with LLMs
von: Jedidi, Nour, et al.
Veröffentlicht: (2026)
von: Jedidi, Nour, et al.
Veröffentlicht: (2026)
QueryGym: A Toolkit for Reproducible LLM-Based Query Reformulation
von: Bigdeli, Amin, et al.
Veröffentlicht: (2025)
von: Bigdeli, Amin, et al.
Veröffentlicht: (2025)
Rank-GRPO: Training LLM-based Conversational Recommender Systems with Reinforcement Learning
von: Zhu, Yaochen, et al.
Veröffentlicht: (2025)
von: Zhu, Yaochen, et al.
Veröffentlicht: (2025)
SARM: LLM-Augmented Semantic Anchor for End-to-End Live-Streaming Ranking
von: Yang, Ruochen, et al.
Veröffentlicht: (2026)
von: Yang, Ruochen, et al.
Veröffentlicht: (2026)
Think When Needed: Model-Aware Reasoning Routing for LLM-based Ranking
von: Guo, Huizhong, et al.
Veröffentlicht: (2026)
von: Guo, Huizhong, et al.
Veröffentlicht: (2026)
LACONIC: Dense-Level Effectiveness for Scalable Sparse Retrieval via a Two-Phase Training Curriculum
von: Xu, Zhichao, et al.
Veröffentlicht: (2026)
von: Xu, Zhichao, et al.
Veröffentlicht: (2026)
Investigating the Robustness of Counterfactual Learning to Rank Models: A Reproducibility Study
von: Niu, Zechun, et al.
Veröffentlicht: (2024)
von: Niu, Zechun, et al.
Veröffentlicht: (2024)
A Reproducible Analysis of Sequential Recommender Systems
von: Betello, Filippo, et al.
Veröffentlicht: (2024)
von: Betello, Filippo, et al.
Veröffentlicht: (2024)
Reproducing and Comparing Distillation Techniques for Cross-Encoders
von: Morand, Victor, et al.
Veröffentlicht: (2026)
von: Morand, Victor, et al.
Veröffentlicht: (2026)
Reproducing Adaptive Reranking for Reasoning-Intensive IR
von: Rathee, Mandeep, et al.
Veröffentlicht: (2026)
von: Rathee, Mandeep, et al.
Veröffentlicht: (2026)
Unifying Multimodal Retrieval via Document Screenshot Embedding
von: Ma, Xueguang, et al.
Veröffentlicht: (2024)
von: Ma, Xueguang, et al.
Veröffentlicht: (2024)
The Ranking Blind Spot: Decision Hijacking in LLM-based Text Ranking
von: Qian, Yaoyao, et al.
Veröffentlicht: (2025)
von: Qian, Yaoyao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
RankLLM: A Python Package for Reranking with LLMs
von: Sharifymoghaddam, Sahel, et al.
Veröffentlicht: (2025) -
Rerank Before You Reason: Analyzing Reranking Tradeoffs through Effective Token Cost in Deep Search Agents
von: Sharifymoghaddam, Sahel, et al.
Veröffentlicht: (2026) -
UniRAG: Universal Retrieval Augmentation for Large Vision Language Models
von: Sharifymoghaddam, Sahel, et al.
Veröffentlicht: (2024) -
Chatbot Arena Meets Nuggets: Towards Explanations and Diagnostics in the Evaluation of LLM Responses
von: Sharifymoghaddam, Sahel, et al.
Veröffentlicht: (2025) -
Ragnarök: A Reusable RAG Framework and Baselines for TREC 2024 Retrieval-Augmented Generation Track
von: Pradeep, Ronak, et al.
Veröffentlicht: (2024)