Building Russian Benchmark for Evaluation of Information Retrieval Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kovalev, Grigory, Tikhomirov, Mikhail, Kozhevnikov, Evgeny, Kornilov, Max, Loukachevitch, Natalia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Wikipedia-based Datasets in Russian Information Retrieval Benchmark RusBEIR
von: Kovalev, Grigory, et al.
Veröffentlicht: (2025)
von: Kovalev, Grigory, et al.
Veröffentlicht: (2025)
Exploring Prompt-Based Methods for Zero-Shot Hypernym Prediction with Large Language Models
von: Tikhomirov, Mikhail, et al.
Veröffentlicht: (2024)
von: Tikhomirov, Mikhail, et al.
Veröffentlicht: (2024)
STELLA: Self-Reflective Terminology-Aware Framework for Building an Aerospace Information Retrieval Benchmark
von: Kim, Bongmin
Veröffentlicht: (2026)
von: Kim, Bongmin
Veröffentlicht: (2026)
Iterative Layer-wise Distillation for Efficient Compression of Large Language Models
von: Kovalev, Grigory, et al.
Veröffentlicht: (2025)
von: Kovalev, Grigory, et al.
Veröffentlicht: (2025)
Benchmarking and Building Zero-Shot Hindi Retrieval Model with Hindi-BEIR and NLLB-E5
von: Acharya, Arkadeep, et al.
Veröffentlicht: (2024)
von: Acharya, Arkadeep, et al.
Veröffentlicht: (2024)
FreshStack: Building Realistic Benchmarks for Evaluating Retrieval on Technical Documents
von: Thakur, Nandan, et al.
Veröffentlicht: (2025)
von: Thakur, Nandan, et al.
Veröffentlicht: (2025)
IFIR: A Comprehensive Benchmark for Evaluating Instruction-Following in Expert-Domain Information Retrieval
von: Song, Tingyu, et al.
Veröffentlicht: (2025)
von: Song, Tingyu, et al.
Veröffentlicht: (2025)
CoIR: A Comprehensive Benchmark for Code Information Retrieval Models
von: Li, Xiangyang, et al.
Veröffentlicht: (2024)
von: Li, Xiangyang, et al.
Veröffentlicht: (2024)
Benchmarking Information Retrieval Models on Complex Retrieval Tasks
von: Killingback, Julian, et al.
Veröffentlicht: (2025)
von: Killingback, Julian, et al.
Veröffentlicht: (2025)
AIR-Bench: Automated Heterogeneous Information Retrieval Benchmark
von: Chen, Jianlyu, et al.
Veröffentlicht: (2024)
von: Chen, Jianlyu, et al.
Veröffentlicht: (2024)
PosIR: Position-Aware Heterogeneous Information Retrieval Benchmark
von: Zeng, Ziyang, et al.
Veröffentlicht: (2026)
von: Zeng, Ziyang, et al.
Veröffentlicht: (2026)
Evaluating Generative Ad Hoc Information Retrieval
von: Gienapp, Lukas, et al.
Veröffentlicht: (2023)
von: Gienapp, Lukas, et al.
Veröffentlicht: (2023)
JUÁ -- A Benchmark for Information Retrieval in Brazilian Legal Text Collections
von: Pereira, Jayr, et al.
Veröffentlicht: (2026)
von: Pereira, Jayr, et al.
Veröffentlicht: (2026)
PersonalAI: A Systematic Comparison of Knowledge Graph Storage and Retrieval Approaches for Personalized LLM agents
von: Menschikov, Mikhail, et al.
Veröffentlicht: (2025)
von: Menschikov, Mikhail, et al.
Veröffentlicht: (2025)
InfoDeepSeek: Benchmarking Agentic Information Seeking for Retrieval-Augmented Generation
von: Xi, Yunjia, et al.
Veröffentlicht: (2025)
von: Xi, Yunjia, et al.
Veröffentlicht: (2025)
Cocktail: A Comprehensive Information Retrieval Benchmark with LLM-Generated Documents Integration
von: Dai, Sunhao, et al.
Veröffentlicht: (2024)
von: Dai, Sunhao, et al.
Veröffentlicht: (2024)
DomainRAG: A Chinese Benchmark for Evaluating Domain-specific Retrieval-Augmented Generation
von: Wang, Shuting, et al.
Veröffentlicht: (2024)
von: Wang, Shuting, et al.
Veröffentlicht: (2024)
DREditor: An Time-efficient Approach for Building a Domain-specific Dense Retrieval Model
von: Huang, Chen, et al.
Veröffentlicht: (2024)
von: Huang, Chen, et al.
Veröffentlicht: (2024)
MIRB: Mathematical Information Retrieval Benchmark
von: Ju, Haocheng, et al.
Veröffentlicht: (2025)
von: Ju, Haocheng, et al.
Veröffentlicht: (2025)
RAR-b: Reasoning as Retrieval Benchmark
von: Xiao, Chenghao, et al.
Veröffentlicht: (2024)
von: Xiao, Chenghao, et al.
Veröffentlicht: (2024)
Evaluating Large Language Models for Cross-Lingual Retrieval
von: Zuo, Longfei, et al.
Veröffentlicht: (2025)
von: Zuo, Longfei, et al.
Veröffentlicht: (2025)
Evaluating Retrieval Quality in Retrieval-Augmented Generation
von: Salemi, Alireza, et al.
Veröffentlicht: (2024)
von: Salemi, Alireza, et al.
Veröffentlicht: (2024)
Anveshana: A New Benchmark Dataset for Cross-Lingual Information Retrieval On English Queries and Sanskrit Documents
von: Jagadeeshan, Manoj Balaji, et al.
Veröffentlicht: (2025)
von: Jagadeeshan, Manoj Balaji, et al.
Veröffentlicht: (2025)
Large Language Models for Information Retrieval: A Survey
von: Zhu, Yutao, et al.
Veröffentlicht: (2023)
von: Zhu, Yutao, et al.
Veröffentlicht: (2023)
Still Fresh? Evaluating Temporal Drift in Retrieval Benchmarks
von: Kuissi, Nathan, et al.
Veröffentlicht: (2026)
von: Kuissi, Nathan, et al.
Veröffentlicht: (2026)
RuOpinionNE-2024: Extraction of Opinion Tuples from Russian News Texts
von: Loukachevitch, Natalia, et al.
Veröffentlicht: (2025)
von: Loukachevitch, Natalia, et al.
Veröffentlicht: (2025)
SAGE: Benchmarking and Improving Retrieval for Deep Research Agents
von: Hu, Tiansheng, et al.
Veröffentlicht: (2026)
von: Hu, Tiansheng, et al.
Veröffentlicht: (2026)
DAPR: A Benchmark on Document-Aware Passage Retrieval
von: Wang, Kexin, et al.
Veröffentlicht: (2023)
von: Wang, Kexin, et al.
Veröffentlicht: (2023)
BERGEN: A Benchmarking Library for Retrieval-Augmented Generation
von: Rau, David, et al.
Veröffentlicht: (2024)
von: Rau, David, et al.
Veröffentlicht: (2024)
Fine-Tuning Large Language Models and Evaluating Retrieval Methods for Improved Question Answering on Building Codes
von: Aqib, Mohammad, et al.
Veröffentlicht: (2025)
von: Aqib, Mohammad, et al.
Veröffentlicht: (2025)
When to Retrieve: Teaching LLMs to Utilize Information Retrieval Effectively
von: Labruna, Tiziano, et al.
Veröffentlicht: (2024)
von: Labruna, Tiziano, et al.
Veröffentlicht: (2024)
An Open-Source Web-Based Tool for Evaluating Open-Source Large Language Models Leveraging Information Retrieval from Custom Documents
von: I, Godfrey
Veröffentlicht: (2025)
von: I, Godfrey
Veröffentlicht: (2025)
Frustratingly Simple Retrieval Improves Challenging, Reasoning-Intensive Benchmarks
von: Lyu, Xinxi, et al.
Veröffentlicht: (2025)
von: Lyu, Xinxi, et al.
Veröffentlicht: (2025)
Distillation for Multilingual Information Retrieval
von: Yang, Eugene, et al.
Veröffentlicht: (2024)
von: Yang, Eugene, et al.
Veröffentlicht: (2024)
FollowTable: A Benchmark for Instruction-Following Table Retrieval
von: Jin, Rihui, et al.
Veröffentlicht: (2026)
von: Jin, Rihui, et al.
Veröffentlicht: (2026)
Hindi-BEIR : A Large Scale Retrieval Benchmark in Hindi
von: Acharya, Arkadeep, et al.
Veröffentlicht: (2024)
von: Acharya, Arkadeep, et al.
Veröffentlicht: (2024)
PJB: A Reasoning-Aware Benchmark for Person-Job Retrieval
von: Wang, Guangzhi, et al.
Veröffentlicht: (2026)
von: Wang, Guangzhi, et al.
Veröffentlicht: (2026)
CORAL: Benchmarking Multi-turn Conversational Retrieval-Augmentation Generation
von: Cheng, Yiruo, et al.
Veröffentlicht: (2024)
von: Cheng, Yiruo, et al.
Veröffentlicht: (2024)
Redefining Retrieval Evaluation in the Era of LLMs
von: Trappolini, Giovanni, et al.
Veröffentlicht: (2025)
von: Trappolini, Giovanni, et al.
Veröffentlicht: (2025)
Automated Evaluation of Retrieval-Augmented Language Models with Task-Specific Exam Generation
von: Guinet, Gauthier, et al.
Veröffentlicht: (2024)
von: Guinet, Gauthier, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Wikipedia-based Datasets in Russian Information Retrieval Benchmark RusBEIR
von: Kovalev, Grigory, et al.
Veröffentlicht: (2025) -
Exploring Prompt-Based Methods for Zero-Shot Hypernym Prediction with Large Language Models
von: Tikhomirov, Mikhail, et al.
Veröffentlicht: (2024) -
STELLA: Self-Reflective Terminology-Aware Framework for Building an Aerospace Information Retrieval Benchmark
von: Kim, Bongmin
Veröffentlicht: (2026) -
Iterative Layer-wise Distillation for Efficient Compression of Large Language Models
von: Kovalev, Grigory, et al.
Veröffentlicht: (2025) -
Benchmarking and Building Zero-Shot Hindi Retrieval Model with Hindi-BEIR and NLLB-E5
von: Acharya, Arkadeep, et al.
Veröffentlicht: (2024)