Benchmarking and Building Zero-Shot Hindi Retrieval Model with Hindi-BEIR and NLLB-E5
Fuente:
arXiv
Saved in:
| Main Authors: | Acharya, Arkadeep, Murthy, Rudra, Kumar, Vishwajeet, Sen, Jaydeep |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hindi-BEIR : A Large Scale Retrieval Benchmark in Hindi
by: Acharya, Arkadeep, et al.
Published: (2024)
by: Acharya, Arkadeep, et al.
Published: (2024)
Mistral-SPLADE: LLMs for better Learned Sparse Retrieval
by: Doshi, Meet, et al.
Published: (2024)
by: Doshi, Meet, et al.
Published: (2024)
Influence Guided Sampling for Domain Adaptation of Text Retrievers
by: Doshi, Meet, et al.
Published: (2026)
by: Doshi, Meet, et al.
Published: (2026)
BEIR-PL: Zero Shot Information Retrieval Benchmark for the Polish Language
by: Wojtasik, Konrad, et al.
Published: (2023)
by: Wojtasik, Konrad, et al.
Published: (2023)
Wikipedia-based Datasets in Russian Information Retrieval Benchmark RusBEIR
by: Kovalev, Grigory, et al.
Published: (2025)
by: Kovalev, Grigory, et al.
Published: (2025)
On the effective transfer of knowledge from English to Hindi Wikipedia
by: Das, Paramita, et al.
Published: (2024)
by: Das, Paramita, et al.
Published: (2024)
LMK > CLS: Landmark Pooling for Dense Embeddings
by: Doshi, Meet, et al.
Published: (2026)
by: Doshi, Meet, et al.
Published: (2026)
Granite Embedding R2 Models
by: Awasthy, Parul, et al.
Published: (2025)
by: Awasthy, Parul, et al.
Published: (2025)
Granite Embedding Models
by: Awasthy, Parul, et al.
Published: (2025)
by: Awasthy, Parul, et al.
Published: (2025)
Building Russian Benchmark for Evaluation of Information Retrieval Models
by: Kovalev, Grigory, et al.
Published: (2025)
by: Kovalev, Grigory, et al.
Published: (2025)
A Representation Sharpening Framework for Zero Shot Dense Retrieval
by: Ashok, Dhananjay, et al.
Published: (2025)
by: Ashok, Dhananjay, et al.
Published: (2025)
ColBERT-XM: A Modular Multi-Vector Representation Model for Zero-Shot Multilingual Information Retrieval
by: Louis, Antoine, et al.
Published: (2024)
by: Louis, Antoine, et al.
Published: (2024)
GATHER: Convergence-Centric Hyper-Entity Retrieval for Zero-Shot Cell-Type Annotation
by: Zhang, Zhonghui, et al.
Published: (2026)
by: Zhang, Zhonghui, et al.
Published: (2026)
ZSE-Cap: A Zero-Shot Ensemble for Image Retrieval and Prompt-Guided Captioning
by: Dinh, Duc-Tai, et al.
Published: (2025)
by: Dinh, Duc-Tai, et al.
Published: (2025)
Investigating Task Arithmetic for Zero-Shot Information Retrieval
by: Braga, Marco, et al.
Published: (2025)
by: Braga, Marco, et al.
Published: (2025)
STELLA: Self-Reflective Terminology-Aware Framework for Building an Aerospace Information Retrieval Benchmark
by: Kim, Bongmin
Published: (2026)
by: Kim, Bongmin
Published: (2026)
NER Retriever: Zero-Shot Named Entity Retrieval with Type-Aware Embeddings
by: Shachar, Or, et al.
Published: (2025)
by: Shachar, Or, et al.
Published: (2025)
BERTopic for Topic Modeling of Hindi Short Texts: A Comparative Study
by: Mutsaddi, Atharva, et al.
Published: (2025)
by: Mutsaddi, Atharva, et al.
Published: (2025)
Large Language Models are Zero-Shot Rankers for Recommender Systems
by: Hou, Yupeng, et al.
Published: (2023)
by: Hou, Yupeng, et al.
Published: (2023)
MILU: A Multi-task Indic Language Understanding Benchmark
by: Verma, Sshubam, et al.
Published: (2024)
by: Verma, Sshubam, et al.
Published: (2024)
M3Retrieve: Benchmarking Multimodal Retrieval for Medicine
by: Acharya, Arkadeep, et al.
Published: (2025)
by: Acharya, Arkadeep, et al.
Published: (2025)
Taxonomy-Guided Zero-Shot Recommendations with LLMs
by: Liang, Yueqing, et al.
Published: (2024)
by: Liang, Yueqing, et al.
Published: (2024)
SLIMER-IT: Zero-Shot NER on Italian Language
by: Zamai, Andrew, et al.
Published: (2024)
by: Zamai, Andrew, et al.
Published: (2024)
Attention in Large Language Models Yields Efficient Zero-Shot Re-Rankers
by: Chen, Shijie, et al.
Published: (2024)
by: Chen, Shijie, et al.
Published: (2024)
Are LLMs Truly Multilingual? Exploring Zero-Shot Multilingual Capability of LLMs for Information Retrieval: An Italian Healthcare Use Case
by: Kembu, Vignesh Kumar, et al.
Published: (2025)
by: Kembu, Vignesh Kumar, et al.
Published: (2025)
Systematic Evaluation of Neural Retrieval Models on the Touché 2020 Argument Retrieval Subset of BEIR
by: Thakur, Nandan, et al.
Published: (2024)
by: Thakur, Nandan, et al.
Published: (2024)
FreshStack: Building Realistic Benchmarks for Evaluating Retrieval on Technical Documents
by: Thakur, Nandan, et al.
Published: (2025)
by: Thakur, Nandan, et al.
Published: (2025)
DREditor: An Time-efficient Approach for Building a Domain-specific Dense Retrieval Model
by: Huang, Chen, et al.
Published: (2024)
by: Huang, Chen, et al.
Published: (2024)
CoIR: A Comprehensive Benchmark for Code Information Retrieval Models
by: Li, Xiangyang, et al.
Published: (2024)
by: Li, Xiangyang, et al.
Published: (2024)
Zero-Shot Contextual Embeddings via Offline Synthetic Corpus Generation
by: Lippmann, Philip, et al.
Published: (2025)
by: Lippmann, Philip, et al.
Published: (2025)
RAR-b: Reasoning as Retrieval Benchmark
by: Xiao, Chenghao, et al.
Published: (2024)
by: Xiao, Chenghao, et al.
Published: (2024)
Adaptive Retrieval-Augmented Generation for Conversational Systems
by: Wang, Xi, et al.
Published: (2024)
by: Wang, Xi, et al.
Published: (2024)
A Cooperative Multi-Agent Framework for Zero-Shot Named Entity Recognition
by: Wang, Zihan, et al.
Published: (2025)
by: Wang, Zihan, et al.
Published: (2025)
Zero-Shot Dense Retrieval with Embeddings from Relevance Feedback
by: Jedidi, Nour, et al.
Published: (2024)
by: Jedidi, Nour, et al.
Published: (2024)
Benchmarking Information Retrieval Models on Complex Retrieval Tasks
by: Killingback, Julian, et al.
Published: (2025)
by: Killingback, Julian, et al.
Published: (2025)
XProvence: Zero-Cost Multilingual Context Pruning for Retrieval-Augmented Generation
by: Mohamed, Youssef, et al.
Published: (2026)
by: Mohamed, Youssef, et al.
Published: (2026)
AIR-Bench: Automated Heterogeneous Information Retrieval Benchmark
by: Chen, Jianlyu, et al.
Published: (2024)
by: Chen, Jianlyu, et al.
Published: (2024)
BERGEN: A Benchmarking Library for Retrieval-Augmented Generation
by: Rau, David, et al.
Published: (2024)
by: Rau, David, et al.
Published: (2024)
SAGE: Benchmarking and Improving Retrieval for Deep Research Agents
by: Hu, Tiansheng, et al.
Published: (2026)
by: Hu, Tiansheng, et al.
Published: (2026)
DAPR: A Benchmark on Document-Aware Passage Retrieval
by: Wang, Kexin, et al.
Published: (2023)
by: Wang, Kexin, et al.
Published: (2023)
Similar Items
-
Hindi-BEIR : A Large Scale Retrieval Benchmark in Hindi
by: Acharya, Arkadeep, et al.
Published: (2024) -
Mistral-SPLADE: LLMs for better Learned Sparse Retrieval
by: Doshi, Meet, et al.
Published: (2024) -
Influence Guided Sampling for Domain Adaptation of Text Retrievers
by: Doshi, Meet, et al.
Published: (2026) -
BEIR-PL: Zero Shot Information Retrieval Benchmark for the Polish Language
by: Wojtasik, Konrad, et al.
Published: (2023) -
Wikipedia-based Datasets in Russian Information Retrieval Benchmark RusBEIR
by: Kovalev, Grigory, et al.
Published: (2025)