INDIC QA BENCHMARK: A Multilingual Benchmark to Evaluate Question Answering capability of LLMs for Indic Languages
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Singh, Abhishek Kumar, kumar, Vishwajeet, Murthy, Rudra, Sen, Jaydeep, Mittal, Ashish, Ramakrishnan, Ganesh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MILU: A Multi-task Indic Language Understanding Benchmark
von: Verma, Sshubam, et al.
Veröffentlicht: (2024)
von: Verma, Sshubam, et al.
Veröffentlicht: (2024)
Hindi-BEIR : A Large Scale Retrieval Benchmark in Hindi
von: Acharya, Arkadeep, et al.
Veröffentlicht: (2024)
von: Acharya, Arkadeep, et al.
Veröffentlicht: (2024)
Benchmarking and Building Zero-Shot Hindi Retrieval Model with Hindi-BEIR and NLLB-E5
von: Acharya, Arkadeep, et al.
Veröffentlicht: (2024)
von: Acharya, Arkadeep, et al.
Veröffentlicht: (2024)
Mistral-SPLADE: LLMs for better Learned Sparse Retrieval
von: Doshi, Meet, et al.
Veröffentlicht: (2024)
von: Doshi, Meet, et al.
Veröffentlicht: (2024)
IndicGenBench: A Multilingual Benchmark to Evaluate Generation Capabilities of LLMs on Indic Languages
von: Singh, Harman, et al.
Veröffentlicht: (2024)
von: Singh, Harman, et al.
Veröffentlicht: (2024)
IndicSQuAD: A Comprehensive Multilingual Question Answering Dataset for Indic Languages
von: Endait, Sharvi, et al.
Veröffentlicht: (2025)
von: Endait, Sharvi, et al.
Veröffentlicht: (2025)
Dynamic Multi-Expert Projectors with Stabilized Routing for Multilingual Speech Recognition
von: Pandey, Isha, et al.
Veröffentlicht: (2026)
von: Pandey, Isha, et al.
Veröffentlicht: (2026)
LittiChoQA: Literary Texts in Indic Languages Chosen for Question Answering
von: Khandelwal, Aarya, et al.
Veröffentlicht: (2026)
von: Khandelwal, Aarya, et al.
Veröffentlicht: (2026)
Multilingual State Space Models for Structured Question Answering in Indic Languages
von: Vats, Arpita, et al.
Veröffentlicht: (2025)
von: Vats, Arpita, et al.
Veröffentlicht: (2025)
L3Cube-IndicQuest: A Benchmark Question Answering Dataset for Evaluating Knowledge of LLMs in Indic Context
von: Rohera, Pritika, et al.
Veröffentlicht: (2024)
von: Rohera, Pritika, et al.
Veröffentlicht: (2024)
Evaluating Monolingual and Multilingual Large Language Models for Greek Question Answering: The DemosQA Benchmark
von: Mastrokostas, Charalampos, et al.
Veröffentlicht: (2026)
von: Mastrokostas, Charalampos, et al.
Veröffentlicht: (2026)
MedExpQA: Multilingual Benchmarking of Large Language Models for Medical Question Answering
von: Alonso, Iñigo, et al.
Veröffentlicht: (2024)
von: Alonso, Iñigo, et al.
Veröffentlicht: (2024)
Influence Guided Sampling for Domain Adaptation of Text Retrievers
von: Doshi, Meet, et al.
Veröffentlicht: (2026)
von: Doshi, Meet, et al.
Veröffentlicht: (2026)
INDIC DIALECT: A Multi Task Benchmark to Evaluate and Translate in Indian Language Dialects
von: Sharma, Tarun, et al.
Veröffentlicht: (2026)
von: Sharma, Tarun, et al.
Veröffentlicht: (2026)
Table Question Answering for Low-resourced Indic Languages
von: Pal, Vaishali, et al.
Veröffentlicht: (2024)
von: Pal, Vaishali, et al.
Veröffentlicht: (2024)
A Three-Pronged Approach to Cross-Lingual Adaptation with Multilingual LLMs
von: Singh, Vaibhav, et al.
Veröffentlicht: (2024)
von: Singh, Vaibhav, et al.
Veröffentlicht: (2024)
MedEthicsQA: A Comprehensive Question Answering Benchmark for Medical Ethics Evaluation of LLMs
von: Wei, Jianhui, et al.
Veröffentlicht: (2025)
von: Wei, Jianhui, et al.
Veröffentlicht: (2025)
IndicDB -- Benchmarking Multilingual Text-to-SQL Capabilities in Indian Languages
von: Dawar, Aviral, et al.
Veröffentlicht: (2026)
von: Dawar, Aviral, et al.
Veröffentlicht: (2026)
Long-context Non-factoid Question Answering in Indic Languages
von: Mishra, Ritwik, et al.
Veröffentlicht: (2025)
von: Mishra, Ritwik, et al.
Veröffentlicht: (2025)
Lost in Transcription: How Speech-to-Text Errors Derail Code Understanding
von: Havare, Jayant, et al.
Veröffentlicht: (2026)
von: Havare, Jayant, et al.
Veröffentlicht: (2026)
M2QA: Multi-domain Multilingual Question Answering
von: Engländer, Leon, et al.
Veröffentlicht: (2024)
von: Engländer, Leon, et al.
Veröffentlicht: (2024)
Leveraging Synthetic Data for Question Answering with Multilingual LLMs in the Agricultural Domain
von: Kaur, Rishemjit, et al.
Veröffentlicht: (2025)
von: Kaur, Rishemjit, et al.
Veröffentlicht: (2025)
CodeVaani: A Multilingual, Voice-Based Code Learning Assistant
von: Havare, Jayant, et al.
Veröffentlicht: (2025)
von: Havare, Jayant, et al.
Veröffentlicht: (2025)
IndicParam: Benchmark to evaluate LLMs on low-resource Indic Languages
von: Maheshwari, Ayush, et al.
Veröffentlicht: (2025)
von: Maheshwari, Ayush, et al.
Veröffentlicht: (2025)
FIND: Toward Multimodal Financial Reasoning and Question Answering for Indic Languages
von: Das, Sarmistha, et al.
Veröffentlicht: (2026)
von: Das, Sarmistha, et al.
Veröffentlicht: (2026)
On the Calibration of Multilingual Question Answering LLMs
von: Yang, Yahan, et al.
Veröffentlicht: (2023)
von: Yang, Yahan, et al.
Veröffentlicht: (2023)
Mathematics Isn't Culture-Free: Probing Cultural Gaps via Entity and Scenario Perturbations
von: Tomar, Aditya, et al.
Veröffentlicht: (2025)
von: Tomar, Aditya, et al.
Veröffentlicht: (2025)
HistoryBankQA: Multilingual Temporal Question Answering on Historical Events
von: Mandal, Biswadip, et al.
Veröffentlicht: (2025)
von: Mandal, Biswadip, et al.
Veröffentlicht: (2025)
DisastQA: A Comprehensive Benchmark for Evaluating Question Answering in Disaster Management
von: Chen, Zhitong, et al.
Veröffentlicht: (2026)
von: Chen, Zhitong, et al.
Veröffentlicht: (2026)
Compound-QA: A Benchmark for Evaluating LLMs on Compound Questions
von: Hou, Yutao, et al.
Veröffentlicht: (2024)
von: Hou, Yutao, et al.
Veröffentlicht: (2024)
MizanQA: Benchmarking Large Language Models on Moroccan Legal Question Answering
von: Bahaj, Adil, et al.
Veröffentlicht: (2025)
von: Bahaj, Adil, et al.
Veröffentlicht: (2025)
Improving Multilingual Neural Machine Translation System for Indic Languages
von: Das, Sudhansu Bala, et al.
Veröffentlicht: (2022)
von: Das, Sudhansu Bala, et al.
Veröffentlicht: (2022)
Evaluating Robustness of LLMs in Question Answering on Multilingual Noisy OCR Data
von: Piryani, Bhawna, et al.
Veröffentlicht: (2025)
von: Piryani, Bhawna, et al.
Veröffentlicht: (2025)
RepLiQA: A Question-Answering Dataset for Benchmarking LLMs on Unseen Reference Content
von: Monteiro, Joao, et al.
Veröffentlicht: (2024)
von: Monteiro, Joao, et al.
Veröffentlicht: (2024)
ClimaQA_SLO - Slovenian Climate Question-Answering Benchmark
von: Ferk Ovčjak, Monika, et al.
Veröffentlicht: (2025)
von: Ferk Ovčjak, Monika, et al.
Veröffentlicht: (2025)
DebateQA: Evaluating Question Answering on Debatable Knowledge
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
CyberMaskQA: A Privacy-Aware Benchmark for Evaluating Large Language Models in Cybersecurity Question Answering
von: Gaddi, Matilda, et al.
Veröffentlicht: (2026)
von: Gaddi, Matilda, et al.
Veröffentlicht: (2026)
EconLogicQA: A Question-Answering Benchmark for Evaluating Large Language Models in Economic Sequential Reasoning
von: Quan, Yinzhu, et al.
Veröffentlicht: (2024)
von: Quan, Yinzhu, et al.
Veröffentlicht: (2024)
FairMedQA: Benchmarking Bias in Large Language Models for Medical Question Answering
von: Xiao, Ying, et al.
Veröffentlicht: (2025)
von: Xiao, Ying, et al.
Veröffentlicht: (2025)
IndicSafe: A Benchmark for Evaluating Multilingual LLM Safety in South Asia
von: Pattnayak, Priyaranjan, et al.
Veröffentlicht: (2026)
von: Pattnayak, Priyaranjan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
MILU: A Multi-task Indic Language Understanding Benchmark
von: Verma, Sshubam, et al.
Veröffentlicht: (2024) -
Hindi-BEIR : A Large Scale Retrieval Benchmark in Hindi
von: Acharya, Arkadeep, et al.
Veröffentlicht: (2024) -
Benchmarking and Building Zero-Shot Hindi Retrieval Model with Hindi-BEIR and NLLB-E5
von: Acharya, Arkadeep, et al.
Veröffentlicht: (2024) -
Mistral-SPLADE: LLMs for better Learned Sparse Retrieval
von: Doshi, Meet, et al.
Veröffentlicht: (2024) -
IndicGenBench: A Multilingual Benchmark to Evaluate Generation Capabilities of LLMs on Indic Languages
von: Singh, Harman, et al.
Veröffentlicht: (2024)