IndicSQuAD: A Comprehensive Multilingual Question Answering Dataset for Indic Languages
Fuente:
arXiv
Saved in:
| Main Authors: | Endait, Sharvi, Ghatage, Ruturaj, Kulkarni, Aditya, Patil, Rajlaxmi, Joshi, Raviraj |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MahaSQuAD: Bridging Linguistic Divides in Marathi Question-Answering
by: Ghatage, Ruturaj, et al.
Published: (2024)
by: Ghatage, Ruturaj, et al.
Published: (2024)
Handling and extracting key entities from customer conversations using Speech recognition and Named Entity recognition
by: Endait, Sharvi, et al.
Published: (2022)
by: Endait, Sharvi, et al.
Published: (2022)
Automated Assessment of Multimodal Answer Sheets in the STEM domain
by: Patil, Rajlaxmi, et al.
Published: (2024)
by: Patil, Rajlaxmi, et al.
Published: (2024)
L3Cube-IndicQuest: A Benchmark Question Answering Dataset for Evaluating Knowledge of LLMs in Indic Context
by: Rohera, Pritika, et al.
Published: (2024)
by: Rohera, Pritika, et al.
Published: (2024)
L3Cube-IndicNews: News-based Short Text and Long Document Classification Datasets in Indic Languages
by: Mirashi, Aishwarya, et al.
Published: (2024)
by: Mirashi, Aishwarya, et al.
Published: (2024)
L3Cube-IndicHeadline-ID: A Dataset for Headline Identification and Semantic Evaluation in Low-Resource Indian Languages
by: Tanksale, Nishant, et al.
Published: (2025)
by: Tanksale, Nishant, et al.
Published: (2025)
Overview of the 2023 ICON Shared Task on Gendered Abuse Detection in Indic Languages
by: Vaidya, Aatman, et al.
Published: (2024)
by: Vaidya, Aatman, et al.
Published: (2024)
IndicSentEval: How Effectively do Multilingual Transformer Models encode Linguistic Properties for Indic Languages?
by: Aravapalli, Akhilesh, et al.
Published: (2024)
by: Aravapalli, Akhilesh, et al.
Published: (2024)
L3Cube-MahaSum: A Comprehensive Dataset and BART Models for Abstractive Text Summarization in Marathi
by: Deshmukh, Pranita, et al.
Published: (2024)
by: Deshmukh, Pranita, et al.
Published: (2024)
IndicMedDialog: A Parallel Multi-Turn Medical Dialogue Dataset for Accessible Healthcare in Indic Languages
by: Nigam, Shubham Kumar, et al.
Published: (2026)
by: Nigam, Shubham Kumar, et al.
Published: (2026)
Multilingual State Space Models for Structured Question Answering in Indic Languages
by: Vats, Arpita, et al.
Published: (2025)
by: Vats, Arpita, et al.
Published: (2025)
Parallel Corpora for Machine Translation in Low-resource Indic Languages: A Comprehensive Review
by: Raja, Rahul, et al.
Published: (2025)
by: Raja, Rahul, et al.
Published: (2025)
Table Question Answering for Low-resourced Indic Languages
by: Pal, Vaishali, et al.
Published: (2024)
by: Pal, Vaishali, et al.
Published: (2024)
Multi-Stage Training for Abusive Comment Detection in Indic Languages
by: Rastogi, Pranshu, et al.
Published: (2026)
by: Rastogi, Pranshu, et al.
Published: (2026)
Long-context Non-factoid Question Answering in Indic Languages
by: Mishra, Ritwik, et al.
Published: (2025)
by: Mishra, Ritwik, et al.
Published: (2025)
L3Cube-MahaEmotions: A Marathi Emotion Recognition Dataset with Synthetic Annotations using CoTR prompting and Large Language Models
by: Kowtal, Nidhi, et al.
Published: (2025)
by: Kowtal, Nidhi, et al.
Published: (2025)
IndicParam: Benchmark to evaluate LLMs on low-resource Indic Languages
by: Maheshwari, Ayush, et al.
Published: (2025)
by: Maheshwari, Ayush, et al.
Published: (2025)
HeySQuAD: A Spoken Question Answering Dataset
by: Wu, Yijing, et al.
Published: (2023)
by: Wu, Yijing, et al.
Published: (2023)
L3Cube-MahaSTS: A Marathi Sentence Similarity Dataset and Models
by: Mirashi, Aishwarya, et al.
Published: (2025)
by: Mirashi, Aishwarya, et al.
Published: (2025)
Hybrid-SQuAD: Hybrid Scholarly Question Answering Dataset
by: Taffa, Tilahun Abedissa, et al.
Published: (2024)
by: Taffa, Tilahun Abedissa, et al.
Published: (2024)
INDIC QA BENCHMARK: A Multilingual Benchmark to Evaluate Question Answering capability of LLMs for Indic Languages
by: Singh, Abhishek Kumar, et al.
Published: (2024)
by: Singh, Abhishek Kumar, et al.
Published: (2024)
FIND: Toward Multimodal Financial Reasoning and Question Answering for Indic Languages
by: Das, Sarmistha, et al.
Published: (2026)
by: Das, Sarmistha, et al.
Published: (2026)
IndicGenBench: A Multilingual Benchmark to Evaluate Generation Capabilities of LLMs on Indic Languages
by: Singh, Harman, et al.
Published: (2024)
by: Singh, Harman, et al.
Published: (2024)
L3Cube-MahaSocialNER: A Social Media based Marathi NER Dataset and BERT models
by: Chaudhari, Harsh, et al.
Published: (2023)
by: Chaudhari, Harsh, et al.
Published: (2023)
Better To Ask in English? Evaluating Factual Accuracy of Multilingual LLMs in English and Low-Resource Languages
by: Rohera, Pritika, et al.
Published: (2025)
by: Rohera, Pritika, et al.
Published: (2025)
SQuARE: Sequential Question Answering Reasoning Engine for Enhanced Chain-of-Thought in Large Language Models
by: Fleischer, Daniel, et al.
Published: (2025)
by: Fleischer, Daniel, et al.
Published: (2025)
Decoding the Diversity: A Review of the Indic AI Research Landscape
by: KJ, Sankalp, et al.
Published: (2024)
by: KJ, Sankalp, et al.
Published: (2024)
IndicVoices: Towards building an Inclusive Multilingual Speech Dataset for Indian Languages
by: Javed, Tahir, et al.
Published: (2024)
by: Javed, Tahir, et al.
Published: (2024)
Leveraging Parameter Efficient Training Methods for Low Resource Text Classification: A Case Study in Marathi
by: Deshmukh, Pranita, et al.
Published: (2024)
by: Deshmukh, Pranita, et al.
Published: (2024)
LittiChoQA: Literary Texts in Indic Languages Chosen for Question Answering
by: Khandelwal, Aarya, et al.
Published: (2026)
by: Khandelwal, Aarya, et al.
Published: (2026)
Curating Stopwords in Marathi: A TF-IDF Approach for Improved Text Analysis and Information Retrieval
by: Chavan, Rohan, et al.
Published: (2024)
by: Chavan, Rohan, et al.
Published: (2024)
Challenges in Adapting Multilingual LLMs to Low-Resource Languages using LoRA PEFT Tuning
by: Khade, Omkar, et al.
Published: (2024)
by: Khade, Omkar, et al.
Published: (2024)
ELR-1000: A Community-Generated Dataset for Endangered Indic Indigenous Languages
by: Joshi, Neha, et al.
Published: (2025)
by: Joshi, Neha, et al.
Published: (2025)
IndicVoices-R: Unlocking a Massive Multilingual Multi-speaker Speech Corpus for Scaling Indian TTS
by: Sankar, Ashwin, et al.
Published: (2024)
by: Sankar, Ashwin, et al.
Published: (2024)
On the Calibration of Multilingual Question Answering LLMs
by: Yang, Yahan, et al.
Published: (2023)
by: Yang, Yahan, et al.
Published: (2023)
Chain-of-Translation Prompting (CoTR): A Novel Prompting Technique for Low Resource Languages
by: Deshpande, Tejas, et al.
Published: (2024)
by: Deshpande, Tejas, et al.
Published: (2024)
AmaSQuAD: A Benchmark for Amharic Extractive Question Answering
by: Hailemariam, Nebiyou Daniel, et al.
Published: (2025)
by: Hailemariam, Nebiyou Daniel, et al.
Published: (2025)
Long Range Named Entity Recognition for Marathi Documents
by: Deshmukh, Pranita, et al.
Published: (2024)
by: Deshmukh, Pranita, et al.
Published: (2024)
Topic Modeling in Marathi
by: Shinde, Sanket, et al.
Published: (2025)
by: Shinde, Sanket, et al.
Published: (2025)
KenSwQuAD -- A Question Answering Dataset for Swahili Low Resource Language
by: Wanjawa, Barack W., et al.
Published: (2022)
by: Wanjawa, Barack W., et al.
Published: (2022)
Similar Items
-
MahaSQuAD: Bridging Linguistic Divides in Marathi Question-Answering
by: Ghatage, Ruturaj, et al.
Published: (2024) -
Handling and extracting key entities from customer conversations using Speech recognition and Named Entity recognition
by: Endait, Sharvi, et al.
Published: (2022) -
Automated Assessment of Multimodal Answer Sheets in the STEM domain
by: Patil, Rajlaxmi, et al.
Published: (2024) -
L3Cube-IndicQuest: A Benchmark Question Answering Dataset for Evaluating Knowledge of LLMs in Indic Context
by: Rohera, Pritika, et al.
Published: (2024) -
L3Cube-IndicNews: News-based Short Text and Long Document Classification Datasets in Indic Languages
by: Mirashi, Aishwarya, et al.
Published: (2024)