LittiChoQA: Literary Texts in Indic Languages Chosen for Question Answering

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Khandelwal, Aarya, Mishra, Ritwik, Shah, Rajiv Ratn
Natura: Preprint
Pubblicazione: 2026
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866915710942314496
author Khandelwal, Aarya
Mishra, Ritwik
Shah, Rajiv Ratn
author_facet Khandelwal, Aarya
Mishra, Ritwik
Shah, Rajiv Ratn
contents Long-context question answering (QA) over literary texts poses significant challenges for modern large language models, particularly in low-resource languages. We address the scarcity of long-context QA resources for Indic languages by introducing LittiChoQA, the largest literary QA dataset to date covering many languages spoken in the Gangetic plains of India. The dataset comprises over 270K automatically generated question-answer pairs with a balanced distribution of factoid and non-factoid questions, generated from naturally authored literary texts collected from the open web. We evaluate multiple multilingual LLMs on non-factoid, abstractive QA, under both full-context and context-shortened settings. Results demonstrate a clear trade-off between performance and efficiency: full-context fine-tuning yields the highest token-level and semantic-level scores, while context shortening substantially improves throughput. Among the evaluated models, Krutrim-2 achieves the strongest performance, obtaining a semantic score of 76.1 with full context. While, in shortened context settings it scores 74.9 with answer paragraph selection and 71.4 with vector-based retrieval. Qualitative evaluations further corroborate these findings.
format Preprint
id arxiv_https___arxiv_org_abs_2601_03025
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle LittiChoQA: Literary Texts in Indic Languages Chosen for Question Answering
Khandelwal, Aarya
Mishra, Ritwik
Shah, Rajiv Ratn
Computation and Language
Long-context question answering (QA) over literary texts poses significant challenges for modern large language models, particularly in low-resource languages. We address the scarcity of long-context QA resources for Indic languages by introducing LittiChoQA, the largest literary QA dataset to date covering many languages spoken in the Gangetic plains of India. The dataset comprises over 270K automatically generated question-answer pairs with a balanced distribution of factoid and non-factoid questions, generated from naturally authored literary texts collected from the open web. We evaluate multiple multilingual LLMs on non-factoid, abstractive QA, under both full-context and context-shortened settings. Results demonstrate a clear trade-off between performance and efficiency: full-context fine-tuning yields the highest token-level and semantic-level scores, while context shortening substantially improves throughput. Among the evaluated models, Krutrim-2 achieves the strongest performance, obtaining a semantic score of 76.1 with full context. While, in shortened context settings it scores 74.9 with answer paragraph selection and 71.4 with vector-based retrieval. Qualitative evaluations further corroborate these findings.
title LittiChoQA: Literary Texts in Indic Languages Chosen for Question Answering
topic Computation and Language
url https://arxiv.org/abs/2601.03025