NCTB-QA: A Large-Scale Bangla Educational Question Answering Dataset and Benchmarking Performance
Fuente:
arXiv
Saved in:
| Main Authors: | Eyasir, Abrar, Ahmed, Tahsin, Ibrahim, Muhammad |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Adaptive Context Management for Intelligent Conversational Question Answering
by: Perera, Manoj Madushanka, et al.
Published: (2025)
by: Perera, Manoj Madushanka, et al.
Published: (2025)
Contextually Aware E-Commerce Product Question Answering using RAG
by: Tangarajan, Praveen, et al.
Published: (2025)
by: Tangarajan, Praveen, et al.
Published: (2025)
A Comparative Study of Language Models for Khmer Retrieval-Augmented Question Answering
by: Ros, Sereiwathna, et al.
Published: (2026)
by: Ros, Sereiwathna, et al.
Published: (2026)
ESGBench: A Benchmark for Explainable ESG Question Answering in Corporate Sustainability Reports
by: George, Sherine, et al.
Published: (2025)
by: George, Sherine, et al.
Published: (2025)
Local Hybrid Retrieval-Augmented Document QA
by: Astrino, Paolo
Published: (2025)
by: Astrino, Paolo
Published: (2025)
Agentic Retrieval-Augmented Generation for Financial Document Question Answering
by: Shu, Yang, et al.
Published: (2026)
by: Shu, Yang, et al.
Published: (2026)
PubMed Reasoner: Dynamic Reasoning-based Retrieval for Evidence-Grounded Biomedical Question Answering
by: Zhang, Yiqing, et al.
Published: (2026)
by: Zhang, Yiqing, et al.
Published: (2026)
MAB-DQA: Addressing Query Aspect Importance in Document Question Answering with Multi-Armed Bandits
by: Xiang, Yixin, et al.
Published: (2026)
by: Xiang, Yixin, et al.
Published: (2026)
FinBERT-QA: Financial Question Answering with pre-trained BERT Language Models
by: Yuan, Bithiah
Published: (2025)
by: Yuan, Bithiah
Published: (2025)
Benchmarking Google Embeddings 2 against Open-Source Models for Multilingual Dense Retrieval and RAG Systems
by: Cirillo, Stefano, et al.
Published: (2026)
by: Cirillo, Stefano, et al.
Published: (2026)
HOME-KGQA: A Benchmark Dataset for Multimodal Knowledge Graph Question Answering on Household Daily Activities
by: Egami, Shusaku, et al.
Published: (2026)
by: Egami, Shusaku, et al.
Published: (2026)
NameBERT: Scaling Name-Based Nationality Classification with LLM-Augmented Open Academic Data
by: Ming, Cong, et al.
Published: (2026)
by: Ming, Cong, et al.
Published: (2026)
Memory Architectures for Multi-Turn Text-to-SQL: A Benchmark and Empirical Study
by: Tummalapenta, Ravi Kumar, et al.
Published: (2026)
by: Tummalapenta, Ravi Kumar, et al.
Published: (2026)
BridgeRAG: Training-Free Bridge-Conditioned Retrieval for Multi-Hop Question Answering
by: Bacellar, Andre
Published: (2026)
by: Bacellar, Andre
Published: (2026)
The Reasoning Bottleneck in Graph-RAG: Structured Prompting and Context Compression for Multi-Hop QA
by: Zarrinkia, Yasaman, et al.
Published: (2026)
by: Zarrinkia, Yasaman, et al.
Published: (2026)
ChronoMedKG: A Temporally-Grounded Biomedical Knowledge Graph and Benchmark for Clinical Reasoning
by: Ahmed, Md Shamim, et al.
Published: (2026)
by: Ahmed, Md Shamim, et al.
Published: (2026)
MasterSet: A Large-Scale Benchmark for Must-Cite Citation Recommendation in the AI/ML Literature
by: Ratul, Md Toyaha Rahman, et al.
Published: (2026)
by: Ratul, Md Toyaha Rahman, et al.
Published: (2026)
Architecture Matters More Than Scale: A Comparative Study of Retrieval and Memory Augmentation for Financial QA Under SME Compute Constraints
by: Liu, Jianan, et al.
Published: (2026)
by: Liu, Jianan, et al.
Published: (2026)
Flippi: End To End GenAI Assistant for E-Commerce
by: Rajasekar, Anand A., et al.
Published: (2025)
by: Rajasekar, Anand A., et al.
Published: (2025)
NewsScope: Schema-Grounded Cross-Domain News Claim Extraction with Open Models
by: Pandya, Nidhi
Published: (2025)
by: Pandya, Nidhi
Published: (2025)
Evaluation of Chunking Strategies for Effective Text Embedding in Low-Resource Language on Agricultural Documents
by: Chhoun, Sovandara, et al.
Published: (2026)
by: Chhoun, Sovandara, et al.
Published: (2026)
Stage-Audit: Auditable Source-Frontier Discovery for Cross-Wiki Tables
by: Shen, Chen
Published: (2026)
by: Shen, Chen
Published: (2026)
FACTUM: Mechanistic Detection of Citation Hallucination in Long-Form RAG
by: Dassen, Maxime, et al.
Published: (2026)
by: Dassen, Maxime, et al.
Published: (2026)
Extending AI for Research to the Humanities: A Multi-Agent Framework for Evidence-Grounded Scholarship
by: Pan, Yating, et al.
Published: (2026)
by: Pan, Yating, et al.
Published: (2026)
Answer Generation for Questions With Multiple Information Sources in E-Commerce
by: Rajasekar, Anand A., et al.
Published: (2021)
by: Rajasekar, Anand A., et al.
Published: (2021)
Controllable Evidence Selection in Retrieval-Augmented Question Answering via Deterministic Utility Gating
by: Unda, Victor P.
Published: (2026)
by: Unda, Victor P.
Published: (2026)
Real-World En Call Center Transcripts Dataset with PII Redaction
by: Dao, Ha, et al.
Published: (2025)
by: Dao, Ha, et al.
Published: (2025)
Augmented Relevance Datasets with Fine-Tuned Small LLMs
by: Fitte-Rey, Quentin, et al.
Published: (2025)
by: Fitte-Rey, Quentin, et al.
Published: (2025)
ArcheType: A Novel Framework for Open-Source Column Type Annotation using Large Language Models
by: Feuer, Benjamin, et al.
Published: (2023)
by: Feuer, Benjamin, et al.
Published: (2023)
Token-Oriented Object Notation vs JSON: A Benchmark of Plain and Constrained Decoding Generation
by: Matveev, Ivan
Published: (2026)
by: Matveev, Ivan
Published: (2026)
A Question Answering Dataset for Temporal-Sensitive Retrieval-Augmented Generation
by: Chen, Ziyang, et al.
Published: (2025)
by: Chen, Ziyang, et al.
Published: (2025)
LLM-as-a-Judge: Rapid Evaluation of Legal Document Recommendation for Retrieval-Augmented Generation
by: Pradhan, Anu, et al.
Published: (2025)
by: Pradhan, Anu, et al.
Published: (2025)
Temporal Decay of Co-Citation Predictability: A 20-Year Statute Retrieval Benchmark from 396M Ukrainian Court Citations
by: Ovcharov, Volodymyr
Published: (2026)
by: Ovcharov, Volodymyr
Published: (2026)
Prompt Perturbation in Retrieval-Augmented Generation based Large Language Models
by: Hu, Zhibo, et al.
Published: (2024)
by: Hu, Zhibo, et al.
Published: (2024)
IndoBERT-Sentiment: Context-Conditioned Sentiment Classification for Indonesian Text
by: Saputra, Muhammad Apriandito Arya, et al.
Published: (2026)
by: Saputra, Muhammad Apriandito Arya, et al.
Published: (2026)
IndoBERT-Relevancy: A Context-Conditioned Relevancy Classifier for Indonesian Text
by: Saputra, Muhammad Apriandito Arya, et al.
Published: (2026)
by: Saputra, Muhammad Apriandito Arya, et al.
Published: (2026)
Mitigating Hallucinations in Large Language Models via Self-Refinement-Enhanced Knowledge Retrieval
by: Niu, Mengjia, et al.
Published: (2024)
by: Niu, Mengjia, et al.
Published: (2024)
EQUATOR: A Deterministic Framework for Evaluating LLM Reasoning with Open-Ended Questions. # v1.0.0-beta
by: Bernard, Raymond, et al.
Published: (2024)
by: Bernard, Raymond, et al.
Published: (2024)
GISTBench: Evaluating LLM User Understanding via Evidence-Based Interest Verification
by: Fostiropoulos, Iordanis, et al.
Published: (2026)
by: Fostiropoulos, Iordanis, et al.
Published: (2026)
AI-Friendly LaTeX: Using LaTeX Code as a Knowledge Source for Retrieval-Augmented Generation
by: Verhoeff, Tom
Published: (2026)
by: Verhoeff, Tom
Published: (2026)
Similar Items
-
Towards Adaptive Context Management for Intelligent Conversational Question Answering
by: Perera, Manoj Madushanka, et al.
Published: (2025) -
Contextually Aware E-Commerce Product Question Answering using RAG
by: Tangarajan, Praveen, et al.
Published: (2025) -
A Comparative Study of Language Models for Khmer Retrieval-Augmented Question Answering
by: Ros, Sereiwathna, et al.
Published: (2026) -
ESGBench: A Benchmark for Explainable ESG Question Answering in Corporate Sustainability Reports
by: George, Sherine, et al.
Published: (2025) -
Local Hybrid Retrieval-Augmented Document QA
by: Astrino, Paolo
Published: (2025)