A Collection of Question Answering Datasets for Norwegian
Fuente:
arXiv
Saved in:
| Main Authors: | Mikhailov, Vladislav, Mæhlum, Petter, Langø, Victoria Ovedie Chruickshank, Velldal, Erik, Øvrelid, Lilja |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
NorEval: A Norwegian Language Understanding and Generation Evaluation Benchmark
by: Mikhailov, Vladislav, et al.
Published: (2025)
by: Mikhailov, Vladislav, et al.
Published: (2025)
Benchmarking Abstractive Summarisation: A Dataset of Human-authored Summaries of Norwegian News Articles
by: Touileb, Samia, et al.
Published: (2025)
by: Touileb, Samia, et al.
Published: (2025)
Mixed Feelings: Cross-Domain Sentiment Classification of Patient Feedback
by: Rønningstad, Egil, et al.
Published: (2025)
by: Rønningstad, Egil, et al.
Published: (2025)
Fluent Alignment with Disfluent Judges: Post-training for Lower-resource Languages
by: Samuel, David, et al.
Published: (2025)
by: Samuel, David, et al.
Published: (2025)
Multi-label Scandinavian Language Identification (SLIDE)
by: Fedorova, Mariia, et al.
Published: (2025)
by: Fedorova, Mariia, et al.
Published: (2025)
It's Difficult to be Neutral -- Human and LLM-based Sentiment Annotation of Patient Comments
by: Mæhlum, Petter, et al.
Published: (2024)
by: Mæhlum, Petter, et al.
Published: (2024)
Event-based evaluation of abstractive news summarization
by: You, Huiling, et al.
Published: (2025)
by: You, Huiling, et al.
Published: (2025)
Entity-Level Sentiment: More than the Sum of Its Parts
by: Rønningstad, Egil, et al.
Published: (2024)
by: Rønningstad, Egil, et al.
Published: (2024)
Small Languages, Big Models: A Study of Continual Training on Languages of Norway
by: Samuel, David, et al.
Published: (2024)
by: Samuel, David, et al.
Published: (2024)
Polish-ASTE: Aspect-Sentiment Triplet Extraction Datasets for Polish
by: Lango, Marta, et al.
Published: (2025)
by: Lango, Marta, et al.
Published: (2025)
The Impact of Copyrighted Material on Large Language Models: A Norwegian Perspective
by: de la Rosa, Javier, et al.
Published: (2024)
by: de la Rosa, Javier, et al.
Published: (2024)
Compositional Generalization with Grounded Language Models
by: Wold, Sondre, et al.
Published: (2024)
by: Wold, Sondre, et al.
Published: (2024)
HeySQuAD: A Spoken Question Answering Dataset
by: Wu, Yijing, et al.
Published: (2023)
by: Wu, Yijing, et al.
Published: (2023)
A Dataset of Open-Domain Question Answering with Multiple-Span Answers
by: Luo, Zhiyi, et al.
Published: (2024)
by: Luo, Zhiyi, et al.
Published: (2024)
Automatic Dataset Generation for Knowledge Intensive Question Answering Tasks
by: Yuen, Sizhe, et al.
Published: (2025)
by: Yuen, Sizhe, et al.
Published: (2025)
Hybrid-SQuAD: Hybrid Scholarly Question Answering Dataset
by: Taffa, Tilahun Abedissa, et al.
Published: (2024)
by: Taffa, Tilahun Abedissa, et al.
Published: (2024)
BanglaQuAD: A Bengali Open-domain Question Answering Dataset
by: Rony, Md Rashad Al Hasan, et al.
Published: (2024)
by: Rony, Md Rashad Al Hasan, et al.
Published: (2024)
FinTextQA: A Dataset for Long-form Financial Question Answering
by: Chen, Jian, et al.
Published: (2024)
by: Chen, Jian, et al.
Published: (2024)
Building a Rich Dataset to Empower the Persian Question Answering Systems
by: Yazdinejad, Mohsen, et al.
Published: (2024)
by: Yazdinejad, Mohsen, et al.
Published: (2024)
MediQAl: A French Medical Question Answering Dataset for Knowledge and Reasoning Evaluation
by: Bazoge, Adrien
Published: (2025)
by: Bazoge, Adrien
Published: (2025)
BoundingDocs: a Unified Dataset for Document Question Answering with Spatial Annotations
by: Giovannini, Simone, et al.
Published: (2025)
by: Giovannini, Simone, et al.
Published: (2025)
LLM Agents Implement an NLG System from Scratch: Building Interpretable Rule-Based RDF-to-Text Generators
by: Lango, Mateusz, et al.
Published: (2025)
by: Lango, Mateusz, et al.
Published: (2025)
From Chat Logs to Collective Insights: Aggregative Question Answering
by: Zhang, Wentao, et al.
Published: (2025)
by: Zhang, Wentao, et al.
Published: (2025)
RepLiQA: A Question-Answering Dataset for Benchmarking LLMs on Unseen Reference Content
by: Monteiro, Joao, et al.
Published: (2024)
by: Monteiro, Joao, et al.
Published: (2024)
CasiMedicos-Arg: A Medical Question Answering Dataset Annotated with Explanatory Argumentative Structures
by: Sviridova, Ekaterina, et al.
Published: (2024)
by: Sviridova, Ekaterina, et al.
Published: (2024)
VLQA: The First Comprehensive, Large, and High-Quality Vietnamese Dataset for Legal Question Answering
by: Nguyen, Tan-Minh, et al.
Published: (2025)
by: Nguyen, Tan-Minh, et al.
Published: (2025)
Declarative Knowledge Distillation from Large Language Models for Visual Question Answering Datasets
by: Eiter, Thomas, et al.
Published: (2024)
by: Eiter, Thomas, et al.
Published: (2024)
SPIQA: A Dataset for Multimodal Question Answering on Scientific Papers
by: Pramanick, Shraman, et al.
Published: (2024)
by: Pramanick, Shraman, et al.
Published: (2024)
SciQAG: A Framework for Auto-Generated Science Question Answering Dataset with Fine-grained Evaluation
by: Wan, Yuwei, et al.
Published: (2024)
by: Wan, Yuwei, et al.
Published: (2024)
RETQA: A Large-Scale Open-Domain Tabular Question Answering Dataset for Real Estate Sector
by: Wang, Zhensheng, et al.
Published: (2024)
by: Wang, Zhensheng, et al.
Published: (2024)
Faithful and Plausible Natural Language Explanations for Image Classification: A Pipeline Approach
by: Wojciechowski, Adam, et al.
Published: (2024)
by: Wojciechowski, Adam, et al.
Published: (2024)
Efficient Medical Question Answering with Knowledge-Augmented Question Generation
by: Khlaut, Julien, et al.
Published: (2024)
by: Khlaut, Julien, et al.
Published: (2024)
PeerQA: A Scientific Question Answering Dataset from Peer Reviews
by: Baumgärtner, Tim, et al.
Published: (2025)
by: Baumgärtner, Tim, et al.
Published: (2025)
Can Language Models Analyze Data? Evaluating Large Language Models for Question Answering over Datasets
by: Xenofontos, Andreas, et al.
Published: (2026)
by: Xenofontos, Andreas, et al.
Published: (2026)
From National Curricula to Cultural Awareness: Constructing Open-Ended Culture-Specific Question Answering Dataset
by: Yoo, Haneul, et al.
Published: (2026)
by: Yoo, Haneul, et al.
Published: (2026)
Knowledge Graph-Guided Multi-Agent Distillation for Reliable Industrial Question Answering with Datasets
by: Pan, Jiqun, et al.
Published: (2025)
by: Pan, Jiqun, et al.
Published: (2025)
Augmenting Question Answering with A Hybrid RAG Approach
by: Yang, Tianyi, et al.
Published: (2026)
by: Yang, Tianyi, et al.
Published: (2026)
A Benchmark for Long-Form Medical Question Answering
by: Hosseini, Pedram, et al.
Published: (2024)
by: Hosseini, Pedram, et al.
Published: (2024)
QPaug: Question and Passage Augmentation for Open-Domain Question Answering of LLMs
by: Kim, Minsang, et al.
Published: (2024)
by: Kim, Minsang, et al.
Published: (2024)
QACP: An Annotated Question Answering Dataset for Assisting Chinese Python Programming Learners
by: Xiao, Rui, et al.
Published: (2024)
by: Xiao, Rui, et al.
Published: (2024)
Similar Items
-
NorEval: A Norwegian Language Understanding and Generation Evaluation Benchmark
by: Mikhailov, Vladislav, et al.
Published: (2025) -
Benchmarking Abstractive Summarisation: A Dataset of Human-authored Summaries of Norwegian News Articles
by: Touileb, Samia, et al.
Published: (2025) -
Mixed Feelings: Cross-Domain Sentiment Classification of Patient Feedback
by: Rønningstad, Egil, et al.
Published: (2025) -
Fluent Alignment with Disfluent Judges: Post-training for Lower-resource Languages
by: Samuel, David, et al.
Published: (2025) -
Multi-label Scandinavian Language Identification (SLIDE)
by: Fedorova, Mariia, et al.
Published: (2025)