SciQAG: A Framework for Auto-Generated Science Question Answering Dataset with Fine-grained Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Wan, Yuwei, Liu, Yixuan, Ajith, Aswathy, Grazian, Clara, Hoex, Bram, Zhang, Wenjie, Kit, Chunyu, Xie, Tong, Foster, Ian |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ByteScience: Bridging Unstructured Scientific Literature and Structured Data with Auto Fine-tuned Large Language Model in Token Granularity
by: Xie, Tong, et al.
Published: (2024)
by: Xie, Tong, et al.
Published: (2024)
DARWIN 1.5: Large Language Models as Materials Science Adapted Learners
by: Xie, Tong, et al.
Published: (2024)
by: Xie, Tong, et al.
Published: (2024)
From Tokens to Materials: Leveraging Language Models for Scientific Discovery
by: Wan, Yuwei, et al.
Published: (2024)
by: Wan, Yuwei, et al.
Published: (2024)
Construction and Application of Materials Knowledge Graph in Multidisciplinary Materials Science via Large Language Model
by: Ye, Yanpeng, et al.
Published: (2024)
by: Ye, Yanpeng, et al.
Published: (2024)
VeriSciQA: An Auto-Verified Dataset for Scientific Visual Question Answering
by: Li, Yuyi, et al.
Published: (2025)
by: Li, Yuyi, et al.
Published: (2025)
DrugMCTS: a drug repurposing framework combining multi-agent, RAG and Monte Carlo Tree Search
by: Yang, Zerui, et al.
Published: (2025)
by: Yang, Zerui, et al.
Published: (2025)
SciEGQA: A Dataset for Scientific Evidence-Grounded Question Answering and Reasoning
by: Yu, Wenhan, et al.
Published: (2025)
by: Yu, Wenhan, et al.
Published: (2025)
Cinéaste: A Fine-grained Contextual Movie Question Answering Benchmark
by: Shah, Nisarg A., et al.
Published: (2025)
by: Shah, Nisarg A., et al.
Published: (2025)
AutoFish: Dataset and Benchmark for Fine-grained Analysis of Fish
by: Bengtson, Stefan Hein, et al.
Published: (2025)
by: Bengtson, Stefan Hein, et al.
Published: (2025)
TransLaw: A Large-Scale Dataset and Multi-Agent Benchmark Simulating Professional Translation of Hong Kong Case Law
by: Xuan, Xi, et al.
Published: (2025)
by: Xuan, Xi, et al.
Published: (2025)
Memory Injections: Correcting Multi-Hop Reasoning Failures during Inference in Transformer-Based Language Models
by: Sakarvadia, Mansi, et al.
Published: (2023)
by: Sakarvadia, Mansi, et al.
Published: (2023)
Memory-Centric Embodied Question Answering
by: Zhai, Mingliang, et al.
Published: (2025)
by: Zhai, Mingliang, et al.
Published: (2025)
ETVA: Evaluation of Text-to-Video Alignment via Fine-grained Question Generation and Answering
by: Guan, Kaisi, et al.
Published: (2025)
by: Guan, Kaisi, et al.
Published: (2025)
EVQAScore: A Fine-grained Metric for Video Question Answering Data Quality Evaluation
by: Liang, Hao, et al.
Published: (2024)
by: Liang, Hao, et al.
Published: (2024)
RAMoEA-QA: Hierarchical Specialization for Robust Respiratory Audio Question Answering
by: Bertolino, Gaia A., et al.
Published: (2026)
by: Bertolino, Gaia A., et al.
Published: (2026)
Synthetic Dataset Creation and Fine-Tuning of Transformer Models for Question Answering in Serbian
by: Cvetanović, Aleksa, et al.
Published: (2024)
by: Cvetanović, Aleksa, et al.
Published: (2024)
Fine-Tuning vs. RAG for Multi-Hop Question Answering with Novel Knowledge
by: Yang, Zhuoyi, et al.
Published: (2026)
by: Yang, Zhuoyi, et al.
Published: (2026)
Deep Model Merging: The Sister of Neural Network Interpretability -- A Survey
by: Khan, Arham, et al.
Published: (2024)
by: Khan, Arham, et al.
Published: (2024)
AutoViVQA: A Large-Scale Automatically Constructed Dataset for Vietnamese Visual Question Answering
by: Tuong, Nguyen Anh, et al.
Published: (2026)
by: Tuong, Nguyen Anh, et al.
Published: (2026)
Advances in Bayesian random partition models: A comprehensive review
by: Grazian, Clara
Published: (2023)
by: Grazian, Clara
Published: (2023)
Bayesian Consistency for Long Memory Processes: A Semiparametric Perspective
by: Grazian, Clara
Published: (2024)
by: Grazian, Clara
Published: (2024)
Approximate Bayesian Computation with Statistical Distances for Model Selection
by: Grazian, Clara
Published: (2024)
by: Grazian, Clara
Published: (2024)
RA-QA: A Benchmarking System for Respiratory Audio Question Answering Under Real-World Heterogeneity
by: Bertolino, Gaia A., et al.
Published: (2026)
by: Bertolino, Gaia A., et al.
Published: (2026)
Prompting Video-Language Foundation Models with Domain-specific Fine-grained Heuristics for Video Question Answering
by: Yu, Ting, et al.
Published: (2024)
by: Yu, Ting, et al.
Published: (2024)
Texts or Images? A Fine-grained Analysis on the Effectiveness of Input Representations and Models for Table Question Answering
by: Zhou, Wei, et al.
Published: (2025)
by: Zhou, Wei, et al.
Published: (2025)
Mitigating Memorization In Language Models
by: Sakarvadia, Mansi, et al.
Published: (2024)
by: Sakarvadia, Mansi, et al.
Published: (2024)
Multi-Sourced Compositional Generalization in Visual Question Answering
by: Li, Chuanhao, et al.
Published: (2025)
by: Li, Chuanhao, et al.
Published: (2025)
Towards Fine-Grained Video Question Answering
by: Dai, Wei, et al.
Published: (2025)
by: Dai, Wei, et al.
Published: (2025)
AnimeAdapter: Fine-grained and Consistent Zero-shot Anime Character Generation
by: Han, Yixuan
Published: (2026)
by: Han, Yixuan
Published: (2026)
LSHBloom: Memory-efficient, Extreme-scale Document Deduplication
by: Khan, Arham, et al.
Published: (2024)
by: Khan, Arham, et al.
Published: (2024)
PolQA: Polish Question Answering Dataset
by: Rybak, Piotr, et al.
Published: (2022)
by: Rybak, Piotr, et al.
Published: (2022)
A Collection of Question Answering Datasets for Norwegian
by: Mikhailov, Vladislav, et al.
Published: (2025)
by: Mikhailov, Vladislav, et al.
Published: (2025)
Improved carrier collection efficiency in CZTS solar cells by Li‐enhanced liquid‐phase‐assisted grain growth
by: Xiaojie Yuan, et al.
Published: (2024)
by: Xiaojie Yuan, et al.
Published: (2024)
LLaMa-SciQ: An Educational Chatbot for Answering Science MCQ
by: Allard, Marc-Antoine, et al.
Published: (2024)
by: Allard, Marc-Antoine, et al.
Published: (2024)
Knowledge-Aware Diverse Reranking for Cross-Source Question Answering
by: Zhou, Tong
Published: (2025)
by: Zhou, Tong
Published: (2025)
Instance-Level Trojan Attacks on Visual Question Answering via Adversarial Learning in Neuron Activation Space
by: Sun, Yuwei, et al.
Published: (2023)
by: Sun, Yuwei, et al.
Published: (2023)
DialogAgent: An Auto-engagement Agent for Code Question Answering Data Production
by: Liang, Xiaoyun, et al.
Published: (2024)
by: Liang, Xiaoyun, et al.
Published: (2024)
A Dataset for Spatiotemporal-Sensitive POI Question Answering
by: Han, Xiao, et al.
Published: (2025)
by: Han, Xiao, et al.
Published: (2025)
FoQA: A Faroese Question-Answering Dataset
by: Simonsen, Annika, et al.
Published: (2025)
by: Simonsen, Annika, et al.
Published: (2025)
TIGQA:An Expert Annotated Question Answering Dataset in Tigrinya
by: Teklehaymanot, Hailay, et al.
Published: (2024)
by: Teklehaymanot, Hailay, et al.
Published: (2024)
Similar Items
-
ByteScience: Bridging Unstructured Scientific Literature and Structured Data with Auto Fine-tuned Large Language Model in Token Granularity
by: Xie, Tong, et al.
Published: (2024) -
DARWIN 1.5: Large Language Models as Materials Science Adapted Learners
by: Xie, Tong, et al.
Published: (2024) -
From Tokens to Materials: Leveraging Language Models for Scientific Discovery
by: Wan, Yuwei, et al.
Published: (2024) -
Construction and Application of Materials Knowledge Graph in Multidisciplinary Materials Science via Large Language Model
by: Ye, Yanpeng, et al.
Published: (2024) -
VeriSciQA: An Auto-Verified Dataset for Scientific Visual Question Answering
by: Li, Yuyi, et al.
Published: (2025)