SciQAG: A Framework for Auto-Generated Science Question Answering Dataset with Fine-grained Evaluation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wan, Yuwei, Liu, Yixuan, Ajith, Aswathy, Grazian, Clara, Hoex, Bram, Zhang, Wenjie, Kit, Chunyu, Xie, Tong, Foster, Ian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ByteScience: Bridging Unstructured Scientific Literature and Structured Data with Auto Fine-tuned Large Language Model in Token Granularity
von: Xie, Tong, et al.
Veröffentlicht: (2024)
von: Xie, Tong, et al.
Veröffentlicht: (2024)
DARWIN 1.5: Large Language Models as Materials Science Adapted Learners
von: Xie, Tong, et al.
Veröffentlicht: (2024)
von: Xie, Tong, et al.
Veröffentlicht: (2024)
From Tokens to Materials: Leveraging Language Models for Scientific Discovery
von: Wan, Yuwei, et al.
Veröffentlicht: (2024)
von: Wan, Yuwei, et al.
Veröffentlicht: (2024)
Construction and Application of Materials Knowledge Graph in Multidisciplinary Materials Science via Large Language Model
von: Ye, Yanpeng, et al.
Veröffentlicht: (2024)
von: Ye, Yanpeng, et al.
Veröffentlicht: (2024)
VeriSciQA: An Auto-Verified Dataset for Scientific Visual Question Answering
von: Li, Yuyi, et al.
Veröffentlicht: (2025)
von: Li, Yuyi, et al.
Veröffentlicht: (2025)
DrugMCTS: a drug repurposing framework combining multi-agent, RAG and Monte Carlo Tree Search
von: Yang, Zerui, et al.
Veröffentlicht: (2025)
von: Yang, Zerui, et al.
Veröffentlicht: (2025)
SciEGQA: A Dataset for Scientific Evidence-Grounded Question Answering and Reasoning
von: Yu, Wenhan, et al.
Veröffentlicht: (2025)
von: Yu, Wenhan, et al.
Veröffentlicht: (2025)
Cinéaste: A Fine-grained Contextual Movie Question Answering Benchmark
von: Shah, Nisarg A., et al.
Veröffentlicht: (2025)
von: Shah, Nisarg A., et al.
Veröffentlicht: (2025)
AutoFish: Dataset and Benchmark for Fine-grained Analysis of Fish
von: Bengtson, Stefan Hein, et al.
Veröffentlicht: (2025)
von: Bengtson, Stefan Hein, et al.
Veröffentlicht: (2025)
TransLaw: A Large-Scale Dataset and Multi-Agent Benchmark Simulating Professional Translation of Hong Kong Case Law
von: Xuan, Xi, et al.
Veröffentlicht: (2025)
von: Xuan, Xi, et al.
Veröffentlicht: (2025)
Memory Injections: Correcting Multi-Hop Reasoning Failures during Inference in Transformer-Based Language Models
von: Sakarvadia, Mansi, et al.
Veröffentlicht: (2023)
von: Sakarvadia, Mansi, et al.
Veröffentlicht: (2023)
Memory-Centric Embodied Question Answering
von: Zhai, Mingliang, et al.
Veröffentlicht: (2025)
von: Zhai, Mingliang, et al.
Veröffentlicht: (2025)
ETVA: Evaluation of Text-to-Video Alignment via Fine-grained Question Generation and Answering
von: Guan, Kaisi, et al.
Veröffentlicht: (2025)
von: Guan, Kaisi, et al.
Veröffentlicht: (2025)
EVQAScore: A Fine-grained Metric for Video Question Answering Data Quality Evaluation
von: Liang, Hao, et al.
Veröffentlicht: (2024)
von: Liang, Hao, et al.
Veröffentlicht: (2024)
RAMoEA-QA: Hierarchical Specialization for Robust Respiratory Audio Question Answering
von: Bertolino, Gaia A., et al.
Veröffentlicht: (2026)
von: Bertolino, Gaia A., et al.
Veröffentlicht: (2026)
Synthetic Dataset Creation and Fine-Tuning of Transformer Models for Question Answering in Serbian
von: Cvetanović, Aleksa, et al.
Veröffentlicht: (2024)
von: Cvetanović, Aleksa, et al.
Veröffentlicht: (2024)
Fine-Tuning vs. RAG for Multi-Hop Question Answering with Novel Knowledge
von: Yang, Zhuoyi, et al.
Veröffentlicht: (2026)
von: Yang, Zhuoyi, et al.
Veröffentlicht: (2026)
Deep Model Merging: The Sister of Neural Network Interpretability -- A Survey
von: Khan, Arham, et al.
Veröffentlicht: (2024)
von: Khan, Arham, et al.
Veröffentlicht: (2024)
AutoViVQA: A Large-Scale Automatically Constructed Dataset for Vietnamese Visual Question Answering
von: Tuong, Nguyen Anh, et al.
Veröffentlicht: (2026)
von: Tuong, Nguyen Anh, et al.
Veröffentlicht: (2026)
Advances in Bayesian random partition models: A comprehensive review
von: Grazian, Clara
Veröffentlicht: (2023)
von: Grazian, Clara
Veröffentlicht: (2023)
Bayesian Consistency for Long Memory Processes: A Semiparametric Perspective
von: Grazian, Clara
Veröffentlicht: (2024)
von: Grazian, Clara
Veröffentlicht: (2024)
Approximate Bayesian Computation with Statistical Distances for Model Selection
von: Grazian, Clara
Veröffentlicht: (2024)
von: Grazian, Clara
Veröffentlicht: (2024)
RA-QA: A Benchmarking System for Respiratory Audio Question Answering Under Real-World Heterogeneity
von: Bertolino, Gaia A., et al.
Veröffentlicht: (2026)
von: Bertolino, Gaia A., et al.
Veröffentlicht: (2026)
Prompting Video-Language Foundation Models with Domain-specific Fine-grained Heuristics for Video Question Answering
von: Yu, Ting, et al.
Veröffentlicht: (2024)
von: Yu, Ting, et al.
Veröffentlicht: (2024)
Texts or Images? A Fine-grained Analysis on the Effectiveness of Input Representations and Models for Table Question Answering
von: Zhou, Wei, et al.
Veröffentlicht: (2025)
von: Zhou, Wei, et al.
Veröffentlicht: (2025)
Mitigating Memorization In Language Models
von: Sakarvadia, Mansi, et al.
Veröffentlicht: (2024)
von: Sakarvadia, Mansi, et al.
Veröffentlicht: (2024)
Multi-Sourced Compositional Generalization in Visual Question Answering
von: Li, Chuanhao, et al.
Veröffentlicht: (2025)
von: Li, Chuanhao, et al.
Veröffentlicht: (2025)
Towards Fine-Grained Video Question Answering
von: Dai, Wei, et al.
Veröffentlicht: (2025)
von: Dai, Wei, et al.
Veröffentlicht: (2025)
AnimeAdapter: Fine-grained and Consistent Zero-shot Anime Character Generation
von: Han, Yixuan
Veröffentlicht: (2026)
von: Han, Yixuan
Veröffentlicht: (2026)
LSHBloom: Memory-efficient, Extreme-scale Document Deduplication
von: Khan, Arham, et al.
Veröffentlicht: (2024)
von: Khan, Arham, et al.
Veröffentlicht: (2024)
PolQA: Polish Question Answering Dataset
von: Rybak, Piotr, et al.
Veröffentlicht: (2022)
von: Rybak, Piotr, et al.
Veröffentlicht: (2022)
A Collection of Question Answering Datasets for Norwegian
von: Mikhailov, Vladislav, et al.
Veröffentlicht: (2025)
von: Mikhailov, Vladislav, et al.
Veröffentlicht: (2025)
Improved carrier collection efficiency in CZTS solar cells by Li‐enhanced liquid‐phase‐assisted grain growth
von: Xiaojie Yuan, et al.
Veröffentlicht: (2024)
von: Xiaojie Yuan, et al.
Veröffentlicht: (2024)
LLaMa-SciQ: An Educational Chatbot for Answering Science MCQ
von: Allard, Marc-Antoine, et al.
Veröffentlicht: (2024)
von: Allard, Marc-Antoine, et al.
Veröffentlicht: (2024)
Knowledge-Aware Diverse Reranking for Cross-Source Question Answering
von: Zhou, Tong
Veröffentlicht: (2025)
von: Zhou, Tong
Veröffentlicht: (2025)
Instance-Level Trojan Attacks on Visual Question Answering via Adversarial Learning in Neuron Activation Space
von: Sun, Yuwei, et al.
Veröffentlicht: (2023)
von: Sun, Yuwei, et al.
Veröffentlicht: (2023)
DialogAgent: An Auto-engagement Agent for Code Question Answering Data Production
von: Liang, Xiaoyun, et al.
Veröffentlicht: (2024)
von: Liang, Xiaoyun, et al.
Veröffentlicht: (2024)
A Dataset for Spatiotemporal-Sensitive POI Question Answering
von: Han, Xiao, et al.
Veröffentlicht: (2025)
von: Han, Xiao, et al.
Veröffentlicht: (2025)
FoQA: A Faroese Question-Answering Dataset
von: Simonsen, Annika, et al.
Veröffentlicht: (2025)
von: Simonsen, Annika, et al.
Veröffentlicht: (2025)
TIGQA:An Expert Annotated Question Answering Dataset in Tigrinya
von: Teklehaymanot, Hailay, et al.
Veröffentlicht: (2024)
von: Teklehaymanot, Hailay, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
ByteScience: Bridging Unstructured Scientific Literature and Structured Data with Auto Fine-tuned Large Language Model in Token Granularity
von: Xie, Tong, et al.
Veröffentlicht: (2024) -
DARWIN 1.5: Large Language Models as Materials Science Adapted Learners
von: Xie, Tong, et al.
Veröffentlicht: (2024) -
From Tokens to Materials: Leveraging Language Models for Scientific Discovery
von: Wan, Yuwei, et al.
Veröffentlicht: (2024) -
Construction and Application of Materials Knowledge Graph in Multidisciplinary Materials Science via Large Language Model
von: Ye, Yanpeng, et al.
Veröffentlicht: (2024) -
VeriSciQA: An Auto-Verified Dataset for Scientific Visual Question Answering
von: Li, Yuyi, et al.
Veröffentlicht: (2025)