PassiveQA: A Three-Action Framework for Epistemically Calibrated Question Answering via Supervised Finetuning
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Baidya, Madhav S |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Detecting the Machine: A Comprehensive Benchmark of AI-Generated Text Detectors Across Architectures, Domains, and Adversarial Conditions
von: Baidya, Madhav S., et al.
Veröffentlicht: (2026)
von: Baidya, Madhav S., et al.
Veröffentlicht: (2026)
SensorQA: A Question Answering Benchmark for Daily-Life Monitoring
von: Reichman, Benjamin, et al.
Veröffentlicht: (2025)
von: Reichman, Benjamin, et al.
Veröffentlicht: (2025)
RAG-BioQA: A Retrieval-Augmented Generation Framework for Long-Form Biomedical Question Answering
von: Panchumarthi, Lovely Yeswanth, et al.
Veröffentlicht: (2025)
von: Panchumarthi, Lovely Yeswanth, et al.
Veröffentlicht: (2025)
Long-Span Question-Answering: Automatic Question Generation and QA-System Ranking via Side-by-Side Evaluation
von: Bohnet, Bernd, et al.
Veröffentlicht: (2024)
von: Bohnet, Bernd, et al.
Veröffentlicht: (2024)
pdfQA: Diverse, Challenging, and Realistic Question Answering over PDFs
von: Schimanski, Tobias, et al.
Veröffentlicht: (2026)
von: Schimanski, Tobias, et al.
Veröffentlicht: (2026)
QA-TOOLBOX: Conversational Question-Answering for process task guidance in manufacturing
von: Manuvinakurike, Ramesh, et al.
Veröffentlicht: (2024)
von: Manuvinakurike, Ramesh, et al.
Veröffentlicht: (2024)
MedExQA: Medical Question Answering Benchmark with Multiple Explanations
von: Kim, Yunsoo, et al.
Veröffentlicht: (2024)
von: Kim, Yunsoo, et al.
Veröffentlicht: (2024)
FinTextQA: A Dataset for Long-form Financial Question Answering
von: Chen, Jian, et al.
Veröffentlicht: (2024)
von: Chen, Jian, et al.
Veröffentlicht: (2024)
DataFrame QA: A Universal LLM Framework on DataFrame Question Answering Without Data Exposure
von: Ye, Junyi, et al.
Veröffentlicht: (2024)
von: Ye, Junyi, et al.
Veröffentlicht: (2024)
Memory-QA: Answering Recall Questions Based on Multimodal Memories
von: Jiang, Hongda, et al.
Veröffentlicht: (2025)
von: Jiang, Hongda, et al.
Veröffentlicht: (2025)
ExpliCIT-QA: Explainable Code-Based Image Table Question Answering
von: Lagos, Maximiliano Hormazábal, et al.
Veröffentlicht: (2025)
von: Lagos, Maximiliano Hormazábal, et al.
Veröffentlicht: (2025)
ResearchQA: Evaluating Scholarly Question Answering at Scale Across 75 Fields with Survey-Mined Questions and Rubrics
von: Yifei, Li S., et al.
Veröffentlicht: (2025)
von: Yifei, Li S., et al.
Veröffentlicht: (2025)
RepLiQA: A Question-Answering Dataset for Benchmarking LLMs on Unseen Reference Content
von: Monteiro, Joao, et al.
Veröffentlicht: (2024)
von: Monteiro, Joao, et al.
Veröffentlicht: (2024)
MedEthicsQA: A Comprehensive Question Answering Benchmark for Medical Ethics Evaluation of LLMs
von: Wei, Jianhui, et al.
Veröffentlicht: (2025)
von: Wei, Jianhui, et al.
Veröffentlicht: (2025)
SemViQA: A Semantic Question Answering System for Vietnamese Information Fact-Checking
von: Tran, Dien X., et al.
Veröffentlicht: (2025)
von: Tran, Dien X., et al.
Veröffentlicht: (2025)
MM-PhyQA: Multimodal Physics Question-Answering With Multi-Image CoT Prompting
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
MFORT-QA: Multi-hop Few-shot Open Rich Table Question Answering
von: Guan, Che, et al.
Veröffentlicht: (2024)
von: Guan, Che, et al.
Veröffentlicht: (2024)
A Semantic-Sampling Framework for Evaluating Calibration in Open-Ended Question Answering
von: Wang, Zhanliang, et al.
Veröffentlicht: (2026)
von: Wang, Zhanliang, et al.
Veröffentlicht: (2026)
Evaluating Monolingual and Multilingual Large Language Models for Greek Question Answering: The DemosQA Benchmark
von: Mastrokostas, Charalampos, et al.
Veröffentlicht: (2026)
von: Mastrokostas, Charalampos, et al.
Veröffentlicht: (2026)
Beyond Factual QA: Mentorship-Oriented Question Answering over Long-Form Multilingual Content
von: Bhalerao, Parth, et al.
Veröffentlicht: (2026)
von: Bhalerao, Parth, et al.
Veröffentlicht: (2026)
CT2C-QA: Multimodal Question Answering over Chinese Text, Table and Chart
von: Zhao, Bowen, et al.
Veröffentlicht: (2024)
von: Zhao, Bowen, et al.
Veröffentlicht: (2024)
RAG-QA Arena: Evaluating Domain Robustness for Long-form Retrieval Augmented Question Answering
von: Han, Rujun, et al.
Veröffentlicht: (2024)
von: Han, Rujun, et al.
Veröffentlicht: (2024)
MapQA: Open-domain Geospatial Question Answering on Map Data
von: Li, Zekun, et al.
Veröffentlicht: (2025)
von: Li, Zekun, et al.
Veröffentlicht: (2025)
PeerQA: A Scientific Question Answering Dataset from Peer Reviews
von: Baumgärtner, Tim, et al.
Veröffentlicht: (2025)
von: Baumgärtner, Tim, et al.
Veröffentlicht: (2025)
FanOutQA: A Multi-Hop, Multi-Document Question Answering Benchmark for Large Language Models
von: Zhu, Andrew, et al.
Veröffentlicht: (2024)
von: Zhu, Andrew, et al.
Veröffentlicht: (2024)
PASemiQA: Plan-Assisted Agent for Question Answering on Semi-Structured Data with Text and Relational Information
von: Yang, Hansi, et al.
Veröffentlicht: (2025)
von: Yang, Hansi, et al.
Veröffentlicht: (2025)
LiTransProQA: an LLM-based Literary Translation evaluation metric with Professional Question Answering
von: Zhang, Ran, et al.
Veröffentlicht: (2025)
von: Zhang, Ran, et al.
Veröffentlicht: (2025)
WikiMixQA: A Multimodal Benchmark for Question Answering over Tables and Charts
von: Foroutan, Negar, et al.
Veröffentlicht: (2025)
von: Foroutan, Negar, et al.
Veröffentlicht: (2025)
MizanQA: Benchmarking Large Language Models on Moroccan Legal Question Answering
von: Bahaj, Adil, et al.
Veröffentlicht: (2025)
von: Bahaj, Adil, et al.
Veröffentlicht: (2025)
MMToM-QA: Multimodal Theory of Mind Question Answering
von: Jin, Chuanyang, et al.
Veröffentlicht: (2024)
von: Jin, Chuanyang, et al.
Veröffentlicht: (2024)
A Knowledge-Injected Curriculum Pretraining Framework for Question Answering
von: Lin, Xin, et al.
Veröffentlicht: (2024)
von: Lin, Xin, et al.
Veröffentlicht: (2024)
iQUEST: An Iterative Question-Guided Framework for Knowledge Base Question Answering
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
Conv-CoA: Improving Open-domain Question Answering in Large Language Models via Conversational Chain-of-Action
von: Pan, Zhenyu, et al.
Veröffentlicht: (2024)
von: Pan, Zhenyu, et al.
Veröffentlicht: (2024)
MedHopQA: A Disease-Centered Multi-Hop Reasoning Benchmark and Evaluation Framework for LLM-Based Biomedical Question Answering
von: Islamaj, Rezarta, et al.
Veröffentlicht: (2026)
von: Islamaj, Rezarta, et al.
Veröffentlicht: (2026)
BR-TaxQA-R: A Dataset for Question Answering with References for Brazilian Personal Income Tax Law, including case law
von: Júnior, Juvenal Domingos, et al.
Veröffentlicht: (2025)
von: Júnior, Juvenal Domingos, et al.
Veröffentlicht: (2025)
QA-Dragon: Query-Aware Dynamic RAG System for Knowledge-Intensive Visual Question Answering
von: Jiang, Zhuohang, et al.
Veröffentlicht: (2025)
von: Jiang, Zhuohang, et al.
Veröffentlicht: (2025)
LLM-MedQA: Enhancing Medical Question Answering through Case Studies in Large Language Models
von: Yang, Hang, et al.
Veröffentlicht: (2024)
von: Yang, Hang, et al.
Veröffentlicht: (2024)
GNN2R: Weakly-Supervised Rationale-Providing Question Answering over Knowledge Graphs
von: Wang, Ruijie, et al.
Veröffentlicht: (2023)
von: Wang, Ruijie, et al.
Veröffentlicht: (2023)
TrustUQA: A Trustful Framework for Unified Structured Data Question Answering
von: Zhang, Wen, et al.
Veröffentlicht: (2024)
von: Zhang, Wen, et al.
Veröffentlicht: (2024)
Compositional Consistency-Guided Decoding for Three-Way Logical Question Answering
von: Huang, Tianyi, et al.
Veröffentlicht: (2026)
von: Huang, Tianyi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Detecting the Machine: A Comprehensive Benchmark of AI-Generated Text Detectors Across Architectures, Domains, and Adversarial Conditions
von: Baidya, Madhav S., et al.
Veröffentlicht: (2026) -
SensorQA: A Question Answering Benchmark for Daily-Life Monitoring
von: Reichman, Benjamin, et al.
Veröffentlicht: (2025) -
RAG-BioQA: A Retrieval-Augmented Generation Framework for Long-Form Biomedical Question Answering
von: Panchumarthi, Lovely Yeswanth, et al.
Veröffentlicht: (2025) -
Long-Span Question-Answering: Automatic Question Generation and QA-System Ranking via Side-by-Side Evaluation
von: Bohnet, Bernd, et al.
Veröffentlicht: (2024) -
pdfQA: Diverse, Challenging, and Realistic Question Answering over PDFs
von: Schimanski, Tobias, et al.
Veröffentlicht: (2026)