CondAmbigQA: A Benchmark and Dataset for Conditional Ambiguous Question Answering
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Zongxi, Li, Yang, Xie, Haoran, Qin, S. Joe |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PolQA: Polish Question Answering Dataset
von: Rybak, Piotr, et al.
Veröffentlicht: (2022)
von: Rybak, Piotr, et al.
Veröffentlicht: (2022)
A Gaze-grounded Visual Question Answering Dataset for Clarifying Ambiguous Japanese Questions
von: Inadumi, Shun, et al.
Veröffentlicht: (2024)
von: Inadumi, Shun, et al.
Veröffentlicht: (2024)
DisastQA: A Comprehensive Benchmark for Evaluating Question Answering in Disaster Management
von: Chen, Zhitong, et al.
Veröffentlicht: (2026)
von: Chen, Zhitong, et al.
Veröffentlicht: (2026)
FoQA: A Faroese Question-Answering Dataset
von: Simonsen, Annika, et al.
Veröffentlicht: (2025)
von: Simonsen, Annika, et al.
Veröffentlicht: (2025)
GRS-QA -- Graph Reasoning-Structured Question Answering Dataset
von: Pahilajani, Anish, et al.
Veröffentlicht: (2024)
von: Pahilajani, Anish, et al.
Veröffentlicht: (2024)
MobQA: A Benchmark Dataset for Semantic Understanding of Human Mobility Data through Question Answering
von: Asano, Hikaru, et al.
Veröffentlicht: (2025)
von: Asano, Hikaru, et al.
Veröffentlicht: (2025)
AfriMed-QA: A Pan-African, Multi-Specialty, Medical Question-Answering Benchmark Dataset
von: Olatunji, Tobi, et al.
Veröffentlicht: (2024)
von: Olatunji, Tobi, et al.
Veröffentlicht: (2024)
KET-QA: A Dataset for Knowledge Enhanced Table Question Answering
von: Hu, Mengkang, et al.
Veröffentlicht: (2024)
von: Hu, Mengkang, et al.
Veröffentlicht: (2024)
ArabicaQA: A Comprehensive Dataset for Arabic Question Answering
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
BEnQA: A Question Answering and Reasoning Benchmark for Bengali and English
von: Shafayat, Sheikh, et al.
Veröffentlicht: (2024)
von: Shafayat, Sheikh, et al.
Veröffentlicht: (2024)
RepLiQA: A Question-Answering Dataset for Benchmarking LLMs on Unseen Reference Content
von: Monteiro, Joao, et al.
Veröffentlicht: (2024)
von: Monteiro, Joao, et al.
Veröffentlicht: (2024)
FinTextQA: A Dataset for Long-form Financial Question Answering
von: Chen, Jian, et al.
Veröffentlicht: (2024)
von: Chen, Jian, et al.
Veröffentlicht: (2024)
SyllabusQA: A Course Logistics Question Answering Dataset
von: Fernandez, Nigel, et al.
Veröffentlicht: (2024)
von: Fernandez, Nigel, et al.
Veröffentlicht: (2024)
DashboardQA: Benchmarking Multimodal Agents for Question Answering on Interactive Dashboards
von: Kartha, Aaryaman, et al.
Veröffentlicht: (2025)
von: Kartha, Aaryaman, et al.
Veröffentlicht: (2025)
RespondeoQA: a Benchmark for Bilingual Latin-English Question Answering
von: Hudspeth, Marisa, et al.
Veröffentlicht: (2026)
von: Hudspeth, Marisa, et al.
Veröffentlicht: (2026)
Factuality or Fiction? Benchmarking Modern LLMs on Ambiguous QA with Citations
von: Patel, Maya, et al.
Veröffentlicht: (2024)
von: Patel, Maya, et al.
Veröffentlicht: (2024)
AmharicStoryQA: A Multicultural Story Question Answering Benchmark in Amharic
von: Azime, Israel Abebe, et al.
Veröffentlicht: (2026)
von: Azime, Israel Abebe, et al.
Veröffentlicht: (2026)
SensorQA: A Question Answering Benchmark for Daily-Life Monitoring
von: Reichman, Benjamin, et al.
Veröffentlicht: (2025)
von: Reichman, Benjamin, et al.
Veröffentlicht: (2025)
ASTRA-QA: A Benchmark for Abstract Question Answering over Documents
von: Wang, Shu, et al.
Veröffentlicht: (2026)
von: Wang, Shu, et al.
Veröffentlicht: (2026)
CUS-QA: Local-Knowledge-Oriented Open-Ended Question Answering Dataset
von: Libovický, Jindřich, et al.
Veröffentlicht: (2025)
von: Libovický, Jindřich, et al.
Veröffentlicht: (2025)
JDocQA: Japanese Document Question Answering Dataset for Generative Language Models
von: Onami, Eri, et al.
Veröffentlicht: (2024)
von: Onami, Eri, et al.
Veröffentlicht: (2024)
ComplexTempQA:A 100m Dataset for Complex Temporal Question Answering
von: Gruber, Raphael, et al.
Veröffentlicht: (2024)
von: Gruber, Raphael, et al.
Veröffentlicht: (2024)
NovelQA: Benchmarking Question Answering on Documents Exceeding 200K Tokens
von: Wang, Cunxiang, et al.
Veröffentlicht: (2024)
von: Wang, Cunxiang, et al.
Veröffentlicht: (2024)
NCTB-QA: A Large-Scale Bangla Educational Question Answering Dataset and Benchmarking Performance
von: Eyasir, Abrar, et al.
Veröffentlicht: (2026)
von: Eyasir, Abrar, et al.
Veröffentlicht: (2026)
MedExQA: Medical Question Answering Benchmark with Multiple Explanations
von: Kim, Yunsoo, et al.
Veröffentlicht: (2024)
von: Kim, Yunsoo, et al.
Veröffentlicht: (2024)
2D Matryoshka Sentence Embeddings
von: Li, Xianming, et al.
Veröffentlicht: (2024)
von: Li, Xianming, et al.
Veröffentlicht: (2024)
ReasonTabQA: A Comprehensive Benchmark for Table Question Answering from Real World Industrial Scenarios
von: Pan, Changzai, et al.
Veröffentlicht: (2026)
von: Pan, Changzai, et al.
Veröffentlicht: (2026)
ReCoQA: A Benchmark for Tool-Augmented and Multi-Step Reasoning in Real Estate Question and Answering
von: Zhang, Yindong, et al.
Veröffentlicht: (2026)
von: Zhang, Yindong, et al.
Veröffentlicht: (2026)
MedExpQA: Multilingual Benchmarking of Large Language Models for Medical Question Answering
von: Alonso, Iñigo, et al.
Veröffentlicht: (2024)
von: Alonso, Iñigo, et al.
Veröffentlicht: (2024)
AMBROSIA: A Benchmark for Parsing Ambiguous Questions into Database Queries
von: Saparina, Irina, et al.
Veröffentlicht: (2024)
von: Saparina, Irina, et al.
Veröffentlicht: (2024)
MedEthicsQA: A Comprehensive Question Answering Benchmark for Medical Ethics Evaluation of LLMs
von: Wei, Jianhui, et al.
Veröffentlicht: (2025)
von: Wei, Jianhui, et al.
Veröffentlicht: (2025)
Diversify-verify-adapt: Efficient and Robust Retrieval-Augmented Ambiguous Question Answering
von: In, Yeonjun, et al.
Veröffentlicht: (2024)
von: In, Yeonjun, et al.
Veröffentlicht: (2024)
PeerQA: A Scientific Question Answering Dataset from Peer Reviews
von: Baumgärtner, Tim, et al.
Veröffentlicht: (2025)
von: Baumgärtner, Tim, et al.
Veröffentlicht: (2025)
DebateQA: Evaluating Question Answering on Debatable Knowledge
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
OMoS-QA: A Dataset for Cross-Lingual Extractive Question Answering in a German Migration Context
von: Kleinle, Steffen, et al.
Veröffentlicht: (2024)
von: Kleinle, Steffen, et al.
Veröffentlicht: (2024)
LaMP-QA: A Benchmark for Personalized Long-form Question Answering
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
DEEPAMBIGQA: Ambiguous Multi-hop Questions for Benchmarking LLM Answer Completeness
von: Ji, Jiabao, et al.
Veröffentlicht: (2025)
von: Ji, Jiabao, et al.
Veröffentlicht: (2025)
VisualSimpleQA: A Benchmark for Decoupled Evaluation of Large Vision-Language Models in Fact-Seeking Question Answering
von: Wang, Yanling, et al.
Veröffentlicht: (2025)
von: Wang, Yanling, et al.
Veröffentlicht: (2025)
PRIV-QA: Privacy-Preserving Question Answering for Cloud Large Language Models
von: Li, Guangwei, et al.
Veröffentlicht: (2025)
von: Li, Guangwei, et al.
Veröffentlicht: (2025)
ChroniclingAmericaQA: A Large-scale Question Answering Dataset based on Historical American Newspaper Pages
von: Piryani, Bhawna, et al.
Veröffentlicht: (2024)
von: Piryani, Bhawna, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
PolQA: Polish Question Answering Dataset
von: Rybak, Piotr, et al.
Veröffentlicht: (2022) -
A Gaze-grounded Visual Question Answering Dataset for Clarifying Ambiguous Japanese Questions
von: Inadumi, Shun, et al.
Veröffentlicht: (2024) -
DisastQA: A Comprehensive Benchmark for Evaluating Question Answering in Disaster Management
von: Chen, Zhitong, et al.
Veröffentlicht: (2026) -
FoQA: A Faroese Question-Answering Dataset
von: Simonsen, Annika, et al.
Veröffentlicht: (2025) -
GRS-QA -- Graph Reasoning-Structured Question Answering Dataset
von: Pahilajani, Anish, et al.
Veröffentlicht: (2024)