NeoQA: Evidence-based Question Answering with Generated News Events
Fuente:
arXiv
Guardado en:
| Autores principales: | Glockner, Max, Jiang, Xiang, Ribeiro, Leonardo F. R., Gurevych, Iryna, Dreyer, Markus |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Grounding Fallacies Misrepresenting Scientific Publications in Evidence
por: Glockner, Max, et al.
Publicado: (2024)
por: Glockner, Max, et al.
Publicado: (2024)
PeerQA: A Scientific Question Answering Dataset from Peer Reviews
por: Baumgärtner, Tim, et al.
Publicado: (2025)
por: Baumgärtner, Tim, et al.
Publicado: (2025)
M2QA: Multi-domain Multilingual Question Answering
por: Engländer, Leon, et al.
Publicado: (2024)
por: Engländer, Leon, et al.
Publicado: (2024)
ConspirED: A Dataset for Cognitive Traits of Conspiracy Theories and Large Language Model Safety
por: Bates, Luke, et al.
Publicado: (2025)
por: Bates, Luke, et al.
Publicado: (2025)
Self-Rationalization in the Wild: A Large Scale Out-of-Distribution Evaluation on NLI-related tasks
por: Yang, Jing, et al.
Publicado: (2025)
por: Yang, Jing, et al.
Publicado: (2025)
Missci: Reconstructing Fallacies in Misrepresented Science
por: Glockner, Max, et al.
Publicado: (2024)
por: Glockner, Max, et al.
Publicado: (2024)
Localizing and Mitigating Errors in Long-form Question Answering
por: Sachdeva, Rachneet, et al.
Publicado: (2024)
por: Sachdeva, Rachneet, et al.
Publicado: (2024)
DARA: Decomposition-Alignment-Reasoning Autonomous Language Agent for Question Answering over Knowledge Graphs
por: Fang, Haishuo, et al.
Publicado: (2024)
por: Fang, Haishuo, et al.
Publicado: (2024)
SciCoQA: Quality Assurance for Scientific Paper--Code Alignment
por: Baumgärtner, Tim, et al.
Publicado: (2026)
por: Baumgärtner, Tim, et al.
Publicado: (2026)
Automatic Reviewers Fail to Detect Faulty Reasoning in Research Papers: A New Counterfactual Evaluation Framework
por: Dycke, Nils, et al.
Publicado: (2025)
por: Dycke, Nils, et al.
Publicado: (2025)
Turning Logic Against Itself : Probing Model Defenses Through Contrastive Questions
por: Sachdeva, Rachneet, et al.
Publicado: (2025)
por: Sachdeva, Rachneet, et al.
Publicado: (2025)
HistoryBankQA: Multilingual Temporal Question Answering on Historical Events
por: Mandal, Biswadip, et al.
Publicado: (2025)
por: Mandal, Biswadip, et al.
Publicado: (2025)
Expert Preference-based Evaluation of Automated Related Work Generation
por: Şahinuç, Furkan, et al.
Publicado: (2025)
por: Şahinuç, Furkan, et al.
Publicado: (2025)
Towards Better Question Generation in QA-based Event Extraction
por: Hong, Zijin, et al.
Publicado: (2024)
por: Hong, Zijin, et al.
Publicado: (2024)
PolQA: Polish Question Answering Dataset
por: Rybak, Piotr, et al.
Publicado: (2022)
por: Rybak, Piotr, et al.
Publicado: (2022)
Citation Failure: Definition, Analysis and Efficient Mitigation
por: Buchmann, Jan, et al.
Publicado: (2025)
por: Buchmann, Jan, et al.
Publicado: (2025)
Like a Good Nearest Neighbor: Practical Content Moderation and Text Classification
por: Bates, Luke, et al.
Publicado: (2023)
por: Bates, Luke, et al.
Publicado: (2023)
Enhancing Depression Detection via Question-wise Modality Fusion
por: Mandal, Aishik, et al.
Publicado: (2025)
por: Mandal, Aishik, et al.
Publicado: (2025)
Dive into the Chasm: Probing the Gap between In- and Cross-Topic Generalization
por: Waldis, Andreas, et al.
Publicado: (2024)
por: Waldis, Andreas, et al.
Publicado: (2024)
IRCoder: Intermediate Representations Make Language Models Robust Multilingual Code Generators
por: Paul, Indraneil, et al.
Publicado: (2024)
por: Paul, Indraneil, et al.
Publicado: (2024)
pdfQA: Diverse, Challenging, and Realistic Question Answering over PDFs
por: Schimanski, Tobias, et al.
Publicado: (2026)
por: Schimanski, Tobias, et al.
Publicado: (2026)
DebateQA: Evaluating Question Answering on Debatable Knowledge
por: Xu, Rongwu, et al.
Publicado: (2024)
por: Xu, Rongwu, et al.
Publicado: (2024)
JDocQA: Japanese Document Question Answering Dataset for Generative Language Models
por: Onami, Eri, et al.
Publicado: (2024)
por: Onami, Eri, et al.
Publicado: (2024)
NewsRECON: News article REtrieval for image CONtextualization
por: Tonglet, Jonathan, et al.
Publicado: (2026)
por: Tonglet, Jonathan, et al.
Publicado: (2026)
Measuring Retrieval Complexity in Question Answering Systems
por: Gabburo, Matteo, et al.
Publicado: (2024)
por: Gabburo, Matteo, et al.
Publicado: (2024)
GRS-QA -- Graph Reasoning-Structured Question Answering Dataset
por: Pahilajani, Anish, et al.
Publicado: (2024)
por: Pahilajani, Anish, et al.
Publicado: (2024)
Hierarchical Latent Structures in Data Generation Process Unify Mechanistic Phenomena across Scale
por: Rohweder, Jonas, et al.
Publicado: (2026)
por: Rohweder, Jonas, et al.
Publicado: (2026)
MAGneT: Coordinated Multi-Agent Generation of Synthetic Multi-Turn Mental Health Counseling Sessions
por: Mandal, Aishik, et al.
Publicado: (2025)
por: Mandal, Aishik, et al.
Publicado: (2025)
FoQA: A Faroese Question-Answering Dataset
por: Simonsen, Annika, et al.
Publicado: (2025)
por: Simonsen, Annika, et al.
Publicado: (2025)
Memory-QA: Answering Recall Questions Based on Multimodal Memories
por: Jiang, Hongda, et al.
Publicado: (2025)
por: Jiang, Hongda, et al.
Publicado: (2025)
Constrained C-Test Generation via Mixed-Integer Programming
por: Lee, Ji-Ung, et al.
Publicado: (2024)
por: Lee, Ji-Ung, et al.
Publicado: (2024)
Commitment Checklist: Auditing Author Commitments in Peer Review
por: Chen, Chung-Chi, et al.
Publicado: (2026)
por: Chen, Chung-Chi, et al.
Publicado: (2026)
DOCE: Finding the Sweet Spot for Execution-Based Code Generation
por: Li, Haau-Sing, et al.
Publicado: (2024)
por: Li, Haau-Sing, et al.
Publicado: (2024)
Systematic Task Exploration with LLMs: A Study in Citation Text Generation
por: Şahinuç, Furkan, et al.
Publicado: (2024)
por: Şahinuç, Furkan, et al.
Publicado: (2024)
MMToM-QA: Multimodal Theory of Mind Question Answering
por: Jin, Chuanyang, et al.
Publicado: (2024)
por: Jin, Chuanyang, et al.
Publicado: (2024)
DashboardQA: Benchmarking Multimodal Agents for Question Answering on Interactive Dashboards
por: Kartha, Aaryaman, et al.
Publicado: (2025)
por: Kartha, Aaryaman, et al.
Publicado: (2025)
RespondeoQA: a Benchmark for Bilingual Latin-English Question Answering
por: Hudspeth, Marisa, et al.
Publicado: (2026)
por: Hudspeth, Marisa, et al.
Publicado: (2026)
BEnQA: A Question Answering and Reasoning Benchmark for Bengali and English
por: Shafayat, Sheikh, et al.
Publicado: (2024)
por: Shafayat, Sheikh, et al.
Publicado: (2024)
KET-QA: A Dataset for Knowledge Enhanced Table Question Answering
por: Hu, Mengkang, et al.
Publicado: (2024)
por: Hu, Mengkang, et al.
Publicado: (2024)
NewsQs: Multi-Source Question Generation for the Inquiring Mind
por: Hwang, Alyssa, et al.
Publicado: (2024)
por: Hwang, Alyssa, et al.
Publicado: (2024)
Ejemplares similares
-
Grounding Fallacies Misrepresenting Scientific Publications in Evidence
por: Glockner, Max, et al.
Publicado: (2024) -
PeerQA: A Scientific Question Answering Dataset from Peer Reviews
por: Baumgärtner, Tim, et al.
Publicado: (2025) -
M2QA: Multi-domain Multilingual Question Answering
por: Engländer, Leon, et al.
Publicado: (2024) -
ConspirED: A Dataset for Cognitive Traits of Conspiracy Theories and Large Language Model Safety
por: Bates, Luke, et al.
Publicado: (2025) -
Self-Rationalization in the Wild: A Large Scale Out-of-Distribution Evaluation on NLI-related tasks
por: Yang, Jing, et al.
Publicado: (2025)