Wrong Answers Can Also Be Useful: PlausibleQA -- A Large-Scale QA Dataset with Answer Plausibility Scores
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mozafari, Jamshid, Abdallah, Abdelrahman, Piryani, Bhawna, Jatowt, Adam |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Question Difficulty Estimation for Large Language Models via Answer Plausibility Scoring
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
HintEval: A Comprehensive Framework for Hint Generation and Evaluation for Questions
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2025)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2025)
DeAR: Dual-Stage Document Reranking with Reasoning Agents via LLM Distillation
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
Exploring Hint Generation Approaches in Open-Domain Question Answering
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
It's High Time: A Survey of Temporal Question Answering
von: Piryani, Bhawna, et al.
Veröffentlicht: (2025)
von: Piryani, Bhawna, et al.
Veröffentlicht: (2025)
Rankify: A Comprehensive Python Toolkit for Retrieval, Re-Ranking, and Retrieval-Augmented Generation
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
How Good are LLM-based Rerankers? An Empirical Analysis of State-of-the-Art Reranking Models
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
Context Convergence Improves Answering Inferential Questions
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
ASRank: Zero-Shot Re-Ranking with Answer Scent for Document Retrieval
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
Evaluating Answer Reranking Strategies in Time-sensitive Question Answering
von: Kardan, Mehmet, et al.
Veröffentlicht: (2025)
von: Kardan, Mehmet, et al.
Veröffentlicht: (2025)
ChroniclingAmericaQA: A Large-scale Question Answering Dataset based on Historical American Newspaper Pages
von: Piryani, Bhawna, et al.
Veröffentlicht: (2024)
von: Piryani, Bhawna, et al.
Veröffentlicht: (2024)
RankArena: A Unified Platform for Evaluating Retrieval, Reranking and RAG with Human and LLM Feedback
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
TempRetriever: Fusion-based Temporal Dense Passage Retrieval for Time-Sensitive Questions
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
Detecting Temporal Ambiguity in Questions
von: Piryani, Bhawna, et al.
Veröffentlicht: (2024)
von: Piryani, Bhawna, et al.
Veröffentlicht: (2024)
WikiHint: A Human-Annotated Dataset for Hint Ranking and Generation
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
ArabicaQA: A Comprehensive Dataset for Arabic Question Answering
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
BracketRank: Large Language Model Document Ranking via Reasoning-based Competitive Elimination
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2026)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2026)
Pretraining Exposure Explains Popularity Judgments in Large Language Models
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
From Retrieval to Generation: Comparing Different Approaches
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
Evaluating Robustness of LLMs in Question Answering on Multilingual Noisy OCR Data
von: Piryani, Bhawna, et al.
Veröffentlicht: (2025)
von: Piryani, Bhawna, et al.
Veröffentlicht: (2025)
SustainableQA: A Comprehensive Question Answering Dataset for Corporate Sustainability and EU Taxonomy Reporting
von: Ali, Mohammed, et al.
Veröffentlicht: (2025)
von: Ali, Mohammed, et al.
Veröffentlicht: (2025)
PARSE: An Open-Domain Reasoning Question Answering Benchmark for Persian
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
Inferential Question Answering
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
DynRank: Improving Passage Retrieval with Dynamic Zero-Shot Prompting Based on Question Classification
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
How often do Answers Change? Estimating Recency Requirements in Question Answering
von: Piryani, Bhawna, et al.
Veröffentlicht: (2026)
von: Piryani, Bhawna, et al.
Veröffentlicht: (2026)
Two-Stage Quranic QA via Ensemble Retrieval and Instruction-Tuned Answer Extraction
von: Basem, Mohamed, et al.
Veröffentlicht: (2025)
von: Basem, Mohamed, et al.
Veröffentlicht: (2025)
Estimating the Usefulness of Clarifying Questions and Answers for Conversational Search
von: Sekulić, Ivan, et al.
Veröffentlicht: (2024)
von: Sekulić, Ivan, et al.
Veröffentlicht: (2024)
A Study into Investigating Temporal Robustness of LLMs
von: Wallat, Jonas, et al.
Veröffentlicht: (2025)
von: Wallat, Jonas, et al.
Veröffentlicht: (2025)
Retrieval Improvements Do Not Guarantee Better Answers: A Study of RAG for AI Policy QA
von: Mathur, Saahil, et al.
Veröffentlicht: (2026)
von: Mathur, Saahil, et al.
Veröffentlicht: (2026)
Analyzing the Role of Context in Forecasting with Large Language Models
von: Mutschlechner, Gerrit, et al.
Veröffentlicht: (2025)
von: Mutschlechner, Gerrit, et al.
Veröffentlicht: (2025)
Navigating Tomorrow: Reliably Assessing Large Language Models Performance on Future Event Prediction
von: Nako, Petraq, et al.
Veröffentlicht: (2025)
von: Nako, Petraq, et al.
Veröffentlicht: (2025)
ExpertGenQA: Open-ended QA generation in Specialized Domains
von: Shahgir, Haz Sameen, et al.
Veröffentlicht: (2025)
von: Shahgir, Haz Sameen, et al.
Veröffentlicht: (2025)
Negative Sampling Techniques in Information Retrieval: A Survey
von: Wischounig, Laurin, et al.
Veröffentlicht: (2026)
von: Wischounig, Laurin, et al.
Veröffentlicht: (2026)
Towards Self-Contained Answers: Entity-Based Answer Rewriting in Conversational Search
von: Sekulić, Ivan, et al.
Veröffentlicht: (2024)
von: Sekulić, Ivan, et al.
Veröffentlicht: (2024)
ComplexTempQA:A 100m Dataset for Complex Temporal Question Answering
von: Gruber, Raphael, et al.
Veröffentlicht: (2024)
von: Gruber, Raphael, et al.
Veröffentlicht: (2024)
Cross-Language Approach for Quranic QA
von: Oshallah, Islam, et al.
Veröffentlicht: (2025)
von: Oshallah, Islam, et al.
Veröffentlicht: (2025)
Enhancing Answer Attribution for Faithful Text Generation with Large Language Models
von: Vladika, Juraj, et al.
Veröffentlicht: (2024)
von: Vladika, Juraj, et al.
Veröffentlicht: (2024)
Optimized Quran Passage Retrieval Using an Expanded QA Dataset and Fine-Tuned Language Models
von: Basem, Mohamed, et al.
Veröffentlicht: (2024)
von: Basem, Mohamed, et al.
Veröffentlicht: (2024)
Neural at ArchEHR-QA 2026: One Method Fits All: Unified Prompt Optimization for Clinical QA over EHRs
von: Majeedi, Abrar, et al.
Veröffentlicht: (2026)
von: Majeedi, Abrar, et al.
Veröffentlicht: (2026)
Context Embeddings for Efficient Answer Generation in RAG
von: Rau, David, et al.
Veröffentlicht: (2024)
von: Rau, David, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Question Difficulty Estimation for Large Language Models via Answer Plausibility Scoring
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026) -
HintEval: A Comprehensive Framework for Hint Generation and Evaluation for Questions
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2025) -
DeAR: Dual-Stage Document Reranking with Reasoning Agents via LLM Distillation
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025) -
Exploring Hint Generation Approaches in Open-Domain Question Answering
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024) -
It's High Time: A Survey of Temporal Question Answering
von: Piryani, Bhawna, et al.
Veröffentlicht: (2025)