When Choices Become Priors: Contrastive Decoding for Scientific Figure Multiple-Choice QA
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Roh, Taeyun, Jo, Eun-yeong, Jang, Wonjune, Kang, Jaewoo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CLAG: Adaptive Memory Organization via Agent-Driven Clustering for Small Language Model Agents
von: Roh, Taeyun, et al.
Veröffentlicht: (2026)
von: Roh, Taeyun, et al.
Veröffentlicht: (2026)
MolDeTox: Evaluating Language Model's Stepwise Fragment Editing for Molecular Detoxification
von: Park, Jueon, et al.
Veröffentlicht: (2026)
von: Park, Jueon, et al.
Veröffentlicht: (2026)
ToxReason: A Benchmark for Mechanistic Chemical Toxicity Reasoning via Adverse Outcome Pathway
von: Park, Jueon, et al.
Veröffentlicht: (2026)
von: Park, Jueon, et al.
Veröffentlicht: (2026)
Generating Plausible Distractors for Multiple-Choice Questions via Student Choice Prediction
von: Lee, Yooseop, et al.
Veröffentlicht: (2025)
von: Lee, Yooseop, et al.
Veröffentlicht: (2025)
From Multiple-Choice to Extractive QA: A Case Study for English and Arabic
von: Lynn, Teresa, et al.
Veröffentlicht: (2024)
von: Lynn, Teresa, et al.
Veröffentlicht: (2024)
Beyond Multiple Choice: Verifiable OpenQA for Robust Vision-Language RFT
von: Liu, Yesheng, et al.
Veröffentlicht: (2025)
von: Liu, Yesheng, et al.
Veröffentlicht: (2025)
Boosting Process-Correct CoT Reasoning by Modeling Solvability of Multiple-Choice QA
von: Schumann, Raphael, et al.
Veröffentlicht: (2025)
von: Schumann, Raphael, et al.
Veröffentlicht: (2025)
Differentiating Choices via Commonality for Multiple-Choice Question Answering
von: Deng, Wenqing, et al.
Veröffentlicht: (2024)
von: Deng, Wenqing, et al.
Veröffentlicht: (2024)
Improving Score Reliability of Multiple Choice Benchmarks with Consistency Evaluation and Altered Answer Choices
von: Cavalin, Paulo, et al.
Veröffentlicht: (2025)
von: Cavalin, Paulo, et al.
Veröffentlicht: (2025)
Question Difficulty Ranking for Multiple-Choice Reading Comprehension
von: Raina, Vatsal, et al.
Veröffentlicht: (2024)
von: Raina, Vatsal, et al.
Veröffentlicht: (2024)
MCQG-SRefine: Multiple Choice Question Generation and Evaluation with Iterative Self-Critique, Correction, and Comparison Feedback
von: Yao, Zonghai, et al.
Veröffentlicht: (2024)
von: Yao, Zonghai, et al.
Veröffentlicht: (2024)
Multiple Choice Questions: Reasoning Makes Large Language Models (LLMs) More Self-Confident, Especially When They are Wrong
von: Fu, Tairan, et al.
Veröffentlicht: (2025)
von: Fu, Tairan, et al.
Veröffentlicht: (2025)
Wait, that's not an option: LLMs Robustness with Incorrect Multiple-Choice Options
von: Góral, Gracjan, et al.
Veröffentlicht: (2024)
von: Góral, Gracjan, et al.
Veröffentlicht: (2024)
Plausibly Problematic Questions in Multiple-Choice Benchmarks for Commonsense Reasoning
von: Palta, Shramay, et al.
Veröffentlicht: (2024)
von: Palta, Shramay, et al.
Veröffentlicht: (2024)
Axiomatic Choice
von: Abramowitz, Ben, et al.
Veröffentlicht: (2025)
von: Abramowitz, Ben, et al.
Veröffentlicht: (2025)
Mind the (DH) Gap! A Contrast in Risky Choices Between Reasoning and Conversational LLMs
von: Ge, Luise, et al.
Veröffentlicht: (2026)
von: Ge, Luise, et al.
Veröffentlicht: (2026)
Extending RLVR to Open-Ended Tasks via Verifiable Multiple-Choice Reformulation
von: Zhang, Mengyu, et al.
Veröffentlicht: (2025)
von: Zhang, Mengyu, et al.
Veröffentlicht: (2025)
Biomedical Entity Linking as Multiple Choice Question Answering
von: Lin, Zhenxi, et al.
Veröffentlicht: (2024)
von: Lin, Zhenxi, et al.
Veröffentlicht: (2024)
Option-ID Based Elimination For Multiple Choice Questions
von: Zhu, Zhenhao, et al.
Veröffentlicht: (2025)
von: Zhu, Zhenhao, et al.
Veröffentlicht: (2025)
Orchestrating LLM Agents for Scientific Research: A Pilot Study of Multiple Choice Question (MCQ) Generation and Evaluation
von: An, Yuan
Veröffentlicht: (2026)
von: An, Yuan
Veröffentlicht: (2026)
Adaptive Sparse Allocation with Mutual Choice & Feature Choice Sparse Autoencoders
von: Ayonrinde, Kola
Veröffentlicht: (2024)
von: Ayonrinde, Kola
Veröffentlicht: (2024)
A Vietnamese Dataset for Text Segmentation and Multiple Choices Reading Comprehension
von: Hai, Toan Nguyen, et al.
Veröffentlicht: (2025)
von: Hai, Toan Nguyen, et al.
Veröffentlicht: (2025)
Automated Generation and Tagging of Knowledge Components from Multiple-Choice Questions
von: Moore, Steven, et al.
Veröffentlicht: (2024)
von: Moore, Steven, et al.
Veröffentlicht: (2024)
Self-Correcting Large Language Models: Generation vs. Multiple Choice
von: Rahmani, Hossein A., et al.
Veröffentlicht: (2025)
von: Rahmani, Hossein A., et al.
Veröffentlicht: (2025)
Applying IRT to Distinguish Between Human and Generative AI Responses to Multiple-Choice Assessments
von: Strugatski, Alona, et al.
Veröffentlicht: (2024)
von: Strugatski, Alona, et al.
Veröffentlicht: (2024)
Answer Matching Outperforms Multiple Choice for Language Model Evaluation
von: Chandak, Nikhil, et al.
Veröffentlicht: (2025)
von: Chandak, Nikhil, et al.
Veröffentlicht: (2025)
Multiple Choice Learning of Low-Rank Adapters for Language Modeling
von: Letzelter, Victor, et al.
Veröffentlicht: (2025)
von: Letzelter, Victor, et al.
Veröffentlicht: (2025)
Beyond Multiple-Choice Accuracy: Real-World Challenges of Implementing Large Language Models in Healthcare
von: Yang, Yifan, et al.
Veröffentlicht: (2024)
von: Yang, Yifan, et al.
Veröffentlicht: (2024)
Feedback Indices to Evaluate LLM Responses to Rebuttals for Multiple Choice Type Questions
von: Dunlap, Justin C., et al.
Veröffentlicht: (2026)
von: Dunlap, Justin C., et al.
Veröffentlicht: (2026)
Answer, Assemble, Ace: Understanding How LMs Answer Multiple Choice Questions
von: Wiegreffe, Sarah, et al.
Veröffentlicht: (2024)
von: Wiegreffe, Sarah, et al.
Veröffentlicht: (2024)
Mitigating Easy Option Bias in Multiple-Choice Question Answering
von: Zhang, Hao, et al.
Veröffentlicht: (2025)
von: Zhang, Hao, et al.
Veröffentlicht: (2025)
PriorRG: Prior-Guided Contrastive Pre-training and Coarse-to-Fine Decoding for Chest X-ray Report Generation
von: Liu, Kang, et al.
Veröffentlicht: (2025)
von: Liu, Kang, et al.
Veröffentlicht: (2025)
Pattern Recognition or Medical Knowledge? The Problem with Multiple-Choice Questions in Medicine
von: Griot, Maxime, et al.
Veröffentlicht: (2024)
von: Griot, Maxime, et al.
Veröffentlicht: (2024)
GeoChallenge: A Multi-Answer Multiple-Choice Benchmark for Geometric Reasoning with Diagrams
von: Zhang, Yushun, et al.
Veröffentlicht: (2026)
von: Zhang, Yushun, et al.
Veröffentlicht: (2026)
Multiple-Choice Question Generation Using Large Language Models: Methodology and Educator Insights
von: Biancini, Giorgio, et al.
Veröffentlicht: (2025)
von: Biancini, Giorgio, et al.
Veröffentlicht: (2025)
Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control
von: Ye, Yuanchang
Veröffentlicht: (2025)
von: Ye, Yuanchang
Veröffentlicht: (2025)
Towards Shutdownable Agents via Stochastic Choice
von: Thornley, Elliott, et al.
Veröffentlicht: (2024)
von: Thornley, Elliott, et al.
Veröffentlicht: (2024)
Guidelines For The Choice Of The Baseline in XAI Attribution Methods
von: Morasso, Cristian, et al.
Veröffentlicht: (2025)
von: Morasso, Cristian, et al.
Veröffentlicht: (2025)
Discourse-Aware Scientific Paper Recommendation via QA-Style Summarization and Multi-Level Contrastive Learning
von: Wang, Shenghua, et al.
Veröffentlicht: (2025)
von: Wang, Shenghua, et al.
Veröffentlicht: (2025)
Anchored Answers: Unravelling Positional Bias in GPT-2's Multiple-Choice Questions
von: Li, Ruizhe, et al.
Veröffentlicht: (2024)
von: Li, Ruizhe, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CLAG: Adaptive Memory Organization via Agent-Driven Clustering for Small Language Model Agents
von: Roh, Taeyun, et al.
Veröffentlicht: (2026) -
MolDeTox: Evaluating Language Model's Stepwise Fragment Editing for Molecular Detoxification
von: Park, Jueon, et al.
Veröffentlicht: (2026) -
ToxReason: A Benchmark for Mechanistic Chemical Toxicity Reasoning via Adverse Outcome Pathway
von: Park, Jueon, et al.
Veröffentlicht: (2026) -
Generating Plausible Distractors for Multiple-Choice Questions via Student Choice Prediction
von: Lee, Yooseop, et al.
Veröffentlicht: (2025) -
From Multiple-Choice to Extractive QA: A Case Study for English and Arabic
von: Lynn, Teresa, et al.
Veröffentlicht: (2024)