LLMs May Perform MCQA by Selecting the Least Incorrect Option
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Haochun, Zhao, Sendong, Qiang, Zewen, Xi, Nuwa, Qin, Bing, Liu, Ting |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Beyond Direct Diagnosis: LLM-based Multi-Specialist Agent Consultation for Automatic Diagnosis
di: Wang, Haochun, et al.
Pubblicazione: (2024)
di: Wang, Haochun, et al.
Pubblicazione: (2024)
Uncovering the Role of Initial Saliency in U-Shaped Attention Bias: Scaling Initial Token Weight for Enhanced Long-Text Processing
di: Qiang, Zewen, et al.
Pubblicazione: (2025)
di: Qiang, Zewen, et al.
Pubblicazione: (2025)
Manifold-based Verbalizer Space Re-embedding for Tuning-free Prompt-based Classification
di: Wang, Haochun, et al.
Pubblicazione: (2023)
di: Wang, Haochun, et al.
Pubblicazione: (2023)
Knowledge-tuning Large Language Models with Structured Medical Knowledge Bases for Reliable Response Generation in Chinese
di: Wang, Haochun, et al.
Pubblicazione: (2023)
di: Wang, Haochun, et al.
Pubblicazione: (2023)
AS-ES Learning: Towards Efficient CoT Learning in Small Models
di: Xi, Nuwa, et al.
Pubblicazione: (2024)
di: Xi, Nuwa, et al.
Pubblicazione: (2024)
ArcAligner: Adaptive Recursive Aligner for Compressed Context Embeddings in RAG
di: Li, Jianbo, et al.
Pubblicazione: (2026)
di: Li, Jianbo, et al.
Pubblicazione: (2026)
Beyond Frameworks: Unpacking Collaboration Strategies in Multi-Agent Systems
di: Wang, Haochun, et al.
Pubblicazione: (2025)
di: Wang, Haochun, et al.
Pubblicazione: (2025)
When Correct Beliefs Collapse: Epistemic Resilience of LLMs under Clinical Pressure
di: Xiao, Boyu, et al.
Pubblicazione: (2026)
di: Xiao, Boyu, et al.
Pubblicazione: (2026)
CoCoA: Collaborative Chain-of-Agents for Parametric-Retrieved Knowledge Synergy
di: Jiang, Yi, et al.
Pubblicazione: (2025)
di: Jiang, Yi, et al.
Pubblicazione: (2025)
Easier to Judge than to Find: Predicting In-Context Learning Success for Demonstration Selection
di: Wang, Haochun, et al.
Pubblicazione: (2026)
di: Wang, Haochun, et al.
Pubblicazione: (2026)
From Artificially Real to Real: Leveraging Pseudo Data from Large Language Models for Low-Resource Molecule Discovery
di: Chen, Yuhan, et al.
Pubblicazione: (2023)
di: Chen, Yuhan, et al.
Pubblicazione: (2023)
Wait, that's not an option: LLMs Robustness with Incorrect Multiple-Choice Options
di: Góral, Gracjan, et al.
Pubblicazione: (2024)
di: Góral, Gracjan, et al.
Pubblicazione: (2024)
MolFusion: Multimodal Fusion Learning for Molecular Representations via Multi-granularity Views
di: Cai, Muzhen, et al.
Pubblicazione: (2024)
di: Cai, Muzhen, et al.
Pubblicazione: (2024)
SL-BiLEM: Structured Learnable Behavior-in-the-Loop Epidemic Modeling for Forecasting and Policy Evaluation
di: Wang, Haochun, et al.
Pubblicazione: (2026)
di: Wang, Haochun, et al.
Pubblicazione: (2026)
MCQA-Eval: Efficient Confidence Evaluation in NLG with Gold-Standard Correctness Labels
di: Liu, Xiaoou, et al.
Pubblicazione: (2025)
di: Liu, Xiaoou, et al.
Pubblicazione: (2025)
GSEM: Graph-based Self-Evolving Memory for Experience Augmented Clinical Reasoning
di: Han, Xiao, et al.
Pubblicazione: (2026)
di: Han, Xiao, et al.
Pubblicazione: (2026)
Orchestrating Intelligence: Confidence-Aware Routing for Efficient Multi-Agent Collaboration across Multi-Scale Models
di: Wang, Jingbo, et al.
Pubblicazione: (2026)
di: Wang, Jingbo, et al.
Pubblicazione: (2026)
META-RAG: Meta-Analysis-Inspired Evidence-Re-Ranking Method for Retrieval-Augmented Generation in Evidence-Based Medicine
di: Sun, Mengzhou, et al.
Pubblicazione: (2025)
di: Sun, Mengzhou, et al.
Pubblicazione: (2025)
GainRAG: Preference Alignment in Retrieval-Augmented Generation through Gain Signal Synthesis
di: Jiang, Yi, et al.
Pubblicazione: (2025)
di: Jiang, Yi, et al.
Pubblicazione: (2025)
OptiSet: Unified Optimizing Set Selection and Ranking for Retrieval-Augmented Generation
di: Jiang, Yi, et al.
Pubblicazione: (2026)
di: Jiang, Yi, et al.
Pubblicazione: (2026)
MolTailor: Tailoring Chemical Molecular Representation to Specific Tasks via Text Prompts
di: Guo, Haoqiang, et al.
Pubblicazione: (2024)
di: Guo, Haoqiang, et al.
Pubblicazione: (2024)
ExpeTrans: LLMs Are Experiential Transfer Learners
di: Gao, Jinglong, et al.
Pubblicazione: (2025)
di: Gao, Jinglong, et al.
Pubblicazione: (2025)
Optimal-Agent-Selection: State-Aware Routing Framework for Efficient Multi-Agent Collaboration
di: Wang, Jingbo, et al.
Pubblicazione: (2025)
di: Wang, Jingbo, et al.
Pubblicazione: (2025)
Sparse but Wrong: Incorrect L0 Leads to Incorrect Features in Sparse Autoencoders
di: Chanin, David, et al.
Pubblicazione: (2025)
di: Chanin, David, et al.
Pubblicazione: (2025)
Can We Verify Step by Step for Incorrect Answer Detection?
di: Xu, Xin, et al.
Pubblicazione: (2024)
di: Xu, Xin, et al.
Pubblicazione: (2024)
Multi-Agent Reasoning with Consistency Verification Improves Uncertainty Calibration in Medical MCQA
di: Martinez, John Ray B.
Pubblicazione: (2026)
di: Martinez, John Ray B.
Pubblicazione: (2026)
Stepwise Guided Policy Optimization: Coloring your Incorrect Reasoning in GRPO
di: Chen, Peter, et al.
Pubblicazione: (2025)
di: Chen, Peter, et al.
Pubblicazione: (2025)
Can LLMs Model Incorrect Student Reasoning? A Case Study on Distractor Generation
di: Zengaffinen, Yanick, et al.
Pubblicazione: (2026)
di: Zengaffinen, Yanick, et al.
Pubblicazione: (2026)
DuFFin: A Dual-Level Fingerprinting Framework for LLMs IP Protection
di: Yan, Yuliang, et al.
Pubblicazione: (2025)
di: Yan, Yuliang, et al.
Pubblicazione: (2025)
Diagnosing and Remedying Knowledge Deficiencies in LLMs via Label-free Curricular Meaningful Learning
di: Xiong, Kai, et al.
Pubblicazione: (2024)
di: Xiong, Kai, et al.
Pubblicazione: (2024)
M-Eval: A Heterogeneity-Based Framework for Multi-evidence Validation in Medical RAG Systems
di: Sun, Mengzhou, et al.
Pubblicazione: (2025)
di: Sun, Mengzhou, et al.
Pubblicazione: (2025)
Self-Evolving GPT: A Lifelong Autonomous Experiential Learner
di: Gao, Jinglong, et al.
Pubblicazione: (2024)
di: Gao, Jinglong, et al.
Pubblicazione: (2024)
Playing Language Game with LLMs Leads to Jailbreaking
di: Peng, Yu, et al.
Pubblicazione: (2024)
di: Peng, Yu, et al.
Pubblicazione: (2024)
Towards Understanding the Influence of Reward Margin on Preference Model Performance
di: Qin, Bowen, et al.
Pubblicazione: (2024)
di: Qin, Bowen, et al.
Pubblicazione: (2024)
Examining Inter-Consistency of Large Language Models Collaboration: An In-depth Analysis via Debate
di: Xiong, Kai, et al.
Pubblicazione: (2023)
di: Xiong, Kai, et al.
Pubblicazione: (2023)
From Latent Signals to Reflection Behavior: Tracing Meta-Cognitive Activation Trajectory in R1-Style LLMs
di: Du, Yanrui, et al.
Pubblicazione: (2026)
di: Du, Yanrui, et al.
Pubblicazione: (2026)
KorMedMCQA-V: A Multimodal Benchmark for Evaluating Vision-Language Models on the Korean Medical Licensing Examination
di: Choi, Byungjin, et al.
Pubblicazione: (2026)
di: Choi, Byungjin, et al.
Pubblicazione: (2026)
Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing
di: Peng, Dan, et al.
Pubblicazione: (2025)
di: Peng, Dan, et al.
Pubblicazione: (2025)
Pruning via Merging: Compressing LLMs via Manifold Alignment Based Layer Merging
di: Liu, Deyuan, et al.
Pubblicazione: (2024)
di: Liu, Deyuan, et al.
Pubblicazione: (2024)
Enhancing Complex Causality Extraction via Improved Subtask Interaction and Knowledge Fusion
di: Gao, Jinglong, et al.
Pubblicazione: (2024)
di: Gao, Jinglong, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Beyond Direct Diagnosis: LLM-based Multi-Specialist Agent Consultation for Automatic Diagnosis
di: Wang, Haochun, et al.
Pubblicazione: (2024) -
Uncovering the Role of Initial Saliency in U-Shaped Attention Bias: Scaling Initial Token Weight for Enhanced Long-Text Processing
di: Qiang, Zewen, et al.
Pubblicazione: (2025) -
Manifold-based Verbalizer Space Re-embedding for Tuning-free Prompt-based Classification
di: Wang, Haochun, et al.
Pubblicazione: (2023) -
Knowledge-tuning Large Language Models with Structured Medical Knowledge Bases for Reliable Response Generation in Chinese
di: Wang, Haochun, et al.
Pubblicazione: (2023) -
AS-ES Learning: Towards Efficient CoT Learning in Small Models
di: Xi, Nuwa, et al.
Pubblicazione: (2024)