Multi-Agent Reasoning with Consistency Verification Improves Uncertainty Calibration in Medical MCQA
Fuente:
arXiv
Guardado en:
| Autor principal: | Martinez, John Ray B. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Finding the Cracks: Improving LLMs Reasoning with Paraphrastic Probing and Consistency Verification
por: Shi, Weili, et al.
Publicado: (2026)
por: Shi, Weili, et al.
Publicado: (2026)
Think Consistently, Reason Efficiently: Energy-Based Calibration for Implicit Chain-of-Thought
por: Chen, Zhikang, et al.
Publicado: (2025)
por: Chen, Zhikang, et al.
Publicado: (2025)
Self-Training Meets Consistency: Improving LLMs' Reasoning with Consistency-Driven Rationale Evaluation
por: Lee, Jaehyeok, et al.
Publicado: (2024)
por: Lee, Jaehyeok, et al.
Publicado: (2024)
Soft Self-Consistency Improves Language Model Agents
por: Wang, Han, et al.
Publicado: (2024)
por: Wang, Han, et al.
Publicado: (2024)
Introducing Verification Task of Set Consistency with Set-Consistency Energy Networks
por: Song, Mooho, et al.
Publicado: (2025)
por: Song, Mooho, et al.
Publicado: (2025)
TinyV: Reducing False Negatives in Verification Improves RL for LLM Reasoning
por: Xu, Zhangchen, et al.
Publicado: (2025)
por: Xu, Zhangchen, et al.
Publicado: (2025)
Black-Box Reliability Certification for AI Agents via Self-Consistency Sampling and Conformal Calibration
por: Mouzouni, Charafeddine
Publicado: (2026)
por: Mouzouni, Charafeddine
Publicado: (2026)
MAVEN: Multi-Agent Verification-Elaboration Network with In-Step Epistemic Auditing
por: Yao, Yinsheng, et al.
Publicado: (2026)
por: Yao, Yinsheng, et al.
Publicado: (2026)
Improving LLM Code Reasoning via Semantic Equivalence Self-Play with Formal Verification
por: Barone, Antonio Valerio Miceli, et al.
Publicado: (2026)
por: Barone, Antonio Valerio Miceli, et al.
Publicado: (2026)
Improving Quantized Model Performance in Qualitative Analysis with Multi-Pass Prompt Verification
por: Adeseye, Aisvarya, et al.
Publicado: (2026)
por: Adeseye, Aisvarya, et al.
Publicado: (2026)
Revisiting Uncertainty Estimation and Calibration of Large Language Models
por: Tao, Linwei, et al.
Publicado: (2025)
por: Tao, Linwei, et al.
Publicado: (2025)
Uncertainty in Language Models: Assessment through Rank-Calibration
por: Huang, Xinmeng, et al.
Publicado: (2024)
por: Huang, Xinmeng, et al.
Publicado: (2024)
On Subjective Uncertainty Quantification and Calibration in Natural Language Generation
por: Wang, Ziyu, et al.
Publicado: (2024)
por: Wang, Ziyu, et al.
Publicado: (2024)
MMedAgent-RL: Optimizing Multi-Agent Collaboration for Multimodal Medical Reasoning
por: Xia, Peng, et al.
Publicado: (2025)
por: Xia, Peng, et al.
Publicado: (2025)
The Consistency Hypothesis in Uncertainty Quantification for Large Language Models
por: Xiao, Quan, et al.
Publicado: (2025)
por: Xiao, Quan, et al.
Publicado: (2025)
Consistency Calibration: Improving Uncertainty Calibration via Consistency among Perturbed Neighbors
por: Tao, Linwei, et al.
Publicado: (2024)
por: Tao, Linwei, et al.
Publicado: (2024)
Verification-Aware Planning for Multi-Agent Systems
por: Xu, Tianyang, et al.
Publicado: (2025)
por: Xu, Tianyang, et al.
Publicado: (2025)
BayesAgent: Bayesian Agentic Reasoning Under Uncertainty via Verbalized Probabilistic Graphical Modeling
por: Huang, Hengguan, et al.
Publicado: (2024)
por: Huang, Hengguan, et al.
Publicado: (2024)
Reinforce LLM Reasoning through Multi-Agent Reflection
por: Yuan, Yurun, et al.
Publicado: (2025)
por: Yuan, Yurun, et al.
Publicado: (2025)
An Assessment of Human vs. Model Uncertainty in Soft-Label Learning and Calibration
por: Pavlovic, Maja, et al.
Publicado: (2026)
por: Pavlovic, Maja, et al.
Publicado: (2026)
Learning from Synthetic Data Improves Multi-hop Reasoning
por: Kabra, Anmol, et al.
Publicado: (2026)
por: Kabra, Anmol, et al.
Publicado: (2026)
Tracing Uncertainty in Language Model "Reasoning"
por: Grünefeld, Nils, et al.
Publicado: (2026)
por: Grünefeld, Nils, et al.
Publicado: (2026)
Forward-Backward Reasoning in Large Language Models for Mathematical Verification
por: Jiang, Weisen, et al.
Publicado: (2023)
por: Jiang, Weisen, et al.
Publicado: (2023)
Aligning LLMs with Human Uncertainty: A Beta-Bernoulli Calibrator for LLM Forecasting
por: Dai, Hui, et al.
Publicado: (2026)
por: Dai, Hui, et al.
Publicado: (2026)
ConsistRM: Improving Generative Reward Models via Consistency-Aware Self-Training
por: Liang, Yu, et al.
Publicado: (2026)
por: Liang, Yu, et al.
Publicado: (2026)
Enhancing Trust in Large Language Models via Uncertainty-Calibrated Fine-Tuning
por: Krishnan, Ranganath, et al.
Publicado: (2024)
por: Krishnan, Ranganath, et al.
Publicado: (2024)
ARM: Discovering Agentic Reasoning Modules for Generalizable Multi-Agent Systems
por: Yao, Bohan, et al.
Publicado: (2025)
por: Yao, Bohan, et al.
Publicado: (2025)
Temporal Consistency for LLM Reasoning Process Error Identification
por: Guo, Jiacheng, et al.
Publicado: (2025)
por: Guo, Jiacheng, et al.
Publicado: (2025)
Beyond Isolation: Multi-Agent Synergy for Improving Knowledge Graph Construction
por: Ye, Hongbin, et al.
Publicado: (2023)
por: Ye, Hongbin, et al.
Publicado: (2023)
Less is More for Improving Automatic Evaluation of Factual Consistency
por: Wang, Tong, et al.
Publicado: (2024)
por: Wang, Tong, et al.
Publicado: (2024)
Forging the Forger: An Attempt to Improve Authorship Verification via Data Augmentation
por: Corbara, Silvia, et al.
Publicado: (2024)
por: Corbara, Silvia, et al.
Publicado: (2024)
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization
por: Kawakami, Wataru, et al.
Publicado: (2025)
por: Kawakami, Wataru, et al.
Publicado: (2025)
Husky: A Unified, Open-Source Language Agent for Multi-Step Reasoning
por: Kim, Joongwon, et al.
Publicado: (2024)
por: Kim, Joongwon, et al.
Publicado: (2024)
Diversity of Thought Elicits Stronger Reasoning Capabilities in Multi-Agent Debate Frameworks
por: Hegazy, Mahmood
Publicado: (2024)
por: Hegazy, Mahmood
Publicado: (2024)
CAMPHOR: Collaborative Agents for Multi-input Planning and High-Order Reasoning On Device
por: Fu, Yicheng, et al.
Publicado: (2024)
por: Fu, Yicheng, et al.
Publicado: (2024)
Stepwise Self-Consistent Mathematical Reasoning with Large Language Models
por: Zhao, Zilong, et al.
Publicado: (2024)
por: Zhao, Zilong, et al.
Publicado: (2024)
Every Response Counts: Quantifying Uncertainty of LLM-based Multi-Agent Systems through Tensor Decomposition
por: Chen, Tiejin, et al.
Publicado: (2026)
por: Chen, Tiejin, et al.
Publicado: (2026)
Self-Verification Dilemma: Experience-Driven Suppression of Overused Checking in LLM Reasoning
por: Long, Quanyu, et al.
Publicado: (2026)
por: Long, Quanyu, et al.
Publicado: (2026)
How Uncertainty Estimation Scales with Sampling in Reasoning Models
por: Del, Maksym, et al.
Publicado: (2026)
por: Del, Maksym, et al.
Publicado: (2026)
Stepwise Verification and Remediation of Student Reasoning Errors with Large Language Model Tutors
por: Daheim, Nico, et al.
Publicado: (2024)
por: Daheim, Nico, et al.
Publicado: (2024)
Ejemplares similares
-
Finding the Cracks: Improving LLMs Reasoning with Paraphrastic Probing and Consistency Verification
por: Shi, Weili, et al.
Publicado: (2026) -
Think Consistently, Reason Efficiently: Energy-Based Calibration for Implicit Chain-of-Thought
por: Chen, Zhikang, et al.
Publicado: (2025) -
Self-Training Meets Consistency: Improving LLMs' Reasoning with Consistency-Driven Rationale Evaluation
por: Lee, Jaehyeok, et al.
Publicado: (2024) -
Soft Self-Consistency Improves Language Model Agents
por: Wang, Han, et al.
Publicado: (2024) -
Introducing Verification Task of Set Consistency with Set-Consistency Energy Networks
por: Song, Mooho, et al.
Publicado: (2025)