Finding the Cracks: Improving LLMs Reasoning with Paraphrastic Probing and Consistency Verification
Fuente:
arXiv
Guardado en:
| Autores principales: | Shi, Weili, Guo, Dongliang, Yang, Lehan, Wang, Tianlong, Yuan, Hanzhang, Li, Sheng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multi-Agent Reasoning with Consistency Verification Improves Uncertainty Calibration in Medical MCQA
por: Martinez, John Ray B.
Publicado: (2026)
por: Martinez, John Ray B.
Publicado: (2026)
Self-Training Meets Consistency: Improving LLMs' Reasoning with Consistency-Driven Rationale Evaluation
por: Lee, Jaehyeok, et al.
Publicado: (2024)
por: Lee, Jaehyeok, et al.
Publicado: (2024)
Probing to Refine: Reinforcement Distillation of LLMs via Explanatory Inversion
por: Tan, Zhen, et al.
Publicado: (2026)
por: Tan, Zhen, et al.
Publicado: (2026)
Bridging Internal Probability and Self-Consistency for Effective and Efficient LLM Reasoning
por: Zhou, Zhi, et al.
Publicado: (2025)
por: Zhou, Zhi, et al.
Publicado: (2025)
TinyV: Reducing False Negatives in Verification Improves RL for LLM Reasoning
por: Xu, Zhangchen, et al.
Publicado: (2025)
por: Xu, Zhangchen, et al.
Publicado: (2025)
PiCO: Peer Review in LLMs based on the Consistency Optimization
por: Ning, Kun-Peng, et al.
Publicado: (2024)
por: Ning, Kun-Peng, et al.
Publicado: (2024)
Forward-Backward Reasoning in Large Language Models for Mathematical Verification
por: Jiang, Weisen, et al.
Publicado: (2023)
por: Jiang, Weisen, et al.
Publicado: (2023)
Introducing Verification Task of Set Consistency with Set-Consistency Energy Networks
por: Song, Mooho, et al.
Publicado: (2025)
por: Song, Mooho, et al.
Publicado: (2025)
Prompt Repetition Improves Non-Reasoning LLMs
por: Leviathan, Yaniv, et al.
Publicado: (2025)
por: Leviathan, Yaniv, et al.
Publicado: (2025)
Temporal Consistency for LLM Reasoning Process Error Identification
por: Guo, Jiacheng, et al.
Publicado: (2025)
por: Guo, Jiacheng, et al.
Publicado: (2025)
Can GRPO Help LLMs Transcend Their Pretraining Origin?
por: Ni, Kangqi, et al.
Publicado: (2025)
por: Ni, Kangqi, et al.
Publicado: (2025)
GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
por: Duan, Jinhao, et al.
Publicado: (2024)
por: Duan, Jinhao, et al.
Publicado: (2024)
Probe Pruning: Accelerating LLMs through Dynamic Pruning via Model-Probing
por: Le, Qi, et al.
Publicado: (2025)
por: Le, Qi, et al.
Publicado: (2025)
When Models Know More Than They Say: Probing Analogical Reasoning in LLMs
por: McGovern, Hope, et al.
Publicado: (2026)
por: McGovern, Hope, et al.
Publicado: (2026)
Improving LLM Code Reasoning via Semantic Equivalence Self-Play with Formal Verification
por: Barone, Antonio Valerio Miceli, et al.
Publicado: (2026)
por: Barone, Antonio Valerio Miceli, et al.
Publicado: (2026)
ConsistRM: Improving Generative Reward Models via Consistency-Aware Self-Training
por: Liang, Yu, et al.
Publicado: (2026)
por: Liang, Yu, et al.
Publicado: (2026)
Stepwise Self-Consistent Mathematical Reasoning with Large Language Models
por: Zhao, Zilong, et al.
Publicado: (2024)
por: Zhao, Zilong, et al.
Publicado: (2024)
Distillation Contrastive Decoding: Improving LLMs Reasoning with Contrastive Decoding and Distillation
por: Phan, Phuc, et al.
Publicado: (2024)
por: Phan, Phuc, et al.
Publicado: (2024)
Self-Verification Dilemma: Experience-Driven Suppression of Overused Checking in LLM Reasoning
por: Long, Quanyu, et al.
Publicado: (2026)
por: Long, Quanyu, et al.
Publicado: (2026)
MATH-Perturb: Benchmarking LLMs' Math Reasoning Abilities against Hard Perturbations
por: Huang, Kaixuan, et al.
Publicado: (2025)
por: Huang, Kaixuan, et al.
Publicado: (2025)
MolReasoner: Toward Effective and Interpretable Reasoning for Molecular LLMs
por: Zhao, Guojiang, et al.
Publicado: (2025)
por: Zhao, Guojiang, et al.
Publicado: (2025)
DLO: Dynamic Layer Operation for Efficient Vertical Scaling of LLMs
por: Tan, Zhen, et al.
Publicado: (2024)
por: Tan, Zhen, et al.
Publicado: (2024)
Probing Scientific General Intelligence of LLMs with Scientist-Aligned Workflows
por: Xu, Wanghan, et al.
Publicado: (2025)
por: Xu, Wanghan, et al.
Publicado: (2025)
UProp: Investigating the Uncertainty Propagation of LLMs in Multi-Step Agentic Decision-Making
por: Duan, Jinhao, et al.
Publicado: (2025)
por: Duan, Jinhao, et al.
Publicado: (2025)
Fantastic Reasoning Behaviors and Where to Find Them: Unsupervised Discovery of the Reasoning Process
por: Zhang, Zhenyu, et al.
Publicado: (2025)
por: Zhang, Zhenyu, et al.
Publicado: (2025)
Quokka: Accelerating Program Verification with LLMs via Invariant Synthesis
por: Wei, Anjiang, et al.
Publicado: (2025)
por: Wei, Anjiang, et al.
Publicado: (2025)
Rewarding Graph Reasoning Process makes LLMs more Generalized Reasoners
por: Peng, Miao, et al.
Publicado: (2025)
por: Peng, Miao, et al.
Publicado: (2025)
Learning to Correct for QA Reasoning with Black-box LLMs
por: Kim, Jaehyung, et al.
Publicado: (2024)
por: Kim, Jaehyung, et al.
Publicado: (2024)
ReConcile: Round-Table Conference Improves Reasoning via Consensus among Diverse LLMs
por: Chen, Justin Chih-Yao, et al.
Publicado: (2023)
por: Chen, Justin Chih-Yao, et al.
Publicado: (2023)
Hidden in the Haystack: Smaller Needles are More Difficult for LLMs to Find
por: Bianchi, Owen, et al.
Publicado: (2025)
por: Bianchi, Owen, et al.
Publicado: (2025)
KS-Lottery: Finding Certified Lottery Tickets for Multilingual Language Models
por: Yuan, Fei, et al.
Publicado: (2024)
por: Yuan, Fei, et al.
Publicado: (2024)
Leveraging Parameter Space Symmetries for Reasoning Skill Transfer in LLMs
por: Horoi, Stefan, et al.
Publicado: (2025)
por: Horoi, Stefan, et al.
Publicado: (2025)
Safe: Enhancing Mathematical Reasoning in Large Language Models via Retrospective Step-aware Formal Verification
por: Liu, Chengwu, et al.
Publicado: (2025)
por: Liu, Chengwu, et al.
Publicado: (2025)
Skill-Based Mixture-of-Experts: Adaptive Routing for Heterogeneous Reasoning via Inferred Skills
por: Chen, Justin Chih-Yao, et al.
Publicado: (2025)
por: Chen, Justin Chih-Yao, et al.
Publicado: (2025)
Benchmarking the Capabilities of Large Language Models in Transportation System Engineering: Accuracy, Consistency, and Reasoning Behaviors
por: Syed, Usman, et al.
Publicado: (2024)
por: Syed, Usman, et al.
Publicado: (2024)
How Numerical Precision Affects Arithmetical Reasoning Capabilities of LLMs
por: Feng, Guhao, et al.
Publicado: (2024)
por: Feng, Guhao, et al.
Publicado: (2024)
Memorization vs. Reasoning: Updating LLMs with New Knowledge
por: Li, Aochong Oliver, et al.
Publicado: (2025)
por: Li, Aochong Oliver, et al.
Publicado: (2025)
CriticBench: Benchmarking LLMs for Critique-Correct Reasoning
por: Lin, Zicheng, et al.
Publicado: (2024)
por: Lin, Zicheng, et al.
Publicado: (2024)
Continuous Approximations for Improving Quantization Aware Training of LLMs
por: Li, He, et al.
Publicado: (2024)
por: Li, He, et al.
Publicado: (2024)
Traversal Verification for Speculative Tree Decoding
por: Weng, Yepeng, et al.
Publicado: (2025)
por: Weng, Yepeng, et al.
Publicado: (2025)
Ejemplares similares
-
Multi-Agent Reasoning with Consistency Verification Improves Uncertainty Calibration in Medical MCQA
por: Martinez, John Ray B.
Publicado: (2026) -
Self-Training Meets Consistency: Improving LLMs' Reasoning with Consistency-Driven Rationale Evaluation
por: Lee, Jaehyeok, et al.
Publicado: (2024) -
Probing to Refine: Reinforcement Distillation of LLMs via Explanatory Inversion
por: Tan, Zhen, et al.
Publicado: (2026) -
Bridging Internal Probability and Self-Consistency for Effective and Efficient LLM Reasoning
por: Zhou, Zhi, et al.
Publicado: (2025) -
TinyV: Reducing False Negatives in Verification Improves RL for LLM Reasoning
por: Xu, Zhangchen, et al.
Publicado: (2025)