LLM Reasoning Engine: Specialized Training for Enhanced Mathematical Reasoning
Fuente:
arXiv
Guardado en:
| Autores principales: | Chen, Shuguang, Lin, Guang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Toward Automated Robustness Evaluation of Mathematical Reasoning
por: Hou, Yutao, et al.
Publicado: (2025)
por: Hou, Yutao, et al.
Publicado: (2025)
FABSVer: Faster Training and Better Self-Verification for LLM Mathematical Reasoning
por: Pan, Haihui, et al.
Publicado: (2026)
por: Pan, Haihui, et al.
Publicado: (2026)
Describe-then-Reason: Improving Multimodal Mathematical Reasoning through Visual Comprehension Training
por: Jia, Mengzhao, et al.
Publicado: (2024)
por: Jia, Mengzhao, et al.
Publicado: (2024)
Self-Enhanced Reasoning Training: Activating Latent Reasoning in Small Models for Enhanced Reasoning Distillation
por: Zhang, Yong, et al.
Publicado: (2025)
por: Zhang, Yong, et al.
Publicado: (2025)
Weaker LLMs' Opinions Also Matter: Mixture of Opinions Enhances LLM's Mathematical Reasoning
por: Chen, Yanan, et al.
Publicado: (2025)
por: Chen, Yanan, et al.
Publicado: (2025)
Enhancing Mathematical Reasoning in LLMs by Stepwise Correction
por: Wu, Zhenyu, et al.
Publicado: (2024)
por: Wu, Zhenyu, et al.
Publicado: (2024)
Curing Miracle Steps in LLM Mathematical Reasoning with Rubric Rewards
por: Yuan, Youliang, et al.
Publicado: (2025)
por: Yuan, Youliang, et al.
Publicado: (2025)
VERGE: Formal Refinement and Guidance Engine for Verifiable LLM Reasoning
por: Singh, Vikash, et al.
Publicado: (2026)
por: Singh, Vikash, et al.
Publicado: (2026)
Learning From Mistakes Makes LLM Better Reasoner
por: An, Shengnan, et al.
Publicado: (2023)
por: An, Shengnan, et al.
Publicado: (2023)
Mathematical Reasoning Enhanced LLM for Formula Derivation: A Case Study on Fiber NLI Modellin
por: Zhang, Yao, et al.
Publicado: (2026)
por: Zhang, Yao, et al.
Publicado: (2026)
Rethinking Expert Trajectory Utilization in LLM Post-training for Mathematical Reasoning
por: Ding, Bowen, et al.
Publicado: (2025)
por: Ding, Bowen, et al.
Publicado: (2025)
Learn Beyond The Answer: Training Language Models with Reflection for Mathematical Reasoning
por: Zhang, Zhihan, et al.
Publicado: (2024)
por: Zhang, Zhihan, et al.
Publicado: (2024)
Quantization Meets Reasoning: Exploring LLM Low-Bit Quantization Degradation for Mathematical Reasoning
por: Li, Zhen, et al.
Publicado: (2025)
por: Li, Zhen, et al.
Publicado: (2025)
Gap-Filling Prompting Enhances Code-Assisted Mathematical Reasoning
por: Mohammadkhani, Mohammad Ghiasvand
Publicado: (2024)
por: Mohammadkhani, Mohammad Ghiasvand
Publicado: (2024)
Does Learning Mathematical Problem-Solving Generalize to Broader Reasoning?
por: Zhou, Ruochen, et al.
Publicado: (2025)
por: Zhou, Ruochen, et al.
Publicado: (2025)
Knowledge Graph-Assisted LLM Post-Training for Enhanced Legal Reasoning
por: Song, Dezhao, et al.
Publicado: (2026)
por: Song, Dezhao, et al.
Publicado: (2026)
Unlocking LLM Safeguards for Low-Resource Languages via Reasoning and Alignment with Minimal Training Data
por: Chen, Zhuowei, et al.
Publicado: (2025)
por: Chen, Zhuowei, et al.
Publicado: (2025)
Reasoning before Comparison: LLM-Enhanced Semantic Similarity Metrics for Domain Specialized Text Analysis
por: Xu, Shaochen, et al.
Publicado: (2024)
por: Xu, Shaochen, et al.
Publicado: (2024)
DeepDistill: Enhancing LLM Reasoning Capabilities via Large-Scale Difficulty-Graded Data Training
por: Tian, Xiaoyu, et al.
Publicado: (2025)
por: Tian, Xiaoyu, et al.
Publicado: (2025)
Step-Controlled DPO: Leveraging Stepwise Error for Enhanced Mathematical Reasoning
por: Lu, Zimu, et al.
Publicado: (2024)
por: Lu, Zimu, et al.
Publicado: (2024)
Math-PUMA: Progressive Upward Multimodal Alignment to Enhance Mathematical Reasoning
por: Zhuang, Wenwen, et al.
Publicado: (2024)
por: Zhuang, Wenwen, et al.
Publicado: (2024)
Beyond Gold Standards: Epistemic Ensemble of LLM Judges for Formal Mathematical Reasoning
por: Zhang, Lan, et al.
Publicado: (2025)
por: Zhang, Lan, et al.
Publicado: (2025)
Beyond Two-Stage Training: Cooperative SFT and RL for LLM Reasoning
por: Chen, Liang, et al.
Publicado: (2025)
por: Chen, Liang, et al.
Publicado: (2025)
Enhancing LLM Reasoning via Non-Human-Like Reasoning Path Preference Optimization
por: Lu, Junjie, et al.
Publicado: (2025)
por: Lu, Junjie, et al.
Publicado: (2025)
Exploring the Mystery of Influential Data for Mathematical Reasoning
por: Ni, Xinzhe, et al.
Publicado: (2024)
por: Ni, Xinzhe, et al.
Publicado: (2024)
Step Guided Reasoning: Improving Mathematical Reasoning using Guidance Generation and Step Reasoning
por: Cao, Lang, et al.
Publicado: (2024)
por: Cao, Lang, et al.
Publicado: (2024)
Unconstrained Model Merging for Enhanced LLM Reasoning
por: Zhang, Yiming, et al.
Publicado: (2024)
por: Zhang, Yiming, et al.
Publicado: (2024)
MathFimer: Enhancing Mathematical Reasoning by Expanding Reasoning Steps through Fill-in-the-Middle Task
por: Yan, Yuchen, et al.
Publicado: (2025)
por: Yan, Yuchen, et al.
Publicado: (2025)
Mitigating Catastrophic Forgetting in Mathematical Reasoning Finetuning through Mixed Training
por: Reynolds, John Graham
Publicado: (2025)
por: Reynolds, John Graham
Publicado: (2025)
Evaluating Mathematical Reasoning Beyond Accuracy
por: Xia, Shijie, et al.
Publicado: (2024)
por: Xia, Shijie, et al.
Publicado: (2024)
DISPO: Enhancing Training Efficiency and Stability in Reinforcement Learning for Large Language Model Mathematical Reasoning
por: Karaman, Batuhan K., et al.
Publicado: (2026)
por: Karaman, Batuhan K., et al.
Publicado: (2026)
From Calculation to Adjudication: Examining LLM judges on Mathematical Reasoning Tasks
por: Stephan, Andreas, et al.
Publicado: (2024)
por: Stephan, Andreas, et al.
Publicado: (2024)
Saturation-Driven Dataset Generation for LLM Mathematical Reasoning in the TPTP Ecosystem
por: Quesnel, Valentin, et al.
Publicado: (2025)
por: Quesnel, Valentin, et al.
Publicado: (2025)
Self-Explore: Enhancing Mathematical Reasoning in Language Models with Fine-grained Rewards
por: Hwang, Hyeonbin, et al.
Publicado: (2024)
por: Hwang, Hyeonbin, et al.
Publicado: (2024)
MultiLingPoT: Enhancing Mathematical Reasoning with Multilingual Program Fine-tuning
por: Li, Nianqi, et al.
Publicado: (2024)
por: Li, Nianqi, et al.
Publicado: (2024)
Can A Gamer Train A Mathematical Reasoning Model?
por: Shin, Andrew
Publicado: (2025)
por: Shin, Andrew
Publicado: (2025)
Olapa-MCoT: Enhancing the Chinese Mathematical Reasoning Capability of LLMs
por: Zhu, Shaojie, et al.
Publicado: (2023)
por: Zhu, Shaojie, et al.
Publicado: (2023)
PolyMath: Evaluating Mathematical Reasoning in Multilingual Contexts
por: Wang, Yiming, et al.
Publicado: (2025)
por: Wang, Yiming, et al.
Publicado: (2025)
Multi-turn Training with Basic Human Feedback Helps Little on LLM Reasoning
por: Liu, Qiang, et al.
Publicado: (2025)
por: Liu, Qiang, et al.
Publicado: (2025)
Towards Robust Mathematical Reasoning
por: Luong, Thang, et al.
Publicado: (2025)
por: Luong, Thang, et al.
Publicado: (2025)
Ejemplares similares
-
Toward Automated Robustness Evaluation of Mathematical Reasoning
por: Hou, Yutao, et al.
Publicado: (2025) -
FABSVer: Faster Training and Better Self-Verification for LLM Mathematical Reasoning
por: Pan, Haihui, et al.
Publicado: (2026) -
Describe-then-Reason: Improving Multimodal Mathematical Reasoning through Visual Comprehension Training
por: Jia, Mengzhao, et al.
Publicado: (2024) -
Self-Enhanced Reasoning Training: Activating Latent Reasoning in Small Models for Enhanced Reasoning Distillation
por: Zhang, Yong, et al.
Publicado: (2025) -
Weaker LLMs' Opinions Also Matter: Mixture of Opinions Enhances LLM's Mathematical Reasoning
por: Chen, Yanan, et al.
Publicado: (2025)