MathCoder2: Better Math Reasoning from Continued Pretraining on Model-translated Mathematical Code
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lu, Zimu, Zhou, Aojun, Wang, Ke, Ren, Houxing, Shi, Weikang, Pan, Junting, Zhan, Mingjie, Li, Hongsheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MathCoder-VL: Bridging Vision and Code for Enhanced Multimodal Mathematical Reasoning
von: Wang, Ke, et al.
Veröffentlicht: (2025)
von: Wang, Ke, et al.
Veröffentlicht: (2025)
MathGenie: Generating Synthetic Data with Question Back-translation for Enhancing Mathematical Reasoning of LLMs
von: Lu, Zimu, et al.
Veröffentlicht: (2024)
von: Lu, Zimu, et al.
Veröffentlicht: (2024)
Step-Controlled DPO: Leveraging Stepwise Error for Enhanced Mathematical Reasoning
von: Lu, Zimu, et al.
Veröffentlicht: (2024)
von: Lu, Zimu, et al.
Veröffentlicht: (2024)
ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation
von: Ren, Houxing, et al.
Veröffentlicht: (2024)
von: Ren, Houxing, et al.
Veröffentlicht: (2024)
Probability-Consistent Preference Optimization for Enhanced LLM Reasoning
von: Yang, Yunqiao, et al.
Veröffentlicht: (2025)
von: Yang, Yunqiao, et al.
Veröffentlicht: (2025)
Alignment with Fill-In-the-Middle for Enhancing Code Generation
von: Ren, Houxing, et al.
Veröffentlicht: (2025)
von: Ren, Houxing, et al.
Veröffentlicht: (2025)
MathCanvas: Intrinsic Visual Chain-of-Thought for Multimodal Mathematical Reasoning
von: Shi, Weikang, et al.
Veröffentlicht: (2025)
von: Shi, Weikang, et al.
Veröffentlicht: (2025)
Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset
von: Wang, Ke, et al.
Veröffentlicht: (2024)
von: Wang, Ke, et al.
Veröffentlicht: (2024)
From Solver to Tutor: Evaluating the Pedagogical Intelligence of LLMs with KMP-Bench
von: Shi, Weikang, et al.
Veröffentlicht: (2026)
von: Shi, Weikang, et al.
Veröffentlicht: (2026)
VoiceAssistant-Eval: Benchmarking AI Assistants across Listening, Speaking, and Viewing
von: Wang, Ke, et al.
Veröffentlicht: (2025)
von: Wang, Ke, et al.
Veröffentlicht: (2025)
WebGen-Bench: Evaluating LLMs on Generating Interactive and Functional Websites from Scratch
von: Lu, Zimu, et al.
Veröffentlicht: (2025)
von: Lu, Zimu, et al.
Veröffentlicht: (2025)
Edit-Based Refinement for Parallel Masked Diffusion Language Models
von: Ren, Houxing, et al.
Veröffentlicht: (2026)
von: Ren, Houxing, et al.
Veröffentlicht: (2026)
WebGen-Agent: Enhancing Interactive Website Generation with Multi-Level Feedback and Step-Level Reinforcement Learning
von: Lu, Zimu, et al.
Veröffentlicht: (2025)
von: Lu, Zimu, et al.
Veröffentlicht: (2025)
FullStack-Agent: Enhancing Agentic Full-Stack Web Coding via Development-Oriented Testing and Repository Back-Translation
von: Lu, Zimu, et al.
Veröffentlicht: (2026)
von: Lu, Zimu, et al.
Veröffentlicht: (2026)
Towards Robust Real-World Spreadsheet Understanding with Multi-Agent Multi-Format Reasoning
von: Ren, Houxing, et al.
Veröffentlicht: (2026)
von: Ren, Houxing, et al.
Veröffentlicht: (2026)
Empowering Character-level Text Infilling by Eliminating Sub-Tokens
von: Ren, Houxing, et al.
Veröffentlicht: (2024)
von: Ren, Houxing, et al.
Veröffentlicht: (2024)
MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?
von: Zhang, Renrui, et al.
Veröffentlicht: (2024)
von: Zhang, Renrui, et al.
Veröffentlicht: (2024)
SlidesGen-Bench: Evaluating Slides Generation via Computational and Quantitative Metrics
von: Yang, Yunqiao, et al.
Veröffentlicht: (2026)
von: Yang, Yunqiao, et al.
Veröffentlicht: (2026)
CMM-Math: A Chinese Multimodal Math Dataset To Evaluate and Enhance the Mathematics Reasoning of Large Multimodal Models
von: Liu, Wentao, et al.
Veröffentlicht: (2024)
von: Liu, Wentao, et al.
Veröffentlicht: (2024)
VisAidMath: Benchmarking Visual-Aided Mathematical Reasoning
von: Ma, Jingkun, et al.
Veröffentlicht: (2024)
von: Ma, Jingkun, et al.
Veröffentlicht: (2024)
GeoMathCode: Understanding Interleaved Math-Code Reasoning for Geometry Problem Solving
von: Zhang, Yingji, et al.
Veröffentlicht: (2026)
von: Zhang, Yingji, et al.
Veröffentlicht: (2026)
MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts
von: Lu, Pan, et al.
Veröffentlicht: (2023)
von: Lu, Pan, et al.
Veröffentlicht: (2023)
PolyMath: Evaluating Mathematical Reasoning in Multilingual Contexts
von: Wang, Yiming, et al.
Veröffentlicht: (2025)
von: Wang, Yiming, et al.
Veröffentlicht: (2025)
MathSmith: Towards Extremely Hard Mathematical Reasoning by Forging Synthetic Problems with a Reinforced Policy
von: Zhan, Shaoxiong, et al.
Veröffentlicht: (2025)
von: Zhan, Shaoxiong, et al.
Veröffentlicht: (2025)
DeepSeekMath-V2: Towards Self-Verifiable Mathematical Reasoning
von: Shao, Zhihong, et al.
Veröffentlicht: (2025)
von: Shao, Zhihong, et al.
Veröffentlicht: (2025)
Code2Math: Can Your Code Agent Effectively Evolve Math Problems Through Exploration?
von: Guo, Dadi, et al.
Veröffentlicht: (2026)
von: Guo, Dadi, et al.
Veröffentlicht: (2026)
DotaMath: Decomposition of Thought with Code Assistance and Self-correction for Mathematical Reasoning
von: Li, Chengpeng, et al.
Veröffentlicht: (2024)
von: Li, Chengpeng, et al.
Veröffentlicht: (2024)
RoMath: A Mathematical Reasoning Benchmark in Romanian
von: Cosma, Adrian, et al.
Veröffentlicht: (2024)
von: Cosma, Adrian, et al.
Veröffentlicht: (2024)
MathPile: A Billion-Token-Scale Pretraining Corpus for Math
von: Wang, Zengzhi, et al.
Veröffentlicht: (2023)
von: Wang, Zengzhi, et al.
Veröffentlicht: (2023)
MuggleMath: Assessing the Impact of Query and Response Augmentation on Math Reasoning
von: Li, Chengpeng, et al.
Veröffentlicht: (2023)
von: Li, Chengpeng, et al.
Veröffentlicht: (2023)
TabularMath: Understanding Math Reasoning over Tables with Large Language Models
von: Tian, Shi-Yu, et al.
Veröffentlicht: (2025)
von: Tian, Shi-Yu, et al.
Veröffentlicht: (2025)
CoinMath: Harnessing the Power of Coding Instruction for Math LLMs
von: Wei, Chengwei, et al.
Veröffentlicht: (2024)
von: Wei, Chengwei, et al.
Veröffentlicht: (2024)
We-Math 2.0: A Versatile MathBook System for Incentivizing Visual Mathematical Reasoning
von: Qiao, Runqi, et al.
Veröffentlicht: (2025)
von: Qiao, Runqi, et al.
Veröffentlicht: (2025)
Is Your Model Really A Good Math Reasoner? Evaluating Mathematical Reasoning with Checklist
von: Zhou, Zihao, et al.
Veröffentlicht: (2024)
von: Zhou, Zihao, et al.
Veröffentlicht: (2024)
Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models
von: Shi, Wenhao, et al.
Veröffentlicht: (2024)
von: Shi, Wenhao, et al.
Veröffentlicht: (2024)
FinanceMath: Knowledge-Intensive Math Reasoning in Finance Domains
von: Zhao, Yilun, et al.
Veröffentlicht: (2023)
von: Zhao, Yilun, et al.
Veröffentlicht: (2023)
MathScale: Scaling Instruction Tuning for Mathematical Reasoning
von: Tang, Zhengyang, et al.
Veröffentlicht: (2024)
von: Tang, Zhengyang, et al.
Veröffentlicht: (2024)
ReliableMath: Benchmark of Reliable Mathematical Reasoning on Large Language Models
von: Xue, Boyang, et al.
Veröffentlicht: (2025)
von: Xue, Boyang, et al.
Veröffentlicht: (2025)
MathHay: An Automated Benchmark for Long-Context Mathematical Reasoning in LLMs
von: Wang, Lei, et al.
Veröffentlicht: (2024)
von: Wang, Lei, et al.
Veröffentlicht: (2024)
Math-PUMA: Progressive Upward Multimodal Alignment to Enhance Mathematical Reasoning
von: Zhuang, Wenwen, et al.
Veröffentlicht: (2024)
von: Zhuang, Wenwen, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MathCoder-VL: Bridging Vision and Code for Enhanced Multimodal Mathematical Reasoning
von: Wang, Ke, et al.
Veröffentlicht: (2025) -
MathGenie: Generating Synthetic Data with Question Back-translation for Enhancing Mathematical Reasoning of LLMs
von: Lu, Zimu, et al.
Veröffentlicht: (2024) -
Step-Controlled DPO: Leveraging Stepwise Error for Enhanced Mathematical Reasoning
von: Lu, Zimu, et al.
Veröffentlicht: (2024) -
ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation
von: Ren, Houxing, et al.
Veröffentlicht: (2024) -
Probability-Consistent Preference Optimization for Enhanced LLM Reasoning
von: Yang, Yunqiao, et al.
Veröffentlicht: (2025)