MathSmith: Towards Extremely Hard Mathematical Reasoning by Forging Synthetic Problems with a Reinforced Policy
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhan, Shaoxiong, Lai, Yanlin, Lu, Ziyu, Lin, Dahua, Yang, Ziqing, Tan, Fei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
More Data or Better Data? A Critical Analysis of Data Selection and Synthesis for Mathematical Reasoning
von: Zhao, Yike, et al.
Veröffentlicht: (2025)
von: Zhao, Yike, et al.
Veröffentlicht: (2025)
MathGenie: Generating Synthetic Data with Question Back-translation for Enhancing Mathematical Reasoning of LLMs
von: Lu, Zimu, et al.
Veröffentlicht: (2024)
von: Lu, Zimu, et al.
Veröffentlicht: (2024)
Verifier-Backed Hard Problem Generation for Mathematical Reasoning
von: Lai, Yuhang, et al.
Veröffentlicht: (2026)
von: Lai, Yuhang, et al.
Veröffentlicht: (2026)
MathOdyssey: Benchmarking Mathematical Problem-Solving Skills in Large Language Models Using Odyssey Math Data
von: Fang, Meng, et al.
Veröffentlicht: (2024)
von: Fang, Meng, et al.
Veröffentlicht: (2024)
PolyMath: Evaluating Mathematical Reasoning in Multilingual Contexts
von: Wang, Yiming, et al.
Veröffentlicht: (2025)
von: Wang, Yiming, et al.
Veröffentlicht: (2025)
InternLM-Math: Open Math Large Language Models Toward Verifiable Reasoning
von: Ying, Huaiyuan, et al.
Veröffentlicht: (2024)
von: Ying, Huaiyuan, et al.
Veröffentlicht: (2024)
Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems
von: Wei, Chengwei, et al.
Veröffentlicht: (2025)
von: Wei, Chengwei, et al.
Veröffentlicht: (2025)
MathBench: Evaluating the Theory and Application Proficiency of LLMs with a Hierarchical Mathematics Benchmark
von: Liu, Hongwei, et al.
Veröffentlicht: (2024)
von: Liu, Hongwei, et al.
Veröffentlicht: (2024)
DeepSeekMath-V2: Towards Self-Verifiable Mathematical Reasoning
von: Shao, Zhihong, et al.
Veröffentlicht: (2025)
von: Shao, Zhihong, et al.
Veröffentlicht: (2025)
MathCoder2: Better Math Reasoning from Continued Pretraining on Model-translated Mathematical Code
von: Lu, Zimu, et al.
Veröffentlicht: (2024)
von: Lu, Zimu, et al.
Veröffentlicht: (2024)
3ViewSense: Spatial and Mental Perspective Reasoning from Orthographic Views in Vision-Language Models
von: Zhan, Shaoxiong, et al.
Veröffentlicht: (2026)
von: Zhan, Shaoxiong, et al.
Veröffentlicht: (2026)
MathMist: A Parallel Multilingual Benchmark Dataset for Mathematical Problem Solving and Reasoning
von: Sobhani, Mahbub E, et al.
Veröffentlicht: (2025)
von: Sobhani, Mahbub E, et al.
Veröffentlicht: (2025)
MathClean: A Benchmark for Synthetic Mathematical Data Cleaning
von: Liang, Hao, et al.
Veröffentlicht: (2025)
von: Liang, Hao, et al.
Veröffentlicht: (2025)
ReliableMath: Benchmark of Reliable Mathematical Reasoning on Large Language Models
von: Xue, Boyang, et al.
Veröffentlicht: (2025)
von: Xue, Boyang, et al.
Veröffentlicht: (2025)
WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct
von: Luo, Haipeng, et al.
Veröffentlicht: (2023)
von: Luo, Haipeng, et al.
Veröffentlicht: (2023)
VisAidMath: Benchmarking Visual-Aided Mathematical Reasoning
von: Ma, Jingkun, et al.
Veröffentlicht: (2024)
von: Ma, Jingkun, et al.
Veröffentlicht: (2024)
MLIR-Forge: A Modular Framework for Language Smiths
von: Ates, Berke, et al.
Veröffentlicht: (2026)
von: Ates, Berke, et al.
Veröffentlicht: (2026)
MathCoder-VL: Bridging Vision and Code for Enhanced Multimodal Mathematical Reasoning
von: Wang, Ke, et al.
Veröffentlicht: (2025)
von: Wang, Ke, et al.
Veröffentlicht: (2025)
RoMath: A Mathematical Reasoning Benchmark in Romanian
von: Cosma, Adrian, et al.
Veröffentlicht: (2024)
von: Cosma, Adrian, et al.
Veröffentlicht: (2024)
VerityMath: Advancing Mathematical Reasoning by Self-Verification Through Unit Consistency
von: Han, Vernon Toh Yan, et al.
Veröffentlicht: (2023)
von: Han, Vernon Toh Yan, et al.
Veröffentlicht: (2023)
From Large to Tiny: Distilling and Refining Mathematical Expertise for Math Word Problems with Weakly Supervision
von: Lin, Qingwen, et al.
Veröffentlicht: (2024)
von: Lin, Qingwen, et al.
Veröffentlicht: (2024)
ReverseMath: Answer Inversion for Scalable and Verifiable Mathematical Problem Generation
von: Zhao, Raoyuan, et al.
Veröffentlicht: (2026)
von: Zhao, Raoyuan, et al.
Veröffentlicht: (2026)
TAPO: Translation Augmented Policy Optimization for Multilingual Mathematical Reasoning
von: Huang, Xu, et al.
Veröffentlicht: (2026)
von: Huang, Xu, et al.
Veröffentlicht: (2026)
MathFusion: Enhancing Mathematical Problem-solving of LLM through Instruction Fusion
von: Pei, Qizhi, et al.
Veröffentlicht: (2025)
von: Pei, Qizhi, et al.
Veröffentlicht: (2025)
Consultant Decoding: Yet Another Synergistic Mechanism
von: Ding, Chuanghao, et al.
Veröffentlicht: (2025)
von: Ding, Chuanghao, et al.
Veröffentlicht: (2025)
MathVC: An LLM-Simulated Multi-Character Virtual Classroom for Mathematics Education
von: Yue, Murong, et al.
Veröffentlicht: (2024)
von: Yue, Murong, et al.
Veröffentlicht: (2024)
GeoMathCode: Understanding Interleaved Math-Code Reasoning for Geometry Problem Solving
von: Zhang, Yingji, et al.
Veröffentlicht: (2026)
von: Zhang, Yingji, et al.
Veröffentlicht: (2026)
MathScale: Scaling Instruction Tuning for Mathematical Reasoning
von: Tang, Zhengyang, et al.
Veröffentlicht: (2024)
von: Tang, Zhengyang, et al.
Veröffentlicht: (2024)
Is Your Model Really A Good Math Reasoner? Evaluating Mathematical Reasoning with Checklist
von: Zhou, Zihao, et al.
Veröffentlicht: (2024)
von: Zhou, Zihao, et al.
Veröffentlicht: (2024)
MathHay: An Automated Benchmark for Long-Context Mathematical Reasoning in LLMs
von: Wang, Lei, et al.
Veröffentlicht: (2024)
von: Wang, Lei, et al.
Veröffentlicht: (2024)
Math-PUMA: Progressive Upward Multimodal Alignment to Enhance Mathematical Reasoning
von: Zhuang, Wenwen, et al.
Veröffentlicht: (2024)
von: Zhuang, Wenwen, et al.
Veröffentlicht: (2024)
CMM-Math: A Chinese Multimodal Math Dataset To Evaluate and Enhance the Mathematics Reasoning of Large Multimodal Models
von: Liu, Wentao, et al.
Veröffentlicht: (2024)
von: Liu, Wentao, et al.
Veröffentlicht: (2024)
Self-consistent Reasoning For Solving Math Word Problems
von: Xiong, Jing, et al.
Veröffentlicht: (2022)
von: Xiong, Jing, et al.
Veröffentlicht: (2022)
MathCanvas: Intrinsic Visual Chain-of-Thought for Multimodal Mathematical Reasoning
von: Shi, Weikang, et al.
Veröffentlicht: (2025)
von: Shi, Weikang, et al.
Veröffentlicht: (2025)
Future Policy Approximation for Offline Reinforcement Learning Improves Mathematical Reasoning
von: Oh, Minjae, et al.
Veröffentlicht: (2025)
von: Oh, Minjae, et al.
Veröffentlicht: (2025)
MathEDU: Feedback Generation on Problem-Solving Processes for Mathematical Learning Support
von: Hsu, Wei-Ling, et al.
Veröffentlicht: (2025)
von: Hsu, Wei-Ling, et al.
Veröffentlicht: (2025)
WarriorMath: Enhancing the Mathematical Ability of Large Language Models with a Defect-aware Framework
von: Chen, Yue, et al.
Veröffentlicht: (2025)
von: Chen, Yue, et al.
Veröffentlicht: (2025)
DART-Math: Difficulty-Aware Rejection Tuning for Mathematical Problem-Solving
von: Tong, Yuxuan, et al.
Veröffentlicht: (2024)
von: Tong, Yuxuan, et al.
Veröffentlicht: (2024)
Solving Math Word Problems via Cooperative Reasoning induced Language Models
von: Zhu, Xinyu, et al.
Veröffentlicht: (2022)
von: Zhu, Xinyu, et al.
Veröffentlicht: (2022)
MuggleMath: Assessing the Impact of Query and Response Augmentation on Math Reasoning
von: Li, Chengpeng, et al.
Veröffentlicht: (2023)
von: Li, Chengpeng, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
More Data or Better Data? A Critical Analysis of Data Selection and Synthesis for Mathematical Reasoning
von: Zhao, Yike, et al.
Veröffentlicht: (2025) -
MathGenie: Generating Synthetic Data with Question Back-translation for Enhancing Mathematical Reasoning of LLMs
von: Lu, Zimu, et al.
Veröffentlicht: (2024) -
Verifier-Backed Hard Problem Generation for Mathematical Reasoning
von: Lai, Yuhang, et al.
Veröffentlicht: (2026) -
MathOdyssey: Benchmarking Mathematical Problem-Solving Skills in Large Language Models Using Odyssey Math Data
von: Fang, Meng, et al.
Veröffentlicht: (2024) -
PolyMath: Evaluating Mathematical Reasoning in Multilingual Contexts
von: Wang, Yiming, et al.
Veröffentlicht: (2025)