ControlMath: Controllable Data Generation Promotes Math Generalist Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Chen, Nuo, Wu, Ning, Chang, Jianhui, Li, Jia |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MuggleMath: Assessing the Impact of Query and Response Augmentation on Math Reasoning
por: Li, Chengpeng, et al.
Publicado: (2023)
por: Li, Chengpeng, et al.
Publicado: (2023)
AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling
por: Liu, Zihan, et al.
Publicado: (2024)
por: Liu, Zihan, et al.
Publicado: (2024)
MegaMath: Pushing the Limits of Open Math Corpora
por: Zhou, Fan, et al.
Publicado: (2025)
por: Zhou, Fan, et al.
Publicado: (2025)
MathPile: A Billion-Token-Scale Pretraining Corpus for Math
por: Wang, Zengzhi, et al.
Publicado: (2023)
por: Wang, Zengzhi, et al.
Publicado: (2023)
OpenMathInstruct-1: A 1.8 Million Math Instruction Tuning Dataset
por: Toshniwal, Shubham, et al.
Publicado: (2024)
por: Toshniwal, Shubham, et al.
Publicado: (2024)
OpenMathInstruct-2: Accelerating AI for Math with Massive Open-Source Instruction Data
por: Toshniwal, Shubham, et al.
Publicado: (2024)
por: Toshniwal, Shubham, et al.
Publicado: (2024)
An Empirical Study of Data Ability Boundary in LLMs' Math Reasoning
por: Chen, Zui, et al.
Publicado: (2024)
por: Chen, Zui, et al.
Publicado: (2024)
Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
por: Albalak, Alon, et al.
Publicado: (2025)
por: Albalak, Alon, et al.
Publicado: (2025)
Nemotron-CC-Math: A 133 Billion-Token-Scale High Quality Math Pretraining Dataset
por: Mahabadi, Rabeeh Karimi, et al.
Publicado: (2025)
por: Mahabadi, Rabeeh Karimi, et al.
Publicado: (2025)
DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
por: Shao, Zhihong, et al.
Publicado: (2024)
por: Shao, Zhihong, et al.
Publicado: (2024)
Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations
por: Wang, Peiyi, et al.
Publicado: (2023)
por: Wang, Peiyi, et al.
Publicado: (2023)
Executable Functional Abstractions: Inferring Generative Programs for Advanced Math Problems
por: Khan, Zaid, et al.
Publicado: (2025)
por: Khan, Zaid, et al.
Publicado: (2025)
Skywork-Math: Data Scaling Laws for Mathematical Reasoning in Large Language Models -- The Story Goes On
por: Zeng, Liang, et al.
Publicado: (2024)
por: Zeng, Liang, et al.
Publicado: (2024)
MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?
por: Zhang, Renrui, et al.
Publicado: (2024)
por: Zhang, Renrui, et al.
Publicado: (2024)
MATH-Perturb: Benchmarking LLMs' Math Reasoning Abilities against Hard Perturbations
por: Huang, Kaixuan, et al.
Publicado: (2025)
por: Huang, Kaixuan, et al.
Publicado: (2025)
MathScale: Scaling Instruction Tuning for Mathematical Reasoning
por: Tang, Zhengyang, et al.
Publicado: (2024)
por: Tang, Zhengyang, et al.
Publicado: (2024)
To Code or not to Code? Adaptive Tool Integration for Math Language Models via Expectation-Maximization
por: Wang, Haozhe, et al.
Publicado: (2025)
por: Wang, Haozhe, et al.
Publicado: (2025)
Teaching LLMs for Step-Level Automatic Math Correction via Reinforcement Learning
por: Li, Junsong, et al.
Publicado: (2025)
por: Li, Junsong, et al.
Publicado: (2025)
Investigating Bias: A Multilingual Pipeline for Generating, Solving, and Evaluating Math Problems with LLMs
por: Mahran, Mariam, et al.
Publicado: (2025)
por: Mahran, Mariam, et al.
Publicado: (2025)
Physics of Language Models: Part 2.1, Grade-School Math and the Hidden Reasoning Process
por: Ye, Tian, et al.
Publicado: (2024)
por: Ye, Tian, et al.
Publicado: (2024)
Augmenting Math Word Problems via Iterative Question Composing
por: Liu, Haoxiong, et al.
Publicado: (2024)
por: Liu, Haoxiong, et al.
Publicado: (2024)
BrokenMath: A Benchmark for Sycophancy in Theorem Proving with LLMs
por: Petrov, Ivo, et al.
Publicado: (2025)
por: Petrov, Ivo, et al.
Publicado: (2025)
WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct
por: Luo, Haipeng, et al.
Publicado: (2023)
por: Luo, Haipeng, et al.
Publicado: (2023)
MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts
por: Lu, Pan, et al.
Publicado: (2023)
por: Lu, Pan, et al.
Publicado: (2023)
DotaMath: Decomposition of Thought with Code Assistance and Self-correction for Mathematical Reasoning
por: Li, Chengpeng, et al.
Publicado: (2024)
por: Li, Chengpeng, et al.
Publicado: (2024)
Reasoning, Code, or Both? How Large Language Models Handle Variations in Math Questions
por: Kutakh, Matthew
Publicado: (2026)
por: Kutakh, Matthew
Publicado: (2026)
AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning
por: Chen, Yang, et al.
Publicado: (2025)
por: Chen, Yang, et al.
Publicado: (2025)
FLAMES: Improving LLM Math Reasoning via a Fine-Grained Analysis of the Data Synthesis Pipeline
por: Seegmiller, Parker, et al.
Publicado: (2025)
por: Seegmiller, Parker, et al.
Publicado: (2025)
MathGAP: Out-of-Distribution Evaluation on Problems with Arbitrarily Complex Proofs
por: Opedal, Andreas, et al.
Publicado: (2024)
por: Opedal, Andreas, et al.
Publicado: (2024)
Decomposing Elements of Problem Solving: What "Math" Does RL Teach?
por: Qin, Tian, et al.
Publicado: (2025)
por: Qin, Tian, et al.
Publicado: (2025)
PCL-Reasoner-V1.5: Advancing Math Reasoning with Offline Reinforcement Learning
por: Lu, Yao, et al.
Publicado: (2026)
por: Lu, Yao, et al.
Publicado: (2026)
Physics of Language Models: Part 2.2, How to Learn From Mistakes on Grade-School Math Problems
por: Ye, Tian, et al.
Publicado: (2024)
por: Ye, Tian, et al.
Publicado: (2024)
PARAMANU-GANITA: Can Small Math Language Models Rival with Large Language Models on Mathematical Reasoning?
por: Niyogi, Mitodru, et al.
Publicado: (2024)
por: Niyogi, Mitodru, et al.
Publicado: (2024)
Leveraging Online Olympiad-Level Math Problems for LLMs Training and Contamination-Resistant Evaluation
por: Mahdavi, Sadegh, et al.
Publicado: (2025)
por: Mahdavi, Sadegh, et al.
Publicado: (2025)
AgentMath: Empowering Mathematical Reasoning for Large Language Models via Tool-Augmented Agent
por: Luo, Haipeng, et al.
Publicado: (2025)
por: Luo, Haipeng, et al.
Publicado: (2025)
Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement
por: Yang, An, et al.
Publicado: (2024)
por: Yang, An, et al.
Publicado: (2024)
FormalProofBench: Can Models Write Graduate Level Math Proofs That Are Formally Verified?
por: Ravi, Nikil, et al.
Publicado: (2026)
por: Ravi, Nikil, et al.
Publicado: (2026)
Rewarding Graph Reasoning Process makes LLMs more Generalized Reasoners
por: Peng, Miao, et al.
Publicado: (2025)
por: Peng, Miao, et al.
Publicado: (2025)
Confucius3-Math: A Lightweight High-Performance Reasoning LLM for Chinese K-12 Mathematics Learning
por: Wu, Lixin, et al.
Publicado: (2025)
por: Wu, Lixin, et al.
Publicado: (2025)
EasyMath: A 0-shot Math Benchmark for SLMs
por: Karki, Drishya, et al.
Publicado: (2025)
por: Karki, Drishya, et al.
Publicado: (2025)
Ejemplares similares
-
MuggleMath: Assessing the Impact of Query and Response Augmentation on Math Reasoning
por: Li, Chengpeng, et al.
Publicado: (2023) -
AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling
por: Liu, Zihan, et al.
Publicado: (2024) -
MegaMath: Pushing the Limits of Open Math Corpora
por: Zhou, Fan, et al.
Publicado: (2025) -
MathPile: A Billion-Token-Scale Pretraining Corpus for Math
por: Wang, Zengzhi, et al.
Publicado: (2023) -
OpenMathInstruct-1: A 1.8 Million Math Instruction Tuning Dataset
por: Toshniwal, Shubham, et al.
Publicado: (2024)