MathScale: Scaling Instruction Tuning for Mathematical Reasoning
Fuente:
arXiv
Guardado en:
| Autores principales: | Tang, Zhengyang, Zhang, Xingxing, Wang, Benyou, Wei, Furu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Skywork-Math: Data Scaling Laws for Mathematical Reasoning in Large Language Models -- The Story Goes On
por: Zeng, Liang, et al.
Publicado: (2024)
por: Zeng, Liang, et al.
Publicado: (2024)
ScaleDiff: Scaling Difficult Problems for Advanced Mathematical Reasoning
por: Pei, Qizhi, et al.
Publicado: (2025)
por: Pei, Qizhi, et al.
Publicado: (2025)
Generative Representational Instruction Tuning
por: Muennighoff, Niklas, et al.
Publicado: (2024)
por: Muennighoff, Niklas, et al.
Publicado: (2024)
MathPile: A Billion-Token-Scale Pretraining Corpus for Math
por: Wang, Zengzhi, et al.
Publicado: (2023)
por: Wang, Zengzhi, et al.
Publicado: (2023)
OpenMathInstruct-1: A 1.8 Million Math Instruction Tuning Dataset
por: Toshniwal, Shubham, et al.
Publicado: (2024)
por: Toshniwal, Shubham, et al.
Publicado: (2024)
DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
por: Shao, Zhihong, et al.
Publicado: (2024)
por: Shao, Zhihong, et al.
Publicado: (2024)
AgentMath: Empowering Mathematical Reasoning for Large Language Models via Tool-Augmented Agent
por: Luo, Haipeng, et al.
Publicado: (2025)
por: Luo, Haipeng, et al.
Publicado: (2025)
WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct
por: Luo, Haipeng, et al.
Publicado: (2023)
por: Luo, Haipeng, et al.
Publicado: (2023)
DotaMath: Decomposition of Thought with Code Assistance and Self-correction for Mathematical Reasoning
por: Li, Chengpeng, et al.
Publicado: (2024)
por: Li, Chengpeng, et al.
Publicado: (2024)
Understanding and Mitigating Spurious Signal Amplification in Test-Time Reinforcement Learning for Math Reasoning
por: Yu, Yongcan, et al.
Publicado: (2026)
por: Yu, Yongcan, et al.
Publicado: (2026)
Scaling Optimal LR Across Token Horizons
por: Bjorck, Johan, et al.
Publicado: (2024)
por: Bjorck, Johan, et al.
Publicado: (2024)
AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset
por: Moshkov, Ivan, et al.
Publicado: (2025)
por: Moshkov, Ivan, et al.
Publicado: (2025)
A Toolbox, Not a Hammer -- Multi-TAG: Scaling Math Reasoning with Multi-Tool Aggregation
por: Yao, Bohan, et al.
Publicado: (2025)
por: Yao, Bohan, et al.
Publicado: (2025)
CoRT: Code-integrated Reasoning within Thinking
por: Li, Chengpeng, et al.
Publicado: (2025)
por: Li, Chengpeng, et al.
Publicado: (2025)
Examining False Positives under Inference Scaling for Mathematical Reasoning
por: Wang, Yu, et al.
Publicado: (2025)
por: Wang, Yu, et al.
Publicado: (2025)
OpenMathInstruct-2: Accelerating AI for Math with Massive Open-Source Instruction Data
por: Toshniwal, Shubham, et al.
Publicado: (2024)
por: Toshniwal, Shubham, et al.
Publicado: (2024)
PhoneWorld: Scaling Phone-Use Agent Environments
por: Tang, Zhengyang, et al.
Publicado: (2026)
por: Tang, Zhengyang, et al.
Publicado: (2026)
AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling
por: Liu, Zihan, et al.
Publicado: (2024)
por: Liu, Zihan, et al.
Publicado: (2024)
Nemotron-CC-Math: A 133 Billion-Token-Scale High Quality Math Pretraining Dataset
por: Mahabadi, Rabeeh Karimi, et al.
Publicado: (2025)
por: Mahabadi, Rabeeh Karimi, et al.
Publicado: (2025)
Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
por: Albalak, Alon, et al.
Publicado: (2025)
por: Albalak, Alon, et al.
Publicado: (2025)
Confucius3-Math: A Lightweight High-Performance Reasoning LLM for Chinese K-12 Mathematics Learning
por: Wu, Lixin, et al.
Publicado: (2025)
por: Wu, Lixin, et al.
Publicado: (2025)
Scaling Reasoning without Attention
por: Zhao, Xueliang, et al.
Publicado: (2025)
por: Zhao, Xueliang, et al.
Publicado: (2025)
MuggleMath: Assessing the Impact of Query and Response Augmentation on Math Reasoning
por: Li, Chengpeng, et al.
Publicado: (2023)
por: Li, Chengpeng, et al.
Publicado: (2023)
PARAMANU-GANITA: Can Small Math Language Models Rival with Large Language Models on Mathematical Reasoning?
por: Niyogi, Mitodru, et al.
Publicado: (2024)
por: Niyogi, Mitodru, et al.
Publicado: (2024)
Scaling Laws for Predicting Downstream Performance in LLMs
por: Chen, Yangyi, et al.
Publicado: (2024)
por: Chen, Yangyi, et al.
Publicado: (2024)
MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts
por: Lu, Pan, et al.
Publicado: (2023)
por: Lu, Pan, et al.
Publicado: (2023)
Multiplex Thinking: Reasoning via Token-wise Branch-and-Merge
por: Tang, Yao, et al.
Publicado: (2026)
por: Tang, Yao, et al.
Publicado: (2026)
Scaling over Scaling: Exploring Test-Time Scaling Plateau in Large Reasoning Models
por: Wang, Jian, et al.
Publicado: (2025)
por: Wang, Jian, et al.
Publicado: (2025)
VisAidMath: Benchmarking Visual-Aided Mathematical Reasoning
por: Ma, Jingkun, et al.
Publicado: (2024)
por: Ma, Jingkun, et al.
Publicado: (2024)
RAG-Instruct: Boosting LLMs with Diverse Retrieval-Augmented Instructions
por: Liu, Wanlong, et al.
Publicado: (2024)
por: Liu, Wanlong, et al.
Publicado: (2024)
Towards Understanding Fine-Tuning Mechanisms of LLMs via Circuit Analysis
por: Wang, Xu, et al.
Publicado: (2025)
por: Wang, Xu, et al.
Publicado: (2025)
AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning
por: Chen, Yang, et al.
Publicado: (2025)
por: Chen, Yang, et al.
Publicado: (2025)
MegaMath: Pushing the Limits of Open Math Corpora
por: Zhou, Fan, et al.
Publicado: (2025)
por: Zhou, Fan, et al.
Publicado: (2025)
MATH-Perturb: Benchmarking LLMs' Math Reasoning Abilities against Hard Perturbations
por: Huang, Kaixuan, et al.
Publicado: (2025)
por: Huang, Kaixuan, et al.
Publicado: (2025)
NPG-Muse: Scaling Long Chain-of-Thought Reasoning with NP-Hard Graph Problems
por: Wang, Yuyao, et al.
Publicado: (2025)
por: Wang, Yuyao, et al.
Publicado: (2025)
Contrastive Instruction Tuning
por: Yan, Tianyi Lorena, et al.
Publicado: (2024)
por: Yan, Tianyi Lorena, et al.
Publicado: (2024)
OptScale: Probabilistic Optimality for Inference-time Scaling
por: Wang, Youkang, et al.
Publicado: (2025)
por: Wang, Youkang, et al.
Publicado: (2025)
ReasonFlux: Hierarchical LLM Reasoning via Scaling Thought Templates
por: Yang, Ling, et al.
Publicado: (2025)
por: Yang, Ling, et al.
Publicado: (2025)
AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy
por: Liu, Zihan, et al.
Publicado: (2025)
por: Liu, Zihan, et al.
Publicado: (2025)
Instruction Tuning for Large Language Models: A Survey
por: Zhang, Shengyu, et al.
Publicado: (2023)
por: Zhang, Shengyu, et al.
Publicado: (2023)
Ejemplares similares
-
Skywork-Math: Data Scaling Laws for Mathematical Reasoning in Large Language Models -- The Story Goes On
por: Zeng, Liang, et al.
Publicado: (2024) -
ScaleDiff: Scaling Difficult Problems for Advanced Mathematical Reasoning
por: Pei, Qizhi, et al.
Publicado: (2025) -
Generative Representational Instruction Tuning
por: Muennighoff, Niklas, et al.
Publicado: (2024) -
MathPile: A Billion-Token-Scale Pretraining Corpus for Math
por: Wang, Zengzhi, et al.
Publicado: (2023) -
OpenMathInstruct-1: A 1.8 Million Math Instruction Tuning Dataset
por: Toshniwal, Shubham, et al.
Publicado: (2024)