Guardado en:
| Autores principales: | Zhang, Yuanhe, Kuzborskij, Ilja, Lee, Jason D., Leng, Chenlei, Liu, Fanghui |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2510.19842 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
LoRA-One: One-Step Full Gradient Could Suffice for Fine-Tuning Large Language Models, Provably and Efficiently
por: Zhang, Yuanhe, et al.
Publicado: (2025)
por: Zhang, Yuanhe, et al.
Publicado: (2025)
To Believe or Not to Believe Your LLM
por: Yadkori, Yasin Abbasi, et al.
Publicado: (2024)
por: Yadkori, Yasin Abbasi, et al.
Publicado: (2024)
DotaMath: Decomposition of Thought with Code Assistance and Self-correction for Mathematical Reasoning
por: Li, Chengpeng, et al.
Publicado: (2024)
por: Li, Chengpeng, et al.
Publicado: (2024)
Sufficient Conditions for Stability of Minimum-Norm Interpolating Deep ReLU Networks
por: Harzli, Ouns El, et al.
Publicado: (2026)
por: Harzli, Ouns El, et al.
Publicado: (2026)
A Novel Spatiotemporal Coupling Graph Convolutional Network
por: Bi, Fanghui
Publicado: (2024)
por: Bi, Fanghui
Publicado: (2024)
Statistical Learning Theory in Lean 4: Empirical Processes from Scratch
por: Zhang, Yuanhe, et al.
Publicado: (2026)
por: Zhang, Yuanhe, et al.
Publicado: (2026)
MathScale: Scaling Instruction Tuning for Mathematical Reasoning
por: Tang, Zhengyang, et al.
Publicado: (2024)
por: Tang, Zhengyang, et al.
Publicado: (2024)
ARIES: Autonomous Reasoning with LLMs on Interactive Thought Graph Environments
por: Gimenes, Pedro, et al.
Publicado: (2025)
por: Gimenes, Pedro, et al.
Publicado: (2025)
Limits of PRM-Guided Tree Search for Mathematical Reasoning with LLMs
por: Cinquin, Tristan, et al.
Publicado: (2025)
por: Cinquin, Tristan, et al.
Publicado: (2025)
MathSE: Improving Multimodal Mathematical Reasoning via Self-Evolving Iterative Reflection and Reward-Guided Fine-Tuning
por: Chen, Jinhao, et al.
Publicado: (2025)
por: Chen, Jinhao, et al.
Publicado: (2025)
Pointwise confidence estimation in the non-linear $\ell^2$-regularized least squares
por: Kuzborskij, Ilja, et al.
Publicado: (2025)
por: Kuzborskij, Ilja, et al.
Publicado: (2025)
Low-rank bias, weight decay, and model merging in neural networks
por: Kuzborskij, Ilja, et al.
Publicado: (2025)
por: Kuzborskij, Ilja, et al.
Publicado: (2025)
MathMixup: Boosting LLM Mathematical Reasoning with Difficulty-Controllable Data Synthesis and Curriculum Learning
por: Li, Xuchen, et al.
Publicado: (2026)
por: Li, Xuchen, et al.
Publicado: (2026)
DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
por: Shao, Zhihong, et al.
Publicado: (2024)
por: Shao, Zhihong, et al.
Publicado: (2024)
Metastable Dynamics of Chain-of-Thought Reasoning: Provable Benefits of Search, RL and Distillation
por: Kim, Juno, et al.
Publicado: (2025)
por: Kim, Juno, et al.
Publicado: (2025)
Guide: Generalized-Prior and Data Encoders for DAG Estimation
por: Roy, Amartya, et al.
Publicado: (2025)
por: Roy, Amartya, et al.
Publicado: (2025)
DAG-AFL:Directed Acyclic Graph-based Asynchronous Federated Learning
por: Zhang, Shuaipeng, et al.
Publicado: (2025)
por: Zhang, Shuaipeng, et al.
Publicado: (2025)
We-Math 2.0: A Versatile MathBook System for Incentivizing Visual Mathematical Reasoning
por: Qiao, Runqi, et al.
Publicado: (2025)
por: Qiao, Runqi, et al.
Publicado: (2025)
Skywork-Math: Data Scaling Laws for Mathematical Reasoning in Large Language Models -- The Story Goes On
por: Zeng, Liang, et al.
Publicado: (2024)
por: Zeng, Liang, et al.
Publicado: (2024)
An Empirical Study of Data Ability Boundary in LLMs' Math Reasoning
por: Chen, Zui, et al.
Publicado: (2024)
por: Chen, Zui, et al.
Publicado: (2024)
Transformers Learn to Implement Multi-step Gradient Descent with Chain of Thought
por: Huang, Jianhao, et al.
Publicado: (2025)
por: Huang, Jianhao, et al.
Publicado: (2025)
WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct
por: Luo, Haipeng, et al.
Publicado: (2023)
por: Luo, Haipeng, et al.
Publicado: (2023)
CoMAT: Chain of Mathematically Annotated Thought Improves Mathematical Reasoning
por: Leang, Joshua Ong Jun, et al.
Publicado: (2024)
por: Leang, Joshua Ong Jun, et al.
Publicado: (2024)
MathNet: a Global Multimodal Benchmark for Mathematical Reasoning and Retrieval
por: Alshammari, Shaden, et al.
Publicado: (2026)
por: Alshammari, Shaden, et al.
Publicado: (2026)
Quantization Meets Reasoning: Exploring and Mitigating Degradation of Low-Bit LLMs in Mathematical Reasoning
por: Li, Zhen, et al.
Publicado: (2025)
por: Li, Zhen, et al.
Publicado: (2025)
AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset
por: Moshkov, Ivan, et al.
Publicado: (2025)
por: Moshkov, Ivan, et al.
Publicado: (2025)
MATH-Perturb: Benchmarking LLMs' Math Reasoning Abilities against Hard Perturbations
por: Huang, Kaixuan, et al.
Publicado: (2025)
por: Huang, Kaixuan, et al.
Publicado: (2025)
Retrieval-of-Thought: Efficient Reasoning via Reusing Thoughts
por: Ahmed, Ammar, et al.
Publicado: (2025)
por: Ahmed, Ammar, et al.
Publicado: (2025)
AgentMath: Empowering Mathematical Reasoning for Large Language Models via Tool-Augmented Agent
por: Luo, Haipeng, et al.
Publicado: (2025)
por: Luo, Haipeng, et al.
Publicado: (2025)
MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts
por: Lu, Pan, et al.
Publicado: (2023)
por: Lu, Pan, et al.
Publicado: (2023)
VisAidMath: Benchmarking Visual-Aided Mathematical Reasoning
por: Ma, Jingkun, et al.
Publicado: (2024)
por: Ma, Jingkun, et al.
Publicado: (2024)
AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling
por: Liu, Zihan, et al.
Publicado: (2024)
por: Liu, Zihan, et al.
Publicado: (2024)
AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning
por: Chen, Yang, et al.
Publicado: (2025)
por: Chen, Yang, et al.
Publicado: (2025)
SeaDAG: Semi-autoregressive Diffusion for Conditional Directed Acyclic Graph Generation
por: Zhou, Xinyi, et al.
Publicado: (2024)
por: Zhou, Xinyi, et al.
Publicado: (2024)
PARAMANU-GANITA: Can Small Math Language Models Rival with Large Language Models on Mathematical Reasoning?
por: Niyogi, Mitodru, et al.
Publicado: (2024)
por: Niyogi, Mitodru, et al.
Publicado: (2024)
MathConstraint: Automated Generation of Verified Combinatorial Reasoning Instances for LLMs
por: Pati, Viresh, et al.
Publicado: (2026)
por: Pati, Viresh, et al.
Publicado: (2026)
Value-Guided Search for Efficient Chain-of-Thought Reasoning
por: Wang, Kaiwen, et al.
Publicado: (2025)
por: Wang, Kaiwen, et al.
Publicado: (2025)
PA3: Policy-Aware Agent Alignment through Chain-of-Thought
por: Dipta, Shubhashis Roy, et al.
Publicado: (2026)
por: Dipta, Shubhashis Roy, et al.
Publicado: (2026)
An Investigation of Robustness of LLMs in Mathematical Reasoning: Benchmarking with Mathematically-Equivalent Transformation of Advanced Mathematical Problems
por: Hao, Yuren, et al.
Publicado: (2025)
por: Hao, Yuren, et al.
Publicado: (2025)
Scaling Graph Chain-of-Thought Reasoning: A Multi-Agent Framework with Efficient LLM Serving
por: Huan, Chengying, et al.
Publicado: (2025)
por: Huan, Chengying, et al.
Publicado: (2025)
Ejemplares similares
-
LoRA-One: One-Step Full Gradient Could Suffice for Fine-Tuning Large Language Models, Provably and Efficiently
por: Zhang, Yuanhe, et al.
Publicado: (2025) -
To Believe or Not to Believe Your LLM
por: Yadkori, Yasin Abbasi, et al.
Publicado: (2024) -
DotaMath: Decomposition of Thought with Code Assistance and Self-correction for Mathematical Reasoning
por: Li, Chengpeng, et al.
Publicado: (2024) -
Sufficient Conditions for Stability of Minimum-Norm Interpolating Deep ReLU Networks
por: Harzli, Ouns El, et al.
Publicado: (2026) -
A Novel Spatiotemporal Coupling Graph Convolutional Network
por: Bi, Fanghui
Publicado: (2024)