Nemotron-Math: Efficient Long-Context Distillation of Mathematical Reasoning from Multi-Mode Supervision
Fuente:
arXiv
Guardado en:
| Autores principales: | Du, Wei, Toshniwal, Shubham, Kisacanin, Branislav, Mahdavi, Sadegh, Moshkov, Ivan, Armstrong, George, Ge, Stephen, Minasyan, Edgar, Chen, Feng, Gitman, Igor |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
OpenMathInstruct-2: Accelerating AI for Math with Massive Open-Source Instruction Data
por: Toshniwal, Shubham, et al.
Publicado: (2024)
por: Toshniwal, Shubham, et al.
Publicado: (2024)
Scaling Generative Verifiers For Natural Language Mathematical Proof Verification And Selection
por: Mahdavi, Sadegh, et al.
Publicado: (2025)
por: Mahdavi, Sadegh, et al.
Publicado: (2025)
OpenMathInstruct-1: A 1.8 Million Math Instruction Tuning Dataset
por: Toshniwal, Shubham, et al.
Publicado: (2024)
por: Toshniwal, Shubham, et al.
Publicado: (2024)
AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset
por: Moshkov, Ivan, et al.
Publicado: (2025)
por: Moshkov, Ivan, et al.
Publicado: (2025)
GenSelect: A Generative Approach to Best-of-N
por: Toshniwal, Shubham, et al.
Publicado: (2025)
por: Toshniwal, Shubham, et al.
Publicado: (2025)
The Challenge of Teaching Reasoning to LLMs Without RL or Distillation
por: Du, Wei, et al.
Publicado: (2025)
por: Du, Wei, et al.
Publicado: (2025)
Learning Generative Selection for Best-of-N
por: Toshniwal, Shubham, et al.
Publicado: (2026)
por: Toshniwal, Shubham, et al.
Publicado: (2026)
Llama-Nemotron: Efficient Reasoning Models
por: Bercovich, Akhiad, et al.
Publicado: (2025)
por: Bercovich, Akhiad, et al.
Publicado: (2025)
AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning
por: Chen, Yang, et al.
Publicado: (2025)
por: Chen, Yang, et al.
Publicado: (2025)
IdentifyMe: A Challenging Long-Context Mention Resolution Benchmark for LLMs
por: Manikantan, Kawshik, et al.
Publicado: (2024)
por: Manikantan, Kawshik, et al.
Publicado: (2024)
Nemotron-CrossThink: Scaling Self-Learning beyond Math Reasoning
por: Akter, Syeda Nahida, et al.
Publicado: (2025)
por: Akter, Syeda Nahida, et al.
Publicado: (2025)
MathHay: An Automated Benchmark for Long-Context Mathematical Reasoning in LLMs
por: Wang, Lei, et al.
Publicado: (2024)
por: Wang, Lei, et al.
Publicado: (2024)
AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy
por: Liu, Zihan, et al.
Publicado: (2025)
por: Liu, Zihan, et al.
Publicado: (2025)
PolyMath: Evaluating Mathematical Reasoning in Multilingual Contexts
por: Wang, Yiming, et al.
Publicado: (2025)
por: Wang, Yiming, et al.
Publicado: (2025)
NeMo-Inspector: A Visualization Tool for LLM Generation Analysis
por: Gitman, Daria, et al.
Publicado: (2025)
por: Gitman, Daria, et al.
Publicado: (2025)
Nemotron-CC-Math: A 133 Billion-Token-Scale High Quality Math Pretraining Dataset
por: Mahabadi, Rabeeh Karimi, et al.
Publicado: (2025)
por: Mahabadi, Rabeeh Karimi, et al.
Publicado: (2025)
MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts
por: Lu, Pan, et al.
Publicado: (2023)
por: Lu, Pan, et al.
Publicado: (2023)
Nemotron-4 340B Technical Report
por: Nvidia, et al.
Publicado: (2024)
por: Nvidia, et al.
Publicado: (2024)
From Large to Tiny: Distilling and Refining Mathematical Expertise for Math Word Problems with Weakly Supervision
por: Lin, Qingwen, et al.
Publicado: (2024)
por: Lin, Qingwen, et al.
Publicado: (2024)
Leveraging Online Olympiad-Level Math Problems for LLMs Training and Contamination-Resistant Evaluation
por: Mahdavi, Sadegh, et al.
Publicado: (2025)
por: Mahdavi, Sadegh, et al.
Publicado: (2025)
Code Pretraining Improves Entity Tracking Abilities of Language Models
por: Kim, Najoung, et al.
Publicado: (2024)
por: Kim, Najoung, et al.
Publicado: (2024)
Nemotron 3 Nano: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning
por: NVIDIA, et al.
Publicado: (2025)
por: NVIDIA, et al.
Publicado: (2025)
NVIDIA Nemotron 3: Efficient and Open Intelligence
por: NVIDIA, et al.
Publicado: (2025)
por: NVIDIA, et al.
Publicado: (2025)
Programs Versus Finite Tree-Programs
por: Moshkov, Mikhail
Publicado: (2025)
por: Moshkov, Mikhail
Publicado: (2025)
Algorithmic Problems for Computation Trees
por: Moshkov, Mikhail
Publicado: (2025)
por: Moshkov, Mikhail
Publicado: (2025)
NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model
por: NVIDIA, et al.
Publicado: (2025)
por: NVIDIA, et al.
Publicado: (2025)
On double coset separability and the Wilson-Zalesskii property
por: Minasyan, Ashot
Publicado: (2022)
por: Minasyan, Ashot
Publicado: (2022)
Property (LR) and an embedding theorem for virtually free groups
por: Minasyan, Ashot
Publicado: (2026)
por: Minasyan, Ashot
Publicado: (2026)
RoMath: A Mathematical Reasoning Benchmark in Romanian
por: Cosma, Adrian, et al.
Publicado: (2024)
por: Cosma, Adrian, et al.
Publicado: (2024)
MathScale: Scaling Instruction Tuning for Mathematical Reasoning
por: Tang, Zhengyang, et al.
Publicado: (2024)
por: Tang, Zhengyang, et al.
Publicado: (2024)
Nemotron Elastic: Towards Efficient Many-in-One Reasoning LLMs
por: Taghibakhshi, Ali, et al.
Publicado: (2025)
por: Taghibakhshi, Ali, et al.
Publicado: (2025)
Advantage Shaping as Surrogate Reward Maximization: Unifying Pass@K Policy Gradients
por: Thrampoulidis, Christos, et al.
Publicado: (2025)
por: Thrampoulidis, Christos, et al.
Publicado: (2025)
Memorization Capacity of Multi-Head Attention in Transformers
por: Mahdavi, Sadegh, et al.
Publicado: (2023)
por: Mahdavi, Sadegh, et al.
Publicado: (2023)
Major Entity Identification: A Generalizable Alternative to Coreference Resolution
por: Manikantan, Kawshik, et al.
Publicado: (2024)
por: Manikantan, Kawshik, et al.
Publicado: (2024)
Nemotron 3 Super: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning
por: NVIDIA, et al.
Publicado: (2026)
por: NVIDIA, et al.
Publicado: (2026)
VisAidMath: Benchmarking Visual-Aided Mathematical Reasoning
por: Ma, Jingkun, et al.
Publicado: (2024)
por: Ma, Jingkun, et al.
Publicado: (2024)
DAG-Math: Graph-of-Thought Guided Mathematical Reasoning in LLMs
por: Zhang, Yuanhe, et al.
Publicado: (2025)
por: Zhang, Yuanhe, et al.
Publicado: (2025)
VisioMath: Benchmarking Figure-based Mathematical Reasoning in LMMs
por: Li, Can, et al.
Publicado: (2025)
por: Li, Can, et al.
Publicado: (2025)
Zero-Shot Motor Imagery BCI via Self-Supervised Contrastive Learning
por: Branislav Ceperkovic
Publicado: (2025)
por: Branislav Ceperkovic
Publicado: (2025)
We-Math 2.0: A Versatile MathBook System for Incentivizing Visual Mathematical Reasoning
por: Qiao, Runqi, et al.
Publicado: (2025)
por: Qiao, Runqi, et al.
Publicado: (2025)
Ejemplares similares
-
OpenMathInstruct-2: Accelerating AI for Math with Massive Open-Source Instruction Data
por: Toshniwal, Shubham, et al.
Publicado: (2024) -
Scaling Generative Verifiers For Natural Language Mathematical Proof Verification And Selection
por: Mahdavi, Sadegh, et al.
Publicado: (2025) -
OpenMathInstruct-1: A 1.8 Million Math Instruction Tuning Dataset
por: Toshniwal, Shubham, et al.
Publicado: (2024) -
AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset
por: Moshkov, Ivan, et al.
Publicado: (2025) -
GenSelect: A Generative Approach to Best-of-N
por: Toshniwal, Shubham, et al.
Publicado: (2025)