Guardado en:
| Autores principales: | Jung, Jaehun, Han, Seungju, Lu, Ximing, Hallinan, Skyler, Acuna, David, Prabhumoye, Shrimai, Patwary, Mostafa, Shoeybi, Mohammad, Catanzaro, Bryan, Choi, Yejin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2505.20161 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Retro-Search: Exploring Untaken Paths for Deeper and Efficient Reasoning
por: Lu, Ximing, et al.
Publicado: (2025)
por: Lu, Ximing, et al.
Publicado: (2025)
Front-Loading Reasoning: The Synergy between Pretraining and Post-Training Data
por: Akter, Syeda Nahida, et al.
Publicado: (2025)
por: Akter, Syeda Nahida, et al.
Publicado: (2025)
Nemotron-CrossThink: Scaling Self-Learning beyond Math Reasoning
por: Akter, Syeda Nahida, et al.
Publicado: (2025)
por: Akter, Syeda Nahida, et al.
Publicado: (2025)
Maximize Your Data's Potential: Enhancing LLM Accuracy with Two-Phase Pretraining
por: Feng, Steven, et al.
Publicado: (2024)
por: Feng, Steven, et al.
Publicado: (2024)
RLP: Reinforcement as a Pretraining Objective
por: Hatamizadeh, Ali, et al.
Publicado: (2025)
por: Hatamizadeh, Ali, et al.
Publicado: (2025)
Nemotron-CC-Math: A 133 Billion-Token-Scale High Quality Math Pretraining Dataset
por: Mahabadi, Rabeeh Karimi, et al.
Publicado: (2025)
por: Mahabadi, Rabeeh Karimi, et al.
Publicado: (2025)
iGRPO: Self-Feedback-Driven LLM Reasoning
por: Hatamizadeh, Ali, et al.
Publicado: (2026)
por: Hatamizadeh, Ali, et al.
Publicado: (2026)
MIND: Math Informed syNthetic Dialogues for Pretraining LLMs
por: Akter, Syeda Nahida, et al.
Publicado: (2024)
por: Akter, Syeda Nahida, et al.
Publicado: (2024)
Data, Data Everywhere: A Guide for Pretraining Dataset Construction
por: Parmar, Jupinder, et al.
Publicado: (2024)
por: Parmar, Jupinder, et al.
Publicado: (2024)
Introspective X Training: Feedback Conditioning Improves Scaling Across all LLM Training Stages
por: Cui, Brandon, et al.
Publicado: (2026)
por: Cui, Brandon, et al.
Publicado: (2026)
Reuse, Don't Retrain: A Recipe for Continued Pretraining of Language Models
por: Parmar, Jupinder, et al.
Publicado: (2024)
por: Parmar, Jupinder, et al.
Publicado: (2024)
FusionFactory: Fusing LLM Capabilities with Multi-LLM Log Data
por: Feng, Tao, et al.
Publicado: (2025)
por: Feng, Tao, et al.
Publicado: (2025)
Socratic-MCTS: Test-Time Visual Reasoning by Asking the Right Questions
por: Acuna, David, et al.
Publicado: (2025)
por: Acuna, David, et al.
Publicado: (2025)
StyleRemix: Interpretable Authorship Obfuscation via Distillation and Perturbation of Style Elements
por: Fisher, Jillian, et al.
Publicado: (2024)
por: Fisher, Jillian, et al.
Publicado: (2024)
The Surprising Effectiveness of Membership Inference with Simple N-Gram Coverage
por: Hallinan, Skyler, et al.
Publicado: (2025)
por: Hallinan, Skyler, et al.
Publicado: (2025)
Decompose, Mix, Adapt: A Unified Framework for Parameter-Efficient Neural Network Recombination and Compression
por: Tasnim, Nazia, et al.
Publicado: (2026)
por: Tasnim, Nazia, et al.
Publicado: (2026)
Verifying the Verifiers: Unveiling Pitfalls and Potentials in Fact Verifiers
por: Seo, Wooseok, et al.
Publicado: (2025)
por: Seo, Wooseok, et al.
Publicado: (2025)
AI as Humanity's Salieri: Quantifying Linguistic Creativity of Language Models via Systematic Attribution of Machine Text against Web Text
por: Lu, Ximing, et al.
Publicado: (2024)
por: Lu, Ximing, et al.
Publicado: (2024)
Trust or Escalate: LLM Judges with Provable Guarantees for Human Agreement
por: Jung, Jaehun, et al.
Publicado: (2024)
por: Jung, Jaehun, et al.
Publicado: (2024)
Long Grounded Thoughts: Synthesizing Visual Problems and Reasoning Chains at Scale
por: Acuna, David, et al.
Publicado: (2025)
por: Acuna, David, et al.
Publicado: (2025)
AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling
por: Liu, Zihan, et al.
Publicado: (2024)
por: Liu, Zihan, et al.
Publicado: (2024)
JAMDEC: Unsupervised Authorship Obfuscation using Constrained Decoding over Small Language Models
por: Fisher, Jillian, et al.
Publicado: (2024)
por: Fisher, Jillian, et al.
Publicado: (2024)
Synthetic Mixed Training: Scaling Parametric Knowledge Acquisition Beyond RAG
por: Han, Seungju, et al.
Publicado: (2026)
por: Han, Seungju, et al.
Publicado: (2026)
On Data Engineering for Scaling LLM Terminal Capabilities
por: Pi, Renjie, et al.
Publicado: (2026)
por: Pi, Renjie, et al.
Publicado: (2026)
Nemotron-CC: Transforming Common Crawl into a Refined Long-Horizon Pretraining Dataset
por: Su, Dan, et al.
Publicado: (2024)
por: Su, Dan, et al.
Publicado: (2024)
AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy
por: Liu, Zihan, et al.
Publicado: (2025)
por: Liu, Zihan, et al.
Publicado: (2025)
Tailoring Self-Rationalizers with Multi-Reward Distillation
por: Ramnath, Sahana, et al.
Publicado: (2023)
por: Ramnath, Sahana, et al.
Publicado: (2023)
Compact Language Models via Pruning and Knowledge Distillation
por: Muralidharan, Saurav, et al.
Publicado: (2024)
por: Muralidharan, Saurav, et al.
Publicado: (2024)
Information-Theoretic Distillation for Reference-less Summarization
por: Jung, Jaehun, et al.
Publicado: (2024)
por: Jung, Jaehun, et al.
Publicado: (2024)
AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning
por: Chen, Yang, et al.
Publicado: (2025)
por: Chen, Yang, et al.
Publicado: (2025)
Impossible Distillation: from Low-Quality Model to High-Quality Dataset & Model for Summarization and Paraphrasing
por: Jung, Jaehun, et al.
Publicado: (2023)
por: Jung, Jaehun, et al.
Publicado: (2023)
AgentKit: Structured LLM Reasoning with Dynamic Graphs
por: Wu, Yue, et al.
Publicado: (2024)
por: Wu, Yue, et al.
Publicado: (2024)
MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs
por: Lin, Sheng-Chieh, et al.
Publicado: (2024)
por: Lin, Sheng-Chieh, et al.
Publicado: (2024)
How to Instruct Your Robot: Dense Language Annotations Power Robot Policy Learning
por: Kim, Bosung, et al.
Publicado: (2026)
por: Kim, Bosung, et al.
Publicado: (2026)
ChatQA 2: Bridging the Gap to Proprietary LLMs in Long Context and RAG Capabilities
por: Xu, Peng, et al.
Publicado: (2024)
por: Xu, Peng, et al.
Publicado: (2024)
NV-Embed: Improved Techniques for Training LLMs as Generalist Embedding Models
por: Lee, Chankyu, et al.
Publicado: (2024)
por: Lee, Chankyu, et al.
Publicado: (2024)
ChatQA: Surpassing GPT-4 on Conversational QA and RAG
por: Liu, Zihan, et al.
Publicado: (2024)
por: Liu, Zihan, et al.
Publicado: (2024)
RAVEN: In-Context Learning with Retrieval-Augmented Encoder-Decoder Language Models
por: Huang, Jie, et al.
Publicado: (2023)
por: Huang, Jie, et al.
Publicado: (2023)
Amulet: Putting Complex Multi-Turn Conversations on the Stand with LLM Juries
por: Ramnath, Sahana, et al.
Publicado: (2025)
por: Ramnath, Sahana, et al.
Publicado: (2025)
ProfBench: Multi-Domain Rubrics requiring Professional Knowledge to Answer and Judge
por: Wang, Zhilin, et al.
Publicado: (2025)
por: Wang, Zhilin, et al.
Publicado: (2025)
Ejemplares similares
-
Retro-Search: Exploring Untaken Paths for Deeper and Efficient Reasoning
por: Lu, Ximing, et al.
Publicado: (2025) -
Front-Loading Reasoning: The Synergy between Pretraining and Post-Training Data
por: Akter, Syeda Nahida, et al.
Publicado: (2025) -
Nemotron-CrossThink: Scaling Self-Learning beyond Math Reasoning
por: Akter, Syeda Nahida, et al.
Publicado: (2025) -
Maximize Your Data's Potential: Enhancing LLM Accuracy with Two-Phase Pretraining
por: Feng, Steven, et al.
Publicado: (2024) -
RLP: Reinforcement as a Pretraining Objective
por: Hatamizadeh, Ali, et al.
Publicado: (2025)