AM-Thinking-v1: Advancing the Frontier of Reasoning at 32B Scale
Fuente:
arXiv
Salvato in:
| Autori principali: | Ji, Yunjie, Tian, Xiaoyu, Zhao, Sitong, Wang, Haotian, Chen, Shuaiting, Peng, Yiping, Zhao, Han, Li, Xiangang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Think Twice: Enhancing LLM Reasoning by Scaling Multi-round Test-time Thinking
di: Tian, Xiaoyu, et al.
Pubblicazione: (2025)
di: Tian, Xiaoyu, et al.
Pubblicazione: (2025)
DeepDistill: Enhancing LLM Reasoning Capabilities via Large-Scale Difficulty-Graded Data Training
di: Tian, Xiaoyu, et al.
Pubblicazione: (2025)
di: Tian, Xiaoyu, et al.
Pubblicazione: (2025)
Leveraging Reasoning Model Answers to Enhance Non-Reasoning Model Capability
di: Wang, Haotian, et al.
Pubblicazione: (2025)
di: Wang, Haotian, et al.
Pubblicazione: (2025)
1.4 Million Open-Source Distilled Reasoning Dataset to Empower Large Language Model Training
di: Zhao, Han, et al.
Pubblicazione: (2025)
di: Zhao, Han, et al.
Pubblicazione: (2025)
Exploring the Potential of Offline RL for Reasoning in LLMs: A Preliminary Study
di: Tian, Xiaoyu, et al.
Pubblicazione: (2025)
di: Tian, Xiaoyu, et al.
Pubblicazione: (2025)
How Difficulty-Aware Staged Reinforcement Learning Enhances LLMs' Reasoning Capabilities: A Preliminary Experimental Study
di: Ji, Yunjie, et al.
Pubblicazione: (2025)
di: Ji, Yunjie, et al.
Pubblicazione: (2025)
Not All Correct Answers Are Equal: Why Your Distillation Source Matters
di: Tian, Xiaoyu, et al.
Pubblicazione: (2025)
di: Tian, Xiaoyu, et al.
Pubblicazione: (2025)
Exploring Efficiency Frontiers of Thinking Budget in Medical Reasoning: Scaling Laws between Computational Resources and Reasoning Quality
di: Bi, Ziqian, et al.
Pubblicazione: (2025)
di: Bi, Ziqian, et al.
Pubblicazione: (2025)
Memorizing is Not Enough: Deep Knowledge Injection Through Reasoning
di: Xu, Ruoxi, et al.
Pubblicazione: (2025)
di: Xu, Ruoxi, et al.
Pubblicazione: (2025)
Navigate through Enigmatic Labyrinth A Survey of Chain of Thought Reasoning: Advances, Frontiers and Future
di: Chu, Zheng, et al.
Pubblicazione: (2023)
di: Chu, Zheng, et al.
Pubblicazione: (2023)
Thinking in Character: Advancing Role-Playing Agents with Role-Aware Reasoning
di: Tang, Yihong, et al.
Pubblicazione: (2025)
di: Tang, Yihong, et al.
Pubblicazione: (2025)
Think Deep, Not Just Long: Measuring LLM Reasoning Effort via Deep-Thinking Tokens
di: Chen, Wei-Lin, et al.
Pubblicazione: (2026)
di: Chen, Wei-Lin, et al.
Pubblicazione: (2026)
MeTHanol: Modularized Thinking Language Models with Intermediate Layer Thinking, Decoding and Bootstrapping Reasoning
di: Xi, Ningyuan, et al.
Pubblicazione: (2024)
di: Xi, Ningyuan, et al.
Pubblicazione: (2024)
SARI: Structured Audio Reasoning via Curriculum-Guided Reinforcement Learning
di: Wen, Cheng, et al.
Pubblicazione: (2025)
di: Wen, Cheng, et al.
Pubblicazione: (2025)
ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests
di: Xu, Shiyi, et al.
Pubblicazione: (2025)
di: Xu, Shiyi, et al.
Pubblicazione: (2025)
To Think or Not To Think, That is The Question for Large Reasoning Models in Theory of Mind Tasks
di: Gong, Nanxu, et al.
Pubblicazione: (2026)
di: Gong, Nanxu, et al.
Pubblicazione: (2026)
ThinkPilot: Steering Reasoning Models via Automated Think-prefixes Optimization
di: Li, Sunzhu, et al.
Pubblicazione: (2025)
di: Li, Sunzhu, et al.
Pubblicazione: (2025)
Seed1.5-Thinking: Advancing Superb Reasoning Models with Reinforcement Learning
di: Seed, ByteDance, et al.
Pubblicazione: (2025)
di: Seed, ByteDance, et al.
Pubblicazione: (2025)
InCoder-32B-Thinking: Industrial Code World Model for Thinking
di: Yang, Jian, et al.
Pubblicazione: (2026)
di: Yang, Jian, et al.
Pubblicazione: (2026)
Thinking in a Crowd: How Auxiliary Information Shapes LLM Reasoning
di: Zhao, Haodong, et al.
Pubblicazione: (2025)
di: Zhao, Haodong, et al.
Pubblicazione: (2025)
Advancing Speech Language Models by Scaling Supervised Fine-Tuning with Over 60,000 Hours of Synthetic Speech Dialogue Data
di: Zhao, Shuaijiang, et al.
Pubblicazione: (2024)
di: Zhao, Shuaijiang, et al.
Pubblicazione: (2024)
Efficient Reasoning with Balanced Thinking
di: Li, Yulin, et al.
Pubblicazione: (2026)
di: Li, Yulin, et al.
Pubblicazione: (2026)
When to Continue Thinking: Adaptive Thinking Mode Switching for Efficient Reasoning
di: Zhang, Xiaoyun, et al.
Pubblicazione: (2025)
di: Zhang, Xiaoyun, et al.
Pubblicazione: (2025)
Think Dense, Not Long: Dynamic Decoupled Conditional Advantage for Efficient Reasoning
di: Peng, Keqin, et al.
Pubblicazione: (2026)
di: Peng, Keqin, et al.
Pubblicazione: (2026)
LexPro-1.0 Technical Report
di: Chen, Haotian, et al.
Pubblicazione: (2025)
di: Chen, Haotian, et al.
Pubblicazione: (2025)
Think More, Hallucinate Less: Mitigating Hallucinations via Dual Process of Fast and Slow Thinking
di: Cheng, Xiaoxue, et al.
Pubblicazione: (2025)
di: Cheng, Xiaoxue, et al.
Pubblicazione: (2025)
Table-R1: Inference-Time Scaling for Table Reasoning
di: Yang, Zheyuan, et al.
Pubblicazione: (2025)
di: Yang, Zheyuan, et al.
Pubblicazione: (2025)
Thinking Economically: A Hierarchical Framework for Adaptive-Complexity Reasoning in LLMs
di: Gao, Yubo, et al.
Pubblicazione: (2026)
di: Gao, Yubo, et al.
Pubblicazione: (2026)
MM-THEBench: Do Reasoning MLLMs Think Reasonably?
di: Huang, Zhidian, et al.
Pubblicazione: (2026)
di: Huang, Zhidian, et al.
Pubblicazione: (2026)
A Survey of Frontiers in LLM Reasoning: Inference Scaling, Learning to Reason, and Agentic Systems
di: Ke, Zixuan, et al.
Pubblicazione: (2025)
di: Ke, Zixuan, et al.
Pubblicazione: (2025)
Thinking with Nothinking Calibration: A New In-Context Learning Paradigm in Reasoning Large Language Models
di: Wu, Haotian, et al.
Pubblicazione: (2025)
di: Wu, Haotian, et al.
Pubblicazione: (2025)
SciResearcher: Scaling Deep Research Agents for Frontier Scientific Reasoning
di: Zheng, Tianshi, et al.
Pubblicazione: (2026)
di: Zheng, Tianshi, et al.
Pubblicazione: (2026)
Atomic Thinking of LLMs: Decoupling and Exploring Mathematical Reasoning Abilities
di: Kuang, Jiayi, et al.
Pubblicazione: (2025)
di: Kuang, Jiayi, et al.
Pubblicazione: (2025)
StreamingThinker: Large Language Models Can Think While Reading
di: Tong, Junlong, et al.
Pubblicazione: (2025)
di: Tong, Junlong, et al.
Pubblicazione: (2025)
AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling
di: Liu, Zihan, et al.
Pubblicazione: (2024)
di: Liu, Zihan, et al.
Pubblicazione: (2024)
Skip-Thinking: Chunk-wise Chain-of-Thought Distillation Enable Smaller Language Models to Reason Better and Faster
di: Chen, Xiao, et al.
Pubblicazione: (2025)
di: Chen, Xiao, et al.
Pubblicazione: (2025)
Thinking Slow, Fast: Scaling Inference Compute with Distilled Reasoners
di: Paliotta, Daniele, et al.
Pubblicazione: (2025)
di: Paliotta, Daniele, et al.
Pubblicazione: (2025)
Efficient Reasoning with Hidden Thinking
di: Shen, Xuan, et al.
Pubblicazione: (2025)
di: Shen, Xuan, et al.
Pubblicazione: (2025)
DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
di: DeepSeek-AI, et al.
Pubblicazione: (2025)
di: DeepSeek-AI, et al.
Pubblicazione: (2025)
MatryoshkaThinking: Recursive Test-Time Scaling Enables Efficient Reasoning
di: Chen, Hongwei, et al.
Pubblicazione: (2025)
di: Chen, Hongwei, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Think Twice: Enhancing LLM Reasoning by Scaling Multi-round Test-time Thinking
di: Tian, Xiaoyu, et al.
Pubblicazione: (2025) -
DeepDistill: Enhancing LLM Reasoning Capabilities via Large-Scale Difficulty-Graded Data Training
di: Tian, Xiaoyu, et al.
Pubblicazione: (2025) -
Leveraging Reasoning Model Answers to Enhance Non-Reasoning Model Capability
di: Wang, Haotian, et al.
Pubblicazione: (2025) -
1.4 Million Open-Source Distilled Reasoning Dataset to Empower Large Language Model Training
di: Zhao, Han, et al.
Pubblicazione: (2025) -
Exploring the Potential of Offline RL for Reasoning in LLMs: A Preliminary Study
di: Tian, Xiaoyu, et al.
Pubblicazione: (2025)