Gespeichert in:
| Hauptverfasser: | Peng, Keqin, Ouyang, Yuanxin, Liu, Xuebo, Tian, Zhiliang, Han, Ruijian, Yuan, Yancheng, Ding, Liang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.02099 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Revisiting Demonstration Selection Strategies in In-Context Learning
von: Peng, Keqin, et al.
Veröffentlicht: (2024)
von: Peng, Keqin, et al.
Veröffentlicht: (2024)
Enhancing Input-Label Mapping in In-Context Learning with Contrastive Decoding
von: Peng, Keqin, et al.
Veröffentlicht: (2025)
von: Peng, Keqin, et al.
Veröffentlicht: (2025)
Revisiting Overthinking in Long Chain-of-Thought from the Perspective of Self-Doubt
von: Peng, Keqin, et al.
Veröffentlicht: (2025)
von: Peng, Keqin, et al.
Veröffentlicht: (2025)
Read Quietly, Think Aloud: Decoupling Comprehension and Reasoning in LLMs
von: Wang, Yuanxin, et al.
Veröffentlicht: (2025)
von: Wang, Yuanxin, et al.
Veröffentlicht: (2025)
CoRT: Code-integrated Reasoning within Thinking
von: Li, Chengpeng, et al.
Veröffentlicht: (2025)
von: Li, Chengpeng, et al.
Veröffentlicht: (2025)
VAE-Inf: A statistically interpretable generative paradigm for imbalanced classification
von: Wu, Hongfei, et al.
Veröffentlicht: (2026)
von: Wu, Hongfei, et al.
Veröffentlicht: (2026)
Efficient Reasoning with Balanced Thinking
von: Li, Yulin, et al.
Veröffentlicht: (2026)
von: Li, Yulin, et al.
Veröffentlicht: (2026)
Stabilizing Efficient Reasoning with Step-Level Advantage Selection
von: Wang, Han, et al.
Veröffentlicht: (2026)
von: Wang, Han, et al.
Veröffentlicht: (2026)
A Survey on Large Language Model-based Agents for Statistics and Data Science
von: Sun, Maojun, et al.
Veröffentlicht: (2024)
von: Sun, Maojun, et al.
Veröffentlicht: (2024)
REA-RL: Reflection-Aware Online Reinforcement Learning for Efficient Reasoning
von: Deng, Hexuan, et al.
Veröffentlicht: (2025)
von: Deng, Hexuan, et al.
Veröffentlicht: (2025)
BRiTE: Bootstrapping Reinforced Thinking Process to Enhance Language Model Reasoning
von: Zhong, Han, et al.
Veröffentlicht: (2025)
von: Zhong, Han, et al.
Veröffentlicht: (2025)
Train Long, Think Short: Curriculum Learning for Efficient Reasoning
von: Hammoud, Hasan Abed Al Kader, et al.
Veröffentlicht: (2025)
von: Hammoud, Hasan Abed Al Kader, et al.
Veröffentlicht: (2025)
Dynamic Thinking-Token Selection for Efficient Reasoning in Large Reasoning Models
von: Guo, Zhenyuan, et al.
Veröffentlicht: (2026)
von: Guo, Zhenyuan, et al.
Veröffentlicht: (2026)
Longer Context, Deeper Thinking: Uncovering the Role of Long-Context Ability in Reasoning
von: Yang, Wang, et al.
Veröffentlicht: (2025)
von: Yang, Wang, et al.
Veröffentlicht: (2025)
Efficient Reasoning with Hidden Thinking
von: Shen, Xuan, et al.
Veröffentlicht: (2025)
von: Shen, Xuan, et al.
Veröffentlicht: (2025)
Arbitrage: Efficient Reasoning via Advantage-Aware Speculation
von: Maheswaran, Monishwaran, et al.
Veröffentlicht: (2025)
von: Maheswaran, Monishwaran, et al.
Veröffentlicht: (2025)
DB-LLM: Accurate Dual-Binarization for Efficient LLMs
von: Chen, Hong, et al.
Veröffentlicht: (2024)
von: Chen, Hong, et al.
Veröffentlicht: (2024)
LongFlow: Efficient KV Cache Compression for Reasoning Models
von: Su, Yi, et al.
Veröffentlicht: (2026)
von: Su, Yi, et al.
Veröffentlicht: (2026)
Chain of Execution Supervision Promotes General Reasoning in Large Language Models
von: Chen, Nuo, et al.
Veröffentlicht: (2025)
von: Chen, Nuo, et al.
Veröffentlicht: (2025)
Asymmetric Advantage Modulation Calibrates Entropy Dynamics in RLVR
von: Gu, Hengrui, et al.
Veröffentlicht: (2026)
von: Gu, Hengrui, et al.
Veröffentlicht: (2026)
Decoder-Hybrid-Decoder Architecture for Efficient Reasoning with Long Generation
von: Ren, Liliang, et al.
Veröffentlicht: (2025)
von: Ren, Liliang, et al.
Veröffentlicht: (2025)
ThinkRouter: Efficient Reasoning via Routing Thinking between Latent and Discrete Spaces
von: Xu, Xin, et al.
Veröffentlicht: (2026)
von: Xu, Xin, et al.
Veröffentlicht: (2026)
Free Energy-Driven Reinforcement Learning with Adaptive Advantage Shaping for Unsupervised Reasoning in LLMs
von: Huang, Yiming, et al.
Veröffentlicht: (2026)
von: Huang, Yiming, et al.
Veröffentlicht: (2026)
DELTA: Dynamic Layer-Aware Token Attention for Efficient Long-Context Reasoning
von: Zarch, Hossein Entezari, et al.
Veröffentlicht: (2025)
von: Zarch, Hossein Entezari, et al.
Veröffentlicht: (2025)
Every Attention Matters: An Efficient Hybrid Architecture for Long-Context Reasoning
von: Ling Team, et al.
Veröffentlicht: (2025)
von: Ling Team, et al.
Veröffentlicht: (2025)
AAPO: Enhancing the Reasoning Capabilities of LLMs with Advantage Margin
von: Xiong, Jian, et al.
Veröffentlicht: (2025)
von: Xiong, Jian, et al.
Veröffentlicht: (2025)
Thinking-Free Policy Initialization Makes Distilled Reasoning Models More Effective and Efficient Reasoners
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
DenseMamba: State Space Models with Dense Hidden Connection for Efficient Large Language Models
von: He, Wei, et al.
Veröffentlicht: (2024)
von: He, Wei, et al.
Veröffentlicht: (2024)
Stable Adaptive Thinking via Advantage Shaping and Length-Aware Gradient Regulation
von: Xu, Zihang, et al.
Veröffentlicht: (2026)
von: Xu, Zihang, et al.
Veröffentlicht: (2026)
DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs
von: Zhou, Xiabin, et al.
Veröffentlicht: (2024)
von: Zhou, Xiabin, et al.
Veröffentlicht: (2024)
Outcome-Grounded Advantage Reshaping for Fine-Grained Credit Assignment in Mathematical Reasoning
von: Li, Ziheng, et al.
Veröffentlicht: (2026)
von: Li, Ziheng, et al.
Veröffentlicht: (2026)
Multipole Attention for Efficient Long Context Reasoning
von: Hooper, Coleman, et al.
Veröffentlicht: (2025)
von: Hooper, Coleman, et al.
Veröffentlicht: (2025)
AutoL2S: Auto Long-Short Reasoning for Efficient Large Language Models
von: Luo, Feng, et al.
Veröffentlicht: (2025)
von: Luo, Feng, et al.
Veröffentlicht: (2025)
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code
von: Bao, Keqin, et al.
Veröffentlicht: (2025)
von: Bao, Keqin, et al.
Veröffentlicht: (2025)
DSAEval: Evaluating Data Science Agents on a Wide Range of Real-World Data Science Problems
von: Sun, Maojun, et al.
Veröffentlicht: (2026)
von: Sun, Maojun, et al.
Veröffentlicht: (2026)
AgentDropout: Dynamic Agent Elimination for Token-Efficient and High-Performance LLM-Based Multi-Agent Collaboration
von: Wang, Zhexuan, et al.
Veröffentlicht: (2025)
von: Wang, Zhexuan, et al.
Veröffentlicht: (2025)
Atomic Thinking of LLMs: Decoupling and Exploring Mathematical Reasoning Abilities
von: Kuang, Jiayi, et al.
Veröffentlicht: (2025)
von: Kuang, Jiayi, et al.
Veröffentlicht: (2025)
ForesightKV: Optimizing KV Cache Eviction for Reasoning Models by Learning Long-Term Contribution
von: Dong, Zican, et al.
Veröffentlicht: (2026)
von: Dong, Zican, et al.
Veröffentlicht: (2026)
L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2025)
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2025)
Recursively Summarizing Enables Long-Term Dialogue Memory in Large Language Models
von: Wang, Qingyue, et al.
Veröffentlicht: (2023)
von: Wang, Qingyue, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Revisiting Demonstration Selection Strategies in In-Context Learning
von: Peng, Keqin, et al.
Veröffentlicht: (2024) -
Enhancing Input-Label Mapping in In-Context Learning with Contrastive Decoding
von: Peng, Keqin, et al.
Veröffentlicht: (2025) -
Revisiting Overthinking in Long Chain-of-Thought from the Perspective of Self-Doubt
von: Peng, Keqin, et al.
Veröffentlicht: (2025) -
Read Quietly, Think Aloud: Decoupling Comprehension and Reasoning in LLMs
von: Wang, Yuanxin, et al.
Veröffentlicht: (2025) -
CoRT: Code-integrated Reasoning within Thinking
von: Li, Chengpeng, et al.
Veröffentlicht: (2025)