Accordion-Thinking: Self-Regulated Step Summaries for Efficient and Readable LLM Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Zhicheng, Guo, Zhijiang, Huang, Yinya, Wang, Yongxin, Shi, Wenlei, Wang, Yiwei, Liang, Xiaodan, Tang, Jing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Critique to Verify: Accurate and Honest Test-Time Scaling with RL-Trained Verifiers
von: Yang, Zhicheng, et al.
Veröffentlicht: (2025)
von: Yang, Zhicheng, et al.
Veröffentlicht: (2025)
Depth-Breadth Synergy in RLVR: Unlocking LLM Reasoning Gains with Adaptive Exploration
von: Yang, Zhicheng, et al.
Veröffentlicht: (2025)
von: Yang, Zhicheng, et al.
Veröffentlicht: (2025)
TreeRPO: Tree Relative Policy Optimization
von: Yang, Zhicheng, et al.
Veröffentlicht: (2025)
von: Yang, Zhicheng, et al.
Veröffentlicht: (2025)
Prune-OPD: Efficient and Reliable On-Policy Distillation for Long-Horizon Reasoning
von: Yang, Zhicheng, et al.
Veröffentlicht: (2026)
von: Yang, Zhicheng, et al.
Veröffentlicht: (2026)
OptiBench Meets ReSocratic: Measure and Improve LLMs for Optimization Modeling
von: Yang, Zhicheng, et al.
Veröffentlicht: (2024)
von: Yang, Zhicheng, et al.
Veröffentlicht: (2024)
AlignedCoT: Prompting Large Language Models via Native-Speaking Demonstrations
von: Yang, Zhicheng, et al.
Veröffentlicht: (2023)
von: Yang, Zhicheng, et al.
Veröffentlicht: (2023)
AtomThink: Multimodal Slow Thinking with Atomic Step Reasoning
von: Xiang, Kun, et al.
Veröffentlicht: (2024)
von: Xiang, Kun, et al.
Veröffentlicht: (2024)
DeReason: A Difficulty-Aware Curriculum Improves Decoupled SFT-then-RL Training for General Reasoning
von: Hu, Hanxu, et al.
Veröffentlicht: (2026)
von: Hu, Hanxu, et al.
Veröffentlicht: (2026)
FormalAlign: Automated Alignment Evaluation for Autoformalization
von: Lu, Jianqiao, et al.
Veröffentlicht: (2024)
von: Lu, Jianqiao, et al.
Veröffentlicht: (2024)
ATG: Benchmarking Automated Theorem Generation for Generative Language Models
von: Lin, Xiaohan, et al.
Veröffentlicht: (2024)
von: Lin, Xiaohan, et al.
Veröffentlicht: (2024)
MetaCrit: A Critical Thinking Framework for Self-Regulated LLM Reasoning
von: Hou, Xinmeng, et al.
Veröffentlicht: (2025)
von: Hou, Xinmeng, et al.
Veröffentlicht: (2025)
ORMind: A Cognitive-Inspired End-to-End Reasoning Framework for Operations Research
von: Wang, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Wang, Zhiyuan, et al.
Veröffentlicht: (2025)
CARE What Fails: Contrastive Anchored-REflection for Verifiable Multimodal Reasoning
von: Wang, Yongxin, et al.
Veröffentlicht: (2025)
von: Wang, Yongxin, et al.
Veröffentlicht: (2025)
Proving Theorems Recursively
von: Wang, Haiming, et al.
Veröffentlicht: (2024)
von: Wang, Haiming, et al.
Veröffentlicht: (2024)
Process-Driven Autoformalization in Lean 4
von: Lu, Jianqiao, et al.
Veröffentlicht: (2024)
von: Lu, Jianqiao, et al.
Veröffentlicht: (2024)
ViewFusion: Structured Spatial Thinking Chains for Multi-View Reasoning
von: Tao, Xingjian, et al.
Veröffentlicht: (2026)
von: Tao, Xingjian, et al.
Veröffentlicht: (2026)
Uncovering Hidden Correctness in LLM Causal Reasoning via Symbolic Verification
von: He, Paul, et al.
Veröffentlicht: (2026)
von: He, Paul, et al.
Veröffentlicht: (2026)
Thinking with Drafts: Speculative Temporal Reasoning for Efficient Long Video Understanding
von: Hu, Pengfei, et al.
Veröffentlicht: (2025)
von: Hu, Pengfei, et al.
Veröffentlicht: (2025)
Thinking with Geometry: Active Geometry Integration for Spatial Reasoning
von: Li, Haoyuan, et al.
Veröffentlicht: (2026)
von: Li, Haoyuan, et al.
Veröffentlicht: (2026)
CLOMO: Counterfactual Logical Modification with Large Language Models
von: Huang, Yinya, et al.
Veröffentlicht: (2023)
von: Huang, Yinya, et al.
Veröffentlicht: (2023)
Efficient Reasoning with Hidden Thinking
von: Shen, Xuan, et al.
Veröffentlicht: (2025)
von: Shen, Xuan, et al.
Veröffentlicht: (2025)
DQ-LoRe: Dual Queries with Low Rank Approximation Re-ranking for In-Context Learning
von: Xiong, Jing, et al.
Veröffentlicht: (2023)
von: Xiong, Jing, et al.
Veröffentlicht: (2023)
SeePhys: Does Seeing Help Thinking? -- Benchmarking Vision-Based Physics Reasoning
von: Xiang, Kun, et al.
Veröffentlicht: (2025)
von: Xiang, Kun, et al.
Veröffentlicht: (2025)
Read Before You Think: Mitigating LLM Comprehension Failures with Step-by-Step Reading
von: Han, Feijiang, et al.
Veröffentlicht: (2025)
von: Han, Feijiang, et al.
Veröffentlicht: (2025)
Understanding GUI Agent Localization Biases through Logit Sharpness
von: Tao, Xingjian, et al.
Veröffentlicht: (2025)
von: Tao, Xingjian, et al.
Veröffentlicht: (2025)
Are LLMs Really Not Knowledgeable? Mining the Submerged Knowledge in LLMs' Memory
von: Tao, Xingjian, et al.
Veröffentlicht: (2024)
von: Tao, Xingjian, et al.
Veröffentlicht: (2024)
Graph-Augmented Reasoning: Evolving Step-by-Step Knowledge Graph Retrieval for LLM Reasoning
von: Wu, Wenjie, et al.
Veröffentlicht: (2025)
von: Wu, Wenjie, et al.
Veröffentlicht: (2025)
TL;DR: Too Long, Do Re-weighting for Efficient LLM Reasoning Compression
von: Li, Zhong-Zhi, et al.
Veröffentlicht: (2025)
von: Li, Zhong-Zhi, et al.
Veröffentlicht: (2025)
Thinking Short and Right Over Thinking Long: Serving LLM Reasoning Efficiently and Accurately
von: Wang, Yuhang, et al.
Veröffentlicht: (2025)
von: Wang, Yuhang, et al.
Veröffentlicht: (2025)
An Empirical Study of Reasoning Steps in Thinking Code LLMs
von: Xue, Haoran, et al.
Veröffentlicht: (2025)
von: Xue, Haoran, et al.
Veröffentlicht: (2025)
R2IF: Aligning Reasoning with Decisions via Composite Rewards for Interpretable LLM Function Calling
von: Cheng, Aijia, et al.
Veröffentlicht: (2026)
von: Cheng, Aijia, et al.
Veröffentlicht: (2026)
When Inverse Data Outperforms: Exploring the Pitfalls of Mixed Data in Multi-Stage Fine-Tuning
von: Deng, Mengyi, et al.
Veröffentlicht: (2025)
von: Deng, Mengyi, et al.
Veröffentlicht: (2025)
Token-Efficient Prompt Injection Attack: Provoking Cessation in LLM Reasoning via Adaptive Token Compression
von: Cui, Yu, et al.
Veröffentlicht: (2025)
von: Cui, Yu, et al.
Veröffentlicht: (2025)
FVEL: Interactive Formal Verification Environment with Large Language Models via Theorem Proving
von: Lin, Xiaohan, et al.
Veröffentlicht: (2024)
von: Lin, Xiaohan, et al.
Veröffentlicht: (2024)
Do Latent-CoT Models Think Step-by-Step? A Mechanistic Study on Sequential Reasoning Tasks
von: Liang, Jia, et al.
Veröffentlicht: (2026)
von: Liang, Jia, et al.
Veröffentlicht: (2026)
Learning From Correctness Without Prompting Makes LLM Efficient Reasoner
von: Yao, Yuxuan, et al.
Veröffentlicht: (2024)
von: Yao, Yuxuan, et al.
Veröffentlicht: (2024)
Step-wise Rubric Rewards for LLM Reasoning
von: Xie, Weichu, et al.
Veröffentlicht: (2026)
von: Xie, Weichu, et al.
Veröffentlicht: (2026)
SPC: Evolving Self-Play Critic via Adversarial Games for LLM Reasoning
von: Chen, Jiaqi, et al.
Veröffentlicht: (2025)
von: Chen, Jiaqi, et al.
Veröffentlicht: (2025)
Process or Result? Manipulated Ending Tokens Can Mislead Reasoning LLMs to Ignore the Correct Reasoning Steps
von: Cui, Yu, et al.
Veröffentlicht: (2025)
von: Cui, Yu, et al.
Veröffentlicht: (2025)
When Can Large Reasoning Models Save Thinking? Mechanistic Analysis of Behavioral Divergence in Reasoning
von: Zhu, Rongzhi, et al.
Veröffentlicht: (2025)
von: Zhu, Rongzhi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Critique to Verify: Accurate and Honest Test-Time Scaling with RL-Trained Verifiers
von: Yang, Zhicheng, et al.
Veröffentlicht: (2025) -
Depth-Breadth Synergy in RLVR: Unlocking LLM Reasoning Gains with Adaptive Exploration
von: Yang, Zhicheng, et al.
Veröffentlicht: (2025) -
TreeRPO: Tree Relative Policy Optimization
von: Yang, Zhicheng, et al.
Veröffentlicht: (2025) -
Prune-OPD: Efficient and Reliable On-Policy Distillation for Long-Horizon Reasoning
von: Yang, Zhicheng, et al.
Veröffentlicht: (2026) -
OptiBench Meets ReSocratic: Measure and Improve LLMs for Optimization Modeling
von: Yang, Zhicheng, et al.
Veröffentlicht: (2024)