Explore-Execute Chain: Towards an Efficient Structured Reasoning Paradigm
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Kaisen, He, Lixuan, Shah, Rushi, Yang, Kaicheng, Ma, Qinwei, Liu, Dianbo, Lamb, Alex |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Representation Collapsing Problems in Vector Quantization
von: Zhao, Wenhao, et al.
Veröffentlicht: (2024)
von: Zhao, Wenhao, et al.
Veröffentlicht: (2024)
Improving Discrete Optimisation Via Decoupled Straight-Through Estimator
von: Shah, Rushi, et al.
Veröffentlicht: (2024)
von: Shah, Rushi, et al.
Veröffentlicht: (2024)
Chain of Execution Supervision Promotes General Reasoning in Large Language Models
von: Chen, Nuo, et al.
Veröffentlicht: (2025)
von: Chen, Nuo, et al.
Veröffentlicht: (2025)
Efficient Joint Prediction of Multiple Future Tokens
von: Ahn, Kwangjun, et al.
Veröffentlicht: (2025)
von: Ahn, Kwangjun, et al.
Veröffentlicht: (2025)
Early Quantization Shrinks Codebook: A Simple Fix for Diversity-Preserving Tokenization
von: Zhao, Wenhao, et al.
Veröffentlicht: (2026)
von: Zhao, Wenhao, et al.
Veröffentlicht: (2026)
Towards Execution-Grounded Automated AI Research
von: Si, Chenglei, et al.
Veröffentlicht: (2026)
von: Si, Chenglei, et al.
Veröffentlicht: (2026)
Improving Sampling for Masked Diffusion Models via Information Gain
von: Yang, Kaisen, et al.
Veröffentlicht: (2026)
von: Yang, Kaisen, et al.
Veröffentlicht: (2026)
AMFT: Aligning LLM Reasoners by Meta-Learning the Optimal Imitation-Exploration Balance
von: He, Lixuan, et al.
Veröffentlicht: (2025)
von: He, Lixuan, et al.
Veröffentlicht: (2025)
Batched Contextual Reinforcement: A Task-Scaling Law for Efficient Reasoning
von: Yang, Bangji, et al.
Veröffentlicht: (2026)
von: Yang, Bangji, et al.
Veröffentlicht: (2026)
Stepwise Perplexity-Guided Refinement for Efficient Chain-of-Thought Reasoning in Large Language Models
von: Cui, Yingqian, et al.
Veröffentlicht: (2025)
von: Cui, Yingqian, et al.
Veröffentlicht: (2025)
Rep2Text: Decoding Full Text from a Single LLM Token Representation
von: Zhao, Haiyan, et al.
Veröffentlicht: (2025)
von: Zhao, Haiyan, et al.
Veröffentlicht: (2025)
Is Chain-of-Thought Reasoning of LLMs a Mirage? A Data Distribution Lens
von: Zhao, Chengshuai, et al.
Veröffentlicht: (2025)
von: Zhao, Chengshuai, et al.
Veröffentlicht: (2025)
Branch-and-Browse: Efficient and Controllable Web Exploration with Tree-Structured Reasoning and Action Memory
von: He, Shiqi, et al.
Veröffentlicht: (2025)
von: He, Shiqi, et al.
Veröffentlicht: (2025)
Internalizing Outcome Supervision into Process Supervision: A New Paradigm for Reinforcement Learning for Reasoning
von: Ding, Fei, et al.
Veröffentlicht: (2026)
von: Ding, Fei, et al.
Veröffentlicht: (2026)
Performance Asymmetry in Model-Based Reinforcement Learning
von: Lim, Jing Yu, et al.
Veröffentlicht: (2025)
von: Lim, Jing Yu, et al.
Veröffentlicht: (2025)
Value-Guided Search for Efficient Chain-of-Thought Reasoning
von: Wang, Kaiwen, et al.
Veröffentlicht: (2025)
von: Wang, Kaiwen, et al.
Veröffentlicht: (2025)
Stepwise Penalization for Length-Efficient Chain-of-Thought Reasoning
von: Li, Xintong, et al.
Veröffentlicht: (2026)
von: Li, Xintong, et al.
Veröffentlicht: (2026)
Attention Schema-based Attention Control (ASAC): A Cognitive-Inspired Approach for Attention Management in Transformers
von: Saxena, Krati, et al.
Veröffentlicht: (2025)
von: Saxena, Krati, et al.
Veröffentlicht: (2025)
Understanding Reasoning in Chain-of-Thought from the Hopfieldian View
von: Hu, Lijie, et al.
Veröffentlicht: (2024)
von: Hu, Lijie, et al.
Veröffentlicht: (2024)
Verifying Chain-of-Thought Reasoning via Its Computational Graph
von: Zhao, Zheng, et al.
Veröffentlicht: (2025)
von: Zhao, Zheng, et al.
Veröffentlicht: (2025)
Markov Chain of Thought for Efficient Mathematical Reasoning
von: Yang, Wen, et al.
Veröffentlicht: (2024)
von: Yang, Wen, et al.
Veröffentlicht: (2024)
BayesAgent: Bayesian Agentic Reasoning Under Uncertainty via Verbalized Probabilistic Graphical Modeling
von: Huang, Hengguan, et al.
Veröffentlicht: (2024)
von: Huang, Hengguan, et al.
Veröffentlicht: (2024)
Enhancing Auto-regressive Chain-of-Thought through Loop-Aligned Reasoning
von: Yu, Qifan, et al.
Veröffentlicht: (2025)
von: Yu, Qifan, et al.
Veröffentlicht: (2025)
From Reasoning Chains to Verifiable Subproblems: Curriculum Reinforcement Learning Enables Credit Assignment for LLM Reasoning
von: Jiang, Xitai, et al.
Veröffentlicht: (2026)
von: Jiang, Xitai, et al.
Veröffentlicht: (2026)
Think Consistently, Reason Efficiently: Energy-Based Calibration for Implicit Chain-of-Thought
von: Chen, Zhikang, et al.
Veröffentlicht: (2025)
von: Chen, Zhikang, et al.
Veröffentlicht: (2025)
Retro-Search: Exploring Untaken Paths for Deeper and Efficient Reasoning
von: Lu, Ximing, et al.
Veröffentlicht: (2025)
von: Lu, Ximing, et al.
Veröffentlicht: (2025)
From Local to Global: Revisiting Structured Pruning Paradigms for Large Language Models
von: Wang, Ziyan, et al.
Veröffentlicht: (2025)
von: Wang, Ziyan, et al.
Veröffentlicht: (2025)
A Theoretical Understanding of Chain-of-Thought: Coherent Reasoning and Error-Aware Demonstration
von: Cui, Yingqian, et al.
Veröffentlicht: (2024)
von: Cui, Yingqian, et al.
Veröffentlicht: (2024)
Towards Data-efficient Customer Intent Recognition with Prompt-based Learning Paradigm
von: Luo, Hengyu, et al.
Veröffentlicht: (2023)
von: Luo, Hengyu, et al.
Veröffentlicht: (2023)
Fractured Chain-of-Thought Reasoning
von: Liao, Baohao, et al.
Veröffentlicht: (2025)
von: Liao, Baohao, et al.
Veröffentlicht: (2025)
AutoLoRA: Automatically Tuning Matrix Ranks in Low-Rank Adaptation Based on Meta Learning
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2024)
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2024)
Reasoning Theater: Disentangling Model Beliefs from Chain-of-Thought
von: Boppana, Siddharth, et al.
Veröffentlicht: (2026)
von: Boppana, Siddharth, et al.
Veröffentlicht: (2026)
MetaState: Persistent Working Memory Enhances Reasoning in Discrete Diffusion Language Models
von: Xia, Kejing, et al.
Veröffentlicht: (2026)
von: Xia, Kejing, et al.
Veröffentlicht: (2026)
Search, Verify and Feedback: Towards Next Generation Post-training Paradigm of Foundation Models via Verifier Engineering
von: Guan, Xinyan, et al.
Veröffentlicht: (2024)
von: Guan, Xinyan, et al.
Veröffentlicht: (2024)
ETR: Entropy Trend Reward for Efficient Chain-of-Thought Reasoning
von: Xiong, Xuan, et al.
Veröffentlicht: (2026)
von: Xiong, Xuan, et al.
Veröffentlicht: (2026)
Beyond Either-Or Reasoning: Transduction and Induction as Cooperative Problem-Solving Paradigms
von: Zenkner, Janis, et al.
Veröffentlicht: (2025)
von: Zenkner, Janis, et al.
Veröffentlicht: (2025)
Exploring the Limitations of Mamba in COPY and CoT Reasoning
von: Ren, Ruifeng, et al.
Veröffentlicht: (2024)
von: Ren, Ruifeng, et al.
Veröffentlicht: (2024)
Learn to Reason Efficiently with Adaptive Length-based Reward Shaping
von: Liu, Wei, et al.
Veröffentlicht: (2025)
von: Liu, Wei, et al.
Veröffentlicht: (2025)
What makes Reasoning Models Different? Follow the Reasoning Leader for Efficient Decoding
von: Li, Ming, et al.
Veröffentlicht: (2025)
von: Li, Ming, et al.
Veröffentlicht: (2025)
Exploring the Reversal Curse and Other Deductive Logical Reasoning in BERT and GPT-Based Large Language Models
von: Wu, Da, et al.
Veröffentlicht: (2023)
von: Wu, Da, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Representation Collapsing Problems in Vector Quantization
von: Zhao, Wenhao, et al.
Veröffentlicht: (2024) -
Improving Discrete Optimisation Via Decoupled Straight-Through Estimator
von: Shah, Rushi, et al.
Veröffentlicht: (2024) -
Chain of Execution Supervision Promotes General Reasoning in Large Language Models
von: Chen, Nuo, et al.
Veröffentlicht: (2025) -
Efficient Joint Prediction of Multiple Future Tokens
von: Ahn, Kwangjun, et al.
Veröffentlicht: (2025) -
Early Quantization Shrinks Codebook: A Simple Fix for Diversity-Preserving Tokenization
von: Zhao, Wenhao, et al.
Veröffentlicht: (2026)