Gespeichert in:
| Hauptverfasser: | Zhang, Yifan, Du, Wenyu, Jin, Dongming, Fu, Jie, Jin, Zhi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2502.20129 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Weights to Code: Extracting Interpretable Algorithms from the Discrete Transformer
von: Zhang, Yifan, et al.
Veröffentlicht: (2026)
von: Zhang, Yifan, et al.
Veröffentlicht: (2026)
Iteration Head: A Mechanistic Study of Chain-of-Thought
von: Cabannes, Vivien, et al.
Veröffentlicht: (2024)
von: Cabannes, Vivien, et al.
Veröffentlicht: (2024)
Evolving Demonstration Optimization for Chain-of-Thought Feature Transformation
von: Wang, Xinyuan, et al.
Veröffentlicht: (2026)
von: Wang, Xinyuan, et al.
Veröffentlicht: (2026)
Learning from Failures in Multi-Attempt Reinforcement Learning
von: Chung, Stephen, et al.
Veröffentlicht: (2025)
von: Chung, Stephen, et al.
Veröffentlicht: (2025)
Chain of Preference Optimization: Improving Chain-of-Thought Reasoning in LLMs
von: Zhang, Xuan, et al.
Veröffentlicht: (2024)
von: Zhang, Xuan, et al.
Veröffentlicht: (2024)
When Chain-of-Thought Fails, the Solution Hides in the Hidden States
von: Mehrafarin, Houman, et al.
Veröffentlicht: (2026)
von: Mehrafarin, Houman, et al.
Veröffentlicht: (2026)
Dissecting Long-Chain-of-Thought Reasoning Models: An Empirical Study
von: Mu, Yongyu, et al.
Veröffentlicht: (2025)
von: Mu, Yongyu, et al.
Veröffentlicht: (2025)
Rethinking Regularization Methods for Knowledge Graph Completion
von: Li, Linyu, et al.
Veröffentlicht: (2025)
von: Li, Linyu, et al.
Veröffentlicht: (2025)
IntentCoding: Amplifying User Intent in Code Generation
von: Fang, Zheng, et al.
Veröffentlicht: (2026)
von: Fang, Zheng, et al.
Veröffentlicht: (2026)
The Expressive Power of Transformers with Chain of Thought
von: Merrill, William, et al.
Veröffentlicht: (2023)
von: Merrill, William, et al.
Veröffentlicht: (2023)
ExpThink: Experience-Guided Reinforcement Learning for Adaptive Chain-of-Thought Compression
von: Bian, Tingcheng, et al.
Veröffentlicht: (2026)
von: Bian, Tingcheng, et al.
Veröffentlicht: (2026)
Learning to Evolve: Bayesian-Guided Continual Knowledge Graph Embedding
von: Li, Linyu, et al.
Veröffentlicht: (2025)
von: Li, Linyu, et al.
Veröffentlicht: (2025)
Tracking the Feature Dynamics in LLM Training: A Mechanistic Study
von: Xu, Yang, et al.
Veröffentlicht: (2024)
von: Xu, Yang, et al.
Veröffentlicht: (2024)
Compositional Reasoning with Transformers, RNNs, and Chain of Thought
von: Yehudai, Gilad, et al.
Veröffentlicht: (2025)
von: Yehudai, Gilad, et al.
Veröffentlicht: (2025)
Training Nonlinear Transformers for Chain-of-Thought Inference: A Theoretical Generalization Analysis
von: Li, Hongkang, et al.
Veröffentlicht: (2024)
von: Li, Hongkang, et al.
Veröffentlicht: (2024)
Value-Guided Search for Efficient Chain-of-Thought Reasoning
von: Wang, Kaiwen, et al.
Veröffentlicht: (2025)
von: Wang, Kaiwen, et al.
Veröffentlicht: (2025)
Chain of Thought Explanation for Dialogue State Tracking
von: Xu, Lin, et al.
Veröffentlicht: (2024)
von: Xu, Lin, et al.
Veröffentlicht: (2024)
Latent Chain-of-Thought? Decoding the Depth-Recurrent Transformer
von: Lu, Wenquan, et al.
Veröffentlicht: (2025)
von: Lu, Wenquan, et al.
Veröffentlicht: (2025)
From Sparse Dependence to Sparse Attention: Unveiling How Chain-of-Thought Enhances Transformer Sample Efficiency
von: Wen, Kaiyue, et al.
Veröffentlicht: (2024)
von: Wen, Kaiyue, et al.
Veröffentlicht: (2024)
The Expressive Power of Low Precision Softmax Transformers with (Summarized) Chain-of-Thought
von: Brösamle, Moritz, et al.
Veröffentlicht: (2026)
von: Brösamle, Moritz, et al.
Veröffentlicht: (2026)
Feature Extraction and Steering for Enhanced Chain-of-Thought Reasoning in Language Models
von: Li, Zihao, et al.
Veröffentlicht: (2025)
von: Li, Zihao, et al.
Veröffentlicht: (2025)
Stepwise Penalization for Length-Efficient Chain-of-Thought Reasoning
von: Li, Xintong, et al.
Veröffentlicht: (2026)
von: Li, Xintong, et al.
Veröffentlicht: (2026)
Automata Extraction from Transformers
von: Zhang, Yihao, et al.
Veröffentlicht: (2024)
von: Zhang, Yihao, et al.
Veröffentlicht: (2024)
Graph Chain-of-Thought: Augmenting Large Language Models by Reasoning on Graphs
von: Jin, Bowen, et al.
Veröffentlicht: (2024)
von: Jin, Bowen, et al.
Veröffentlicht: (2024)
Constraint-Rectified Training for Efficient Chain-of-Thought
von: Wu, Qinhang, et al.
Veröffentlicht: (2026)
von: Wu, Qinhang, et al.
Veröffentlicht: (2026)
Optimizing Chain-of-Thought Reasoners via Gradient Variance Minimization in Rejection Sampling and RL
von: Yao, Jiarui, et al.
Veröffentlicht: (2025)
von: Yao, Jiarui, et al.
Veröffentlicht: (2025)
Exploring Chain-of-Thought Reasoning for Steerable Pluralistic Alignment
von: Zhang, Yunfan, et al.
Veröffentlicht: (2025)
von: Zhang, Yunfan, et al.
Veröffentlicht: (2025)
OPV: Outcome-based Process Verifier for Efficient Long Chain-of-Thought Verification
von: Wu, Zijian, et al.
Veröffentlicht: (2025)
von: Wu, Zijian, et al.
Veröffentlicht: (2025)
On the Diagram of Thought
von: Zhang, Yifan, et al.
Veröffentlicht: (2024)
von: Zhang, Yifan, et al.
Veröffentlicht: (2024)
A Formal Comparison Between Chain of Thought and Latent Thought
von: Xu, Kevin, et al.
Veröffentlicht: (2025)
von: Xu, Kevin, et al.
Veröffentlicht: (2025)
When More is Less: Understanding Chain-of-Thought Length in LLMs
von: Wu, Yuyang, et al.
Veröffentlicht: (2025)
von: Wu, Yuyang, et al.
Veröffentlicht: (2025)
Demystifying Long Chain-of-Thought Reasoning in LLMs
von: Yeo, Edward, et al.
Veröffentlicht: (2025)
von: Yeo, Edward, et al.
Veröffentlicht: (2025)
Understanding Hidden Computations in Chain-of-Thought Reasoning
von: Bharadwaj, Aryasomayajula Ram
Veröffentlicht: (2024)
von: Bharadwaj, Aryasomayajula Ram
Veröffentlicht: (2024)
Reliable Chain-of-Thought via Prefix Consistency
von: Iwase, Naoto, et al.
Veröffentlicht: (2026)
von: Iwase, Naoto, et al.
Veröffentlicht: (2026)
CoT-UQ: Improving Response-wise Uncertainty Quantification in LLMs with Chain-of-Thought
von: Zhang, Boxuan, et al.
Veröffentlicht: (2025)
von: Zhang, Boxuan, et al.
Veröffentlicht: (2025)
Is Chain-of-Thought Really Not Explainability? Chain-of-Thought Can Be Faithful without Hint Verbalization
von: Zaman, Kerem, et al.
Veröffentlicht: (2025)
von: Zaman, Kerem, et al.
Veröffentlicht: (2025)
Tracking Equivalent Mechanistic Interpretations Across Neural Networks
von: Sun, Alan, et al.
Veröffentlicht: (2026)
von: Sun, Alan, et al.
Veröffentlicht: (2026)
The Role of Logic and Automata in Understanding Transformers
von: Lin, Anthony W., et al.
Veröffentlicht: (2025)
von: Lin, Anthony W., et al.
Veröffentlicht: (2025)
Multilingual OCR-Aware Fine-Tuning and Prompt-Guided Chain-of-Thought Reasoning for Multimodal Large Language Models
von: Xu, Qinwu, et al.
Veröffentlicht: (2026)
von: Xu, Qinwu, et al.
Veröffentlicht: (2026)
Fractured Chain-of-Thought Reasoning
von: Liao, Baohao, et al.
Veröffentlicht: (2025)
von: Liao, Baohao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Weights to Code: Extracting Interpretable Algorithms from the Discrete Transformer
von: Zhang, Yifan, et al.
Veröffentlicht: (2026) -
Iteration Head: A Mechanistic Study of Chain-of-Thought
von: Cabannes, Vivien, et al.
Veröffentlicht: (2024) -
Evolving Demonstration Optimization for Chain-of-Thought Feature Transformation
von: Wang, Xinyuan, et al.
Veröffentlicht: (2026) -
Learning from Failures in Multi-Attempt Reinforcement Learning
von: Chung, Stephen, et al.
Veröffentlicht: (2025) -
Chain of Preference Optimization: Improving Chain-of-Thought Reasoning in LLMs
von: Zhang, Xuan, et al.
Veröffentlicht: (2024)