Chain of Execution Supervision Promotes General Reasoning in Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Nuo, Li, Zehua, Bao, Keqin, Lin, Junyang, Liu, Dayiheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CoRT: Code-integrated Reasoning within Thinking
von: Li, Chengpeng, et al.
Veröffentlicht: (2025)
von: Li, Chengpeng, et al.
Veröffentlicht: (2025)
PassNet: Scaling Large Language Models for Graph Compiler Pass Generation
von: Liu, Yiqun, et al.
Veröffentlicht: (2026)
von: Liu, Yiqun, et al.
Veröffentlicht: (2026)
Self-Evolving Critique Abilities in Large Language Models
von: Tang, Zhengyang, et al.
Veröffentlicht: (2025)
von: Tang, Zhengyang, et al.
Veröffentlicht: (2025)
The Lessons of Developing Process Reward Models in Mathematical Reasoning
von: Zhang, Zhenru, et al.
Veröffentlicht: (2025)
von: Zhang, Zhenru, et al.
Veröffentlicht: (2025)
ProcessBench: Identifying Process Errors in Mathematical Reasoning
von: Zheng, Chujie, et al.
Veröffentlicht: (2024)
von: Zheng, Chujie, et al.
Veröffentlicht: (2024)
ControlMath: Controllable Data Generation Promotes Math Generalist Models
von: Chen, Nuo, et al.
Veröffentlicht: (2024)
von: Chen, Nuo, et al.
Veröffentlicht: (2024)
From Large to Small: Transferring CUDA Optimization Expertise via Reasoning Graph
von: Gong, Junfeng, et al.
Veröffentlicht: (2025)
von: Gong, Junfeng, et al.
Veröffentlicht: (2025)
How Abilities in Large Language Models are Affected by Supervised Fine-tuning Data Composition
von: Dong, Guanting, et al.
Veröffentlicht: (2023)
von: Dong, Guanting, et al.
Veröffentlicht: (2023)
Teaching Language Models to Reason with Tools
von: Li, Chengpeng, et al.
Veröffentlicht: (2025)
von: Li, Chengpeng, et al.
Veröffentlicht: (2025)
Code Simulation Challenges for Large Language Models
von: La Malfa, Emanuele, et al.
Veröffentlicht: (2024)
von: La Malfa, Emanuele, et al.
Veröffentlicht: (2024)
RealCritic: Towards Effectiveness-Driven Evaluation of Language Model Critiques
von: Tang, Zhengyang, et al.
Veröffentlicht: (2025)
von: Tang, Zhengyang, et al.
Veröffentlicht: (2025)
Synthesizing Programmatic Reinforcement Learning Policies with Large Language Model Guided Search
von: Liu, Max, et al.
Veröffentlicht: (2024)
von: Liu, Max, et al.
Veröffentlicht: (2024)
Rewarding Graph Reasoning Process makes LLMs more Generalized Reasoners
von: Peng, Miao, et al.
Veröffentlicht: (2025)
von: Peng, Miao, et al.
Veröffentlicht: (2025)
EquiBench: Benchmarking Large Language Models' Reasoning about Program Semantics via Equivalence Checking
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
Flaming-hot Initiation with Regular Execution Sampling for Large Language Models
von: Chen, Weizhe, et al.
Veröffentlicht: (2024)
von: Chen, Weizhe, et al.
Veröffentlicht: (2024)
Program Semantic Inequivalence Game with Large Language Models
von: Miceli-Barone, Antonio Valerio, et al.
Veröffentlicht: (2025)
von: Miceli-Barone, Antonio Valerio, et al.
Veröffentlicht: (2025)
EnCompass: Enhancing Agent Programming with Search Over Program Execution Paths
von: Li, Zhening, et al.
Veröffentlicht: (2025)
von: Li, Zhening, et al.
Veröffentlicht: (2025)
Explore-Execute Chain: Towards an Efficient Structured Reasoning Paradigm
von: Yang, Kaisen, et al.
Veröffentlicht: (2025)
von: Yang, Kaisen, et al.
Veröffentlicht: (2025)
APPL: A Prompt Programming Language for Harmonious Integration of Programs and Large Language Model Prompts
von: Dong, Honghua, et al.
Veröffentlicht: (2024)
von: Dong, Honghua, et al.
Veröffentlicht: (2024)
NPG-Muse: Scaling Long Chain-of-Thought Reasoning with NP-Hard Graph Problems
von: Wang, Yuyao, et al.
Veröffentlicht: (2025)
von: Wang, Yuyao, et al.
Veröffentlicht: (2025)
Stepwise Perplexity-Guided Refinement for Efficient Chain-of-Thought Reasoning in Large Language Models
von: Cui, Yingqian, et al.
Veröffentlicht: (2025)
von: Cui, Yingqian, et al.
Veröffentlicht: (2025)
From Reasoning to Code: GRPO Optimization for Underrepresented Languages
von: Pennino, Federico, et al.
Veröffentlicht: (2025)
von: Pennino, Federico, et al.
Veröffentlicht: (2025)
Diffusion of Thoughts: Chain-of-Thought Reasoning in Diffusion Language Models
von: Ye, Jiacheng, et al.
Veröffentlicht: (2024)
von: Ye, Jiacheng, et al.
Veröffentlicht: (2024)
JavaBench: A Benchmark of Object-Oriented Code Generation for Evaluating Large Language Models
von: Cao, Jialun, et al.
Veröffentlicht: (2024)
von: Cao, Jialun, et al.
Veröffentlicht: (2024)
Markovian Generation Chains in Large Language Models
von: Geng, Mingmeng, et al.
Veröffentlicht: (2026)
von: Geng, Mingmeng, et al.
Veröffentlicht: (2026)
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code
von: Bao, Keqin, et al.
Veröffentlicht: (2025)
von: Bao, Keqin, et al.
Veröffentlicht: (2025)
Large Language Models Synergize with Automated Machine Learning
von: Xu, Jinglue, et al.
Veröffentlicht: (2024)
von: Xu, Jinglue, et al.
Veröffentlicht: (2024)
Can LLMs Compress (and Decompress)? Evaluating Code Understanding and Execution via Invertibility
von: Maveli, Nickil, et al.
Veröffentlicht: (2026)
von: Maveli, Nickil, et al.
Veröffentlicht: (2026)
Generative Evaluation of Complex Reasoning in Large Language Models
von: Lin, Haowei, et al.
Veröffentlicht: (2025)
von: Lin, Haowei, et al.
Veröffentlicht: (2025)
A Multi-Expert Large Language Model Architecture for Verilog Code Generation
von: Nadimi, Bardia, et al.
Veröffentlicht: (2024)
von: Nadimi, Bardia, et al.
Veröffentlicht: (2024)
What I cannot execute, I do not understand: Training and Evaluating LLMs on Program Execution Traces
von: Armengol-Estapé, Jordi, et al.
Veröffentlicht: (2025)
von: Armengol-Estapé, Jordi, et al.
Veröffentlicht: (2025)
Enhancing Generalization in Chain of Thought Reasoning for Smaller Models
von: Yin, Maxwell J., et al.
Veröffentlicht: (2025)
von: Yin, Maxwell J., et al.
Veröffentlicht: (2025)
Large Language Models for Code Summarization
von: Szalontai, Balázs, et al.
Veröffentlicht: (2024)
von: Szalontai, Balázs, et al.
Veröffentlicht: (2024)
DotaMath: Decomposition of Thought with Code Assistance and Self-correction for Mathematical Reasoning
von: Li, Chengpeng, et al.
Veröffentlicht: (2024)
von: Li, Chengpeng, et al.
Veröffentlicht: (2024)
VeriReason: Reinforcement Learning with Testbench Feedback for Reasoning-Enhanced Verilog Generation
von: Wang, Yiting, et al.
Veröffentlicht: (2025)
von: Wang, Yiting, et al.
Veröffentlicht: (2025)
Toward Adaptive Reasoning in Large Language Models with Thought Rollback
von: Chen, Sijia, et al.
Veröffentlicht: (2024)
von: Chen, Sijia, et al.
Veröffentlicht: (2024)
Tilus: A Tile-Level GPGPU Programming Language for Low-Precision Computation
von: Ding, Yaoyao, et al.
Veröffentlicht: (2025)
von: Ding, Yaoyao, et al.
Veröffentlicht: (2025)
LiteCoOp: Lightweight Multi-LLM Shared-Tree Reasoning for Model-Serving Compiler Optimizations
von: Tang, Annabelle Sujun, et al.
Veröffentlicht: (2026)
von: Tang, Annabelle Sujun, et al.
Veröffentlicht: (2026)
Curriculum Learning for Small Code Language Models
von: Naïr, Marwa, et al.
Veröffentlicht: (2024)
von: Naïr, Marwa, et al.
Veröffentlicht: (2024)
Efficient Reasoning for Large Reasoning Language Models via Certainty-Guided Reflection Suppression
von: Huang, Jiameng, et al.
Veröffentlicht: (2025)
von: Huang, Jiameng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
CoRT: Code-integrated Reasoning within Thinking
von: Li, Chengpeng, et al.
Veröffentlicht: (2025) -
PassNet: Scaling Large Language Models for Graph Compiler Pass Generation
von: Liu, Yiqun, et al.
Veröffentlicht: (2026) -
Self-Evolving Critique Abilities in Large Language Models
von: Tang, Zhengyang, et al.
Veröffentlicht: (2025) -
The Lessons of Developing Process Reward Models in Mathematical Reasoning
von: Zhang, Zhenru, et al.
Veröffentlicht: (2025) -
ProcessBench: Identifying Process Errors in Mathematical Reasoning
von: Zheng, Chujie, et al.
Veröffentlicht: (2024)