Tree-of-Code: A Hybrid Approach for Robust Complex Task Planning and Execution
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ni, Ziyi, Li, Yifan, Dong, Daxiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Tree-of-Code: A Tree-Structured Exploring Framework for End-to-End Code Generation and Execution in Complex Task Handling
von: Ni, Ziyi, et al.
Veröffentlicht: (2024)
von: Ni, Ziyi, et al.
Veröffentlicht: (2024)
SWE-Hub: A Unified Production System for Scalable, Executable Software Engineering Tasks
von: Zeng, Yucheng, et al.
Veröffentlicht: (2026)
von: Zeng, Yucheng, et al.
Veröffentlicht: (2026)
GitTaskBench: A Benchmark for Code Agents Solving Real-World Tasks Through Code Repository Leveraging
von: Ni, Ziyi, et al.
Veröffentlicht: (2025)
von: Ni, Ziyi, et al.
Veröffentlicht: (2025)
RA-Gen: A Controllable Code Generation Framework Using ReAct for Multi-Agent Task Execution
von: Liu, Aofan, et al.
Veröffentlicht: (2025)
von: Liu, Aofan, et al.
Veröffentlicht: (2025)
RedCode: Risky Code Execution and Generation Benchmark for Code Agents
von: Guo, Chengquan, et al.
Veröffentlicht: (2024)
von: Guo, Chengquan, et al.
Veröffentlicht: (2024)
RepoMaster: Autonomous Exploration and Understanding of GitHub Repositories for Complex Task Solving
von: Wang, Huacan, et al.
Veröffentlicht: (2025)
von: Wang, Huacan, et al.
Veröffentlicht: (2025)
Learning Adaptive Parallel Execution for Efficient Code Localization
von: Xu, Ke, et al.
Veröffentlicht: (2026)
von: Xu, Ke, et al.
Veröffentlicht: (2026)
FasterPy: An LLM-based Code Execution Efficiency Optimization Framework
von: Wu, Yue, et al.
Veröffentlicht: (2025)
von: Wu, Yue, et al.
Veröffentlicht: (2025)
An Execution-Verified Multi-Language Benchmark for Code Semantic Reasoning
von: Li, Yikun, et al.
Veröffentlicht: (2026)
von: Li, Yikun, et al.
Veröffentlicht: (2026)
Do Code Semantics Help? A Comprehensive Study on Execution Trace-Based Information for Code Large Language Models
von: Wang, Jian, et al.
Veröffentlicht: (2025)
von: Wang, Jian, et al.
Veröffentlicht: (2025)
Analyzing Chain of Thought (CoT) Approaches in Control Flow Code Deobfuscation Tasks
von: Mohseni, Seyedreza, et al.
Veröffentlicht: (2026)
von: Mohseni, Seyedreza, et al.
Veröffentlicht: (2026)
VisCoder: Fine-Tuning LLMs for Executable Python Visualization Code Generation
von: Ni, Yuansheng, et al.
Veröffentlicht: (2025)
von: Ni, Yuansheng, et al.
Veröffentlicht: (2025)
Integrating Symbolic Execution into the Fine-Tuning of Code-Generating LLMs
von: Sakharova, Marina, et al.
Veröffentlicht: (2025)
von: Sakharova, Marina, et al.
Veröffentlicht: (2025)
CodeRL+: Improving Code Generation via Reinforcement with Execution Semantics Alignment
von: Jiang, Xue, et al.
Veröffentlicht: (2025)
von: Jiang, Xue, et al.
Veröffentlicht: (2025)
Enhancing LLM-Based Code Generation with Complexity Metrics: A Feedback-Driven Approach
von: Sepidband, Melika, et al.
Veröffentlicht: (2025)
von: Sepidband, Melika, et al.
Veröffentlicht: (2025)
Executing as You Generate: Hiding Execution Latency in LLM Code Generation
von: Sun, Zhensu, et al.
Veröffentlicht: (2026)
von: Sun, Zhensu, et al.
Veröffentlicht: (2026)
GeoCode-GPT: A Large Language Model for Geospatial Code Generation Tasks
von: Hou, Shuyang, et al.
Veröffentlicht: (2024)
von: Hou, Shuyang, et al.
Veröffentlicht: (2024)
Constraint-Guided Multi-Agent Decompilation for Executable Binary Recovery
von: Zhang, Yifan, et al.
Veröffentlicht: (2026)
von: Zhang, Yifan, et al.
Veröffentlicht: (2026)
Toward Executable Repository-Level Code Generation via Environment Alignment
von: Pan, Ruwei, et al.
Veröffentlicht: (2026)
von: Pan, Ruwei, et al.
Veröffentlicht: (2026)
MutaGReP: Execution-Free Repository-Grounded Plan Search for Code-Use
von: Khan, Zaid, et al.
Veröffentlicht: (2025)
von: Khan, Zaid, et al.
Veröffentlicht: (2025)
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions
von: Larbi, Maya, et al.
Veröffentlicht: (2025)
von: Larbi, Maya, et al.
Veröffentlicht: (2025)
CASET: Complexity Analysis using Simple Execution Traces for CS* submissions
von: Mehta, Aaryen, et al.
Veröffentlicht: (2024)
von: Mehta, Aaryen, et al.
Veröffentlicht: (2024)
Benchmarking Multimodal LLMs on Code Generation for Complex Interactive Webpages
von: Wu, Fan, et al.
Veröffentlicht: (2026)
von: Wu, Fan, et al.
Veröffentlicht: (2026)
Automatically Generating UI Code from Screenshot: A Divide-and-Conquer-Based Approach
von: Wan, Yuxuan, et al.
Veröffentlicht: (2024)
von: Wan, Yuxuan, et al.
Veröffentlicht: (2024)
ResearchEnvBench: Benchmarking Agents on Environment Synthesis for Research Code Execution
von: Wang, Yubang, et al.
Veröffentlicht: (2026)
von: Wang, Yubang, et al.
Veröffentlicht: (2026)
Survey of GenAI for Automotive Software Development: From Requirements to Executable Code
von: Petrovic, Nenad, et al.
Veröffentlicht: (2025)
von: Petrovic, Nenad, et al.
Veröffentlicht: (2025)
A Benchmark for Localizing Code and Non-Code Issues in Software Projects
von: Zhang, Zejun, et al.
Veröffentlicht: (2025)
von: Zhang, Zejun, et al.
Veröffentlicht: (2025)
Task Abstention for Large Language Models in Code Generation
von: Zhou, Yanke, et al.
Veröffentlicht: (2026)
von: Zhou, Yanke, et al.
Veröffentlicht: (2026)
Learning to Align Human Code Preferences
von: Yin, Xin, et al.
Veröffentlicht: (2025)
von: Yin, Xin, et al.
Veröffentlicht: (2025)
Beyond Execution: Static-Analysis Rewards and Hint-Conditioned Diffusion RL for Code Generation
von: Ouyang, Shuyin, et al.
Veröffentlicht: (2026)
von: Ouyang, Shuyin, et al.
Veröffentlicht: (2026)
RobuNFR: Evaluating the Robustness of Large Language Models on Non-Functional Requirements Aware Code Generation
von: Lin, Feng, et al.
Veröffentlicht: (2025)
von: Lin, Feng, et al.
Veröffentlicht: (2025)
Don't Complete It! Preventing Unhelpful Code Completion for Productive and Sustainable Neural Code Completion Systems
von: Sun, Zhensu, et al.
Veröffentlicht: (2022)
von: Sun, Zhensu, et al.
Veröffentlicht: (2022)
Seed-CTS: Unleashing the Power of Tree Search for Superior Performance in Competitive Coding Tasks
von: Wang, Hao, et al.
Veröffentlicht: (2024)
von: Wang, Hao, et al.
Veröffentlicht: (2024)
CodeFort: Robust Training for Code Generation Models
von: Zhang, Yuhao, et al.
Veröffentlicht: (2024)
von: Zhang, Yuhao, et al.
Veröffentlicht: (2024)
Treefix: Enabling Execution with a Tree of Prefixes
von: Souza, Beatriz, et al.
Veröffentlicht: (2025)
von: Souza, Beatriz, et al.
Veröffentlicht: (2025)
TaskEval: Assessing Difficulty of Code Generation Tasks for Large Language Models
von: Tambon, Florian, et al.
Veröffentlicht: (2024)
von: Tambon, Florian, et al.
Veröffentlicht: (2024)
TransAgent: Enhancing LLM-Based Code Translation via Fine-Grained Execution Alignment
von: Yuan, Zhiqiang, et al.
Veröffentlicht: (2024)
von: Yuan, Zhiqiang, et al.
Veröffentlicht: (2024)
SolidCoder: Bridging the Mental-Reality Gap in LLM Code Generation through Concrete Execution
von: Lee, Woojin, et al.
Veröffentlicht: (2026)
von: Lee, Woojin, et al.
Veröffentlicht: (2026)
Do Machines and Humans Focus on Similar Code? Exploring Explainability of Large Language Models in Code Summarization
von: Li, Jiliang, et al.
Veröffentlicht: (2024)
von: Li, Jiliang, et al.
Veröffentlicht: (2024)
Terminus-4B: Can a Smaller Model Replace Frontier LLMs at Agentic Execution Tasks?
von: Garg, Spandan, et al.
Veröffentlicht: (2026)
von: Garg, Spandan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Tree-of-Code: A Tree-Structured Exploring Framework for End-to-End Code Generation and Execution in Complex Task Handling
von: Ni, Ziyi, et al.
Veröffentlicht: (2024) -
SWE-Hub: A Unified Production System for Scalable, Executable Software Engineering Tasks
von: Zeng, Yucheng, et al.
Veröffentlicht: (2026) -
GitTaskBench: A Benchmark for Code Agents Solving Real-World Tasks Through Code Repository Leveraging
von: Ni, Ziyi, et al.
Veröffentlicht: (2025) -
RA-Gen: A Controllable Code Generation Framework Using ReAct for Multi-Agent Task Execution
von: Liu, Aofan, et al.
Veröffentlicht: (2025) -
RedCode: Risky Code Execution and Generation Benchmark for Code Agents
von: Guo, Chengquan, et al.
Veröffentlicht: (2024)