Seed-CTS: Unleashing the Power of Tree Search for Superior Performance in Competitive Coding Tasks
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wang, Hao, Liu, Boyi, Zhang, Yufeng, Chen, Jie |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Exploring and Unleashing the Power of Large Language Models in Automated Code Translation
par: Yang, Zhen, et autres
Publié: (2024)
par: Yang, Zhen, et autres
Publié: (2024)
DSTC: Direct Preference Learning with Only Self-Generated Tests and Code to Improve Code LMs
par: Liu, Zhihan, et autres
Publié: (2024)
par: Liu, Zhihan, et autres
Publié: (2024)
Exploringand Unleashing the Power of Large Language Models in CI/CD Configuration Translation
par: Wang, Chong, et autres
Publié: (2025)
par: Wang, Chong, et autres
Publié: (2025)
XSearch: Explainable Code Search via Concept-to-Code Alignment
par: Liu, Yiming, et autres
Publié: (2026)
par: Liu, Yiming, et autres
Publié: (2026)
GitTaskBench: A Benchmark for Code Agents Solving Real-World Tasks Through Code Repository Leveraging
par: Ni, Ziyi, et autres
Publié: (2025)
par: Ni, Ziyi, et autres
Publié: (2025)
ChiseLLM: Unleashing the Power of Reasoning LLMs for Chisel Agile Hardware Development
par: Wang, Bowei, et autres
Publié: (2025)
par: Wang, Bowei, et autres
Publié: (2025)
CoRe: Benchmarking LLMs Code Reasoning Capabilities through Static Analysis Tasks
par: Xie, Danning, et autres
Publié: (2025)
par: Xie, Danning, et autres
Publié: (2025)
CodeFuse-CommitEval: Towards Benchmarking LLM's Power on Commit Message and Code Change Inconsistency Detection
par: Zhang, Qingyu, et autres
Publié: (2025)
par: Zhang, Qingyu, et autres
Publié: (2025)
SpareCodeSearch: Searching for Code Context When You Have No Spare GPU
par: Nguyen, Minh
Publié: (2025)
par: Nguyen, Minh
Publié: (2025)
Tree-of-Code: A Tree-Structured Exploring Framework for End-to-End Code Generation and Execution in Complex Task Handling
par: Ni, Ziyi, et autres
Publié: (2024)
par: Ni, Ziyi, et autres
Publié: (2024)
ModiGen: A Large Language Model-Based Workflow for Multi-Task Modelica Code Generation
par: Xiang, Jiahui, et autres
Publié: (2025)
par: Xiang, Jiahui, et autres
Publié: (2025)
ViC: Virtual Compiler Is All You Need For Assembly Code Search
par: Gao, Zeyu, et autres
Publié: (2024)
par: Gao, Zeyu, et autres
Publié: (2024)
Personality-Guided Code Generation Using Large Language Models
par: Guo, Yaoqi, et autres
Publié: (2024)
par: Guo, Yaoqi, et autres
Publié: (2024)
Runtime-Structured Task Decomposition for Agentic Coding Systems
par: Asthana, Shubhi, et autres
Publié: (2026)
par: Asthana, Shubhi, et autres
Publié: (2026)
Task Abstention for Large Language Models in Code Generation
par: Zhou, Yanke, et autres
Publié: (2026)
par: Zhou, Yanke, et autres
Publié: (2026)
An Empirical Study of Knowledge Distillation for Code Understanding Tasks
par: Wang, Ruiqi, et autres
Publié: (2025)
par: Wang, Ruiqi, et autres
Publié: (2025)
Tree-of-Code: A Hybrid Approach for Robust Complex Task Planning and Execution
par: Ni, Ziyi, et autres
Publié: (2024)
par: Ni, Ziyi, et autres
Publié: (2024)
MergeRepair: An Exploratory Study on Merging Task-Specific Adapters in Code LLMs for Automated Program Repair
par: Dehghan, Meghdad, et autres
Publié: (2024)
par: Dehghan, Meghdad, et autres
Publié: (2024)
MarsCode Agent: AI-native Automated Bug Fixing
par: Liu, Yizhou, et autres
Publié: (2024)
par: Liu, Yizhou, et autres
Publié: (2024)
Beyond Retrieval: A Multitask Benchmark and Model for Code Search
par: Xue, Siqiao, et autres
Publié: (2026)
par: Xue, Siqiao, et autres
Publié: (2026)
Software Performance Engineering for Foundation Model-Powered Software
par: Zhang, Haoxiang, et autres
Publié: (2024)
par: Zhang, Haoxiang, et autres
Publié: (2024)
EvoGPT: Leveraging LLM-Driven Seed Diversity to Improve Search-Based Test Suite Generation
par: Broide, Lior, et autres
Publié: (2025)
par: Broide, Lior, et autres
Publié: (2025)
EvoCodeBench: A Human-Performance Benchmark for Self-Evolving LLM-Driven Coding Systems
par: Zhang, Wentao, et autres
Publié: (2026)
par: Zhang, Wentao, et autres
Publié: (2026)
TaskEval: Assessing Difficulty of Code Generation Tasks for Large Language Models
par: Tambon, Florian, et autres
Publié: (2024)
par: Tambon, Florian, et autres
Publié: (2024)
CodeR: Issue Resolving with Multi-Agent and Task Graphs
par: Chen, Dong, et autres
Publié: (2024)
par: Chen, Dong, et autres
Publié: (2024)
Analyzing Message-Code Inconsistency in AI Coding Agent-Authored Pull Requests
par: Gong, Jingzhi, et autres
Publié: (2026)
par: Gong, Jingzhi, et autres
Publié: (2026)
Schedule-and-Calibrate: Utility-Guided Multi-Task Reinforcement Learning for Code LLMs
par: Chen, Yujia, et autres
Publié: (2026)
par: Chen, Yujia, et autres
Publié: (2026)
Identifying Performance-Sensitive Configurations in Software Systems through Code Analysis with LLM Agents
par: Wang, Zehao, et autres
Publié: (2024)
par: Wang, Zehao, et autres
Publié: (2024)
Unveiling Code Pre-Trained Models: Investigating Syntax and Semantics Capacities
par: Ma, Wei, et autres
Publié: (2022)
par: Ma, Wei, et autres
Publié: (2022)
RA-Gen: A Controllable Code Generation Framework Using ReAct for Multi-Agent Task Execution
par: Liu, Aofan, et autres
Publié: (2025)
par: Liu, Aofan, et autres
Publié: (2025)
Goedel-Code-Prover: Hierarchical Proof Search for Open State-of-the-Art Code Verification
par: Li, Zenan, et autres
Publié: (2026)
par: Li, Zenan, et autres
Publié: (2026)
GeoCode-GPT: A Large Language Model for Geospatial Code Generation Tasks
par: Hou, Shuyang, et autres
Publié: (2024)
par: Hou, Shuyang, et autres
Publié: (2024)
Ambiguity Resolution with Human Feedback for Code Writing Tasks
par: Nandan, Aditey, et autres
Publié: (2025)
par: Nandan, Aditey, et autres
Publié: (2025)
Automated Benchmark Generation for Repository-Level Coding Tasks
par: Vergopoulos, Konstantinos, et autres
Publié: (2025)
par: Vergopoulos, Konstantinos, et autres
Publié: (2025)
Bias Testing and Mitigation in LLM-based Code Generation
par: Huang, Dong, et autres
Publié: (2023)
par: Huang, Dong, et autres
Publié: (2023)
VeriContest: A Competitive-Programming Benchmark for Verifiable Code Generation
par: Xie, Zichen, et autres
Publié: (2026)
par: Xie, Zichen, et autres
Publié: (2026)
Beyond Function-Level Search: Repository-Aware Dual-Encoder Code Retrieval with Adversarial Verification
par: Liu, Aofan, et autres
Publié: (2025)
par: Liu, Aofan, et autres
Publié: (2025)
Function-to-Style Guidance of LLMs for Code Translation
par: Zhang, Longhui, et autres
Publié: (2025)
par: Zhang, Longhui, et autres
Publié: (2025)
Squeez: Task-Conditioned Tool-Output Pruning for Coding Agents
par: Kovács, Ádám
Publié: (2026)
par: Kovács, Ádám
Publié: (2026)
InfCode-C++: Intent-Guided Semantic Retrieval and AST-Structured Search for C++ Issue Resolution
par: Dong, Qingao, et autres
Publié: (2025)
par: Dong, Qingao, et autres
Publié: (2025)
Documents similaires
-
Exploring and Unleashing the Power of Large Language Models in Automated Code Translation
par: Yang, Zhen, et autres
Publié: (2024) -
DSTC: Direct Preference Learning with Only Self-Generated Tests and Code to Improve Code LMs
par: Liu, Zhihan, et autres
Publié: (2024) -
Exploringand Unleashing the Power of Large Language Models in CI/CD Configuration Translation
par: Wang, Chong, et autres
Publié: (2025) -
XSearch: Explainable Code Search via Concept-to-Code Alignment
par: Liu, Yiming, et autres
Publié: (2026) -
GitTaskBench: A Benchmark for Code Agents Solving Real-World Tasks Through Code Repository Leveraging
par: Ni, Ziyi, et autres
Publié: (2025)