DARS: Dynamic Action Re-Sampling to Enhance Coding Agent Performance by Adaptive Tree Traversal
Fuente:
arXiv
Salvato in:
| Autori principali: | Aggarwal, Vaibhav, Kamal, Ojasv, Japesh, Abhinav, Jin, Zhijing, Schölkopf, Bernhard |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SWE-Pruner: Self-Adaptive Context Pruning for Coding Agents
di: Wang, Yuhang, et al.
Pubblicazione: (2026)
di: Wang, Yuhang, et al.
Pubblicazione: (2026)
ECO: Enhanced Code Optimization via Performance-Aware Prompting for Code-LLMs
di: Kim, Su-Hyeon, et al.
Pubblicazione: (2025)
di: Kim, Su-Hyeon, et al.
Pubblicazione: (2025)
Automatically Benchmarking LLM Code Agents through Agent-Driven Annotation and Evaluation
di: Fu, Lingyue, et al.
Pubblicazione: (2025)
di: Fu, Lingyue, et al.
Pubblicazione: (2025)
AgentPack: A Dataset of Code Changes, Co-Authored by Agents and Humans
di: Zi, Yangtian, et al.
Pubblicazione: (2025)
di: Zi, Yangtian, et al.
Pubblicazione: (2025)
Is Vibe Coding Safe? Benchmarking Vulnerability of Agent-Generated Code in Real-World Tasks
di: Zhao, Songwen, et al.
Pubblicazione: (2025)
di: Zhao, Songwen, et al.
Pubblicazione: (2025)
Enhancing Code LLMs with Reinforcement Learning in Code Generation: A Survey
di: Wang, Junqiao, et al.
Pubblicazione: (2024)
di: Wang, Junqiao, et al.
Pubblicazione: (2024)
A Critical Study of What Code-LLMs (Do Not) Learn
di: Anand, Abhinav, et al.
Pubblicazione: (2024)
di: Anand, Abhinav, et al.
Pubblicazione: (2024)
From Code Foundation Models to Agents and Applications: A Comprehensive Survey and Practical Guide to Code Intelligence
di: Yang, Jian, et al.
Pubblicazione: (2025)
di: Yang, Jian, et al.
Pubblicazione: (2025)
Beyond pass@k: Redundancy-Aware RLVR for Multi-Sample Code Generation
di: Florian, Le Bronnec, et al.
Pubblicazione: (2026)
di: Florian, Le Bronnec, et al.
Pubblicazione: (2026)
CodeScout: Contextual Problem Statement Enhancement for Software Agents
di: Suri, Manan, et al.
Pubblicazione: (2026)
di: Suri, Manan, et al.
Pubblicazione: (2026)
SIMCOPILOT: Evaluating Large Language Models for Copilot-Style Code Generation
di: Jiang, Mingchao, et al.
Pubblicazione: (2025)
di: Jiang, Mingchao, et al.
Pubblicazione: (2025)
PerfCodeBench: Benchmarking LLMs for System-Level High-Performance Code Optimization
di: Jing, Huihao, et al.
Pubblicazione: (2026)
di: Jing, Huihao, et al.
Pubblicazione: (2026)
Towards Exception Safety Code Generation with Intermediate Representation Agents Framework
di: Zhang, Xuanming, et al.
Pubblicazione: (2024)
di: Zhang, Xuanming, et al.
Pubblicazione: (2024)
EffiSkill: Agent Skill Based Automated Code Efficiency Optimization
di: Wang, Zimu, et al.
Pubblicazione: (2026)
di: Wang, Zimu, et al.
Pubblicazione: (2026)
Dynamic Scaling of Unit Tests for Code Reward Modeling
di: Ma, Zeyao, et al.
Pubblicazione: (2025)
di: Ma, Zeyao, et al.
Pubblicazione: (2025)
ReFuzzer: Feedback-Driven Approach to Enhance Validity of LLM-Generated Test Programs
di: Shree, Iti, et al.
Pubblicazione: (2025)
di: Shree, Iti, et al.
Pubblicazione: (2025)
Agent-Diff: Benchmarking LLM Agents on Enterprise API Tasks via Code Execution with State-Diff-Based Evaluation
di: Pysklo, Hubert M., et al.
Pubblicazione: (2026)
di: Pysklo, Hubert M., et al.
Pubblicazione: (2026)
Seeker: Towards Exception Safety Code Generation with Intermediate Language Agents Framework
di: Zhang, Xuanming, et al.
Pubblicazione: (2024)
di: Zhang, Xuanming, et al.
Pubblicazione: (2024)
Agentic Harness Engineering: Observability-Driven Automatic Evolution of Coding-Agent Harnesses
di: Lin, Jiahang, et al.
Pubblicazione: (2026)
di: Lin, Jiahang, et al.
Pubblicazione: (2026)
MaintainCoder: Maintainable Code Generation Under Dynamic Requirements
di: Wang, Zhengren, et al.
Pubblicazione: (2025)
di: Wang, Zhengren, et al.
Pubblicazione: (2025)
Overeager Coding Agents: Measuring Out-of-Scope Actions on Benign Tasks
di: Qu, Yubin, et al.
Pubblicazione: (2026)
di: Qu, Yubin, et al.
Pubblicazione: (2026)
Large Language Models are Qualified Benchmark Builders: Rebuilding Pre-Training Datasets for Advancing Code Intelligence Tasks
di: Yang, Kang, et al.
Pubblicazione: (2025)
di: Yang, Kang, et al.
Pubblicazione: (2025)
EvoCodeBench: An Evolving Code Generation Benchmark with Domain-Specific Evaluations
di: Li, Jia, et al.
Pubblicazione: (2024)
di: Li, Jia, et al.
Pubblicazione: (2024)
To Diff or Not to Diff? Structure-Aware and Adaptive Output Formats for Efficient LLM-based Code Editing
di: Cheng, Wei, et al.
Pubblicazione: (2026)
di: Cheng, Wei, et al.
Pubblicazione: (2026)
Enhancing Project-Specific Code Completion by Inferring Internal API Information
di: Deng, Le, et al.
Pubblicazione: (2025)
di: Deng, Le, et al.
Pubblicazione: (2025)
EffiLearner: Enhancing Efficiency of Generated Code via Self-Optimization
di: Huang, Dong, et al.
Pubblicazione: (2024)
di: Huang, Dong, et al.
Pubblicazione: (2024)
Enhanced Automated Code Vulnerability Repair using Large Language Models
di: de-Fitero-Dominguez, David, et al.
Pubblicazione: (2024)
di: de-Fitero-Dominguez, David, et al.
Pubblicazione: (2024)
Is Your Benchmark (Still) Useful? Dynamic Benchmarking for Code Language Models
di: Guan, Batu, et al.
Pubblicazione: (2025)
di: Guan, Batu, et al.
Pubblicazione: (2025)
Iterative Refinement of Project-Level Code Context for Precise Code Generation with Compiler Feedback
di: Bi, Zhangqian, et al.
Pubblicazione: (2024)
di: Bi, Zhangqian, et al.
Pubblicazione: (2024)
AdaCCD: Adaptive Semantic Contrasts Discovery Based Cross Lingual Adaptation for Code Clone Detection
di: Du, Yangkai, et al.
Pubblicazione: (2023)
di: Du, Yangkai, et al.
Pubblicazione: (2023)
CodeAgent: Enhancing Code Generation with Tool-Integrated Agent Systems for Real-World Repo-level Coding Challenges
di: Zhang, Kechi, et al.
Pubblicazione: (2024)
di: Zhang, Kechi, et al.
Pubblicazione: (2024)
BeyondSWE: Can Current Code Agent Survive Beyond Single-Repo Bug Fixing?
di: Chen, Guoxin, et al.
Pubblicazione: (2026)
di: Chen, Guoxin, et al.
Pubblicazione: (2026)
ProjectEval: A Benchmark for Programming Agents Automated Evaluation on Project-Level Code Generation
di: Liu, Kaiyuan, et al.
Pubblicazione: (2025)
di: Liu, Kaiyuan, et al.
Pubblicazione: (2025)
Showing LLM-Generated Code Selectively Based on Confidence of LLMs
di: Li, Jia, et al.
Pubblicazione: (2024)
di: Li, Jia, et al.
Pubblicazione: (2024)
R2C2-Coder: Enhancing and Benchmarking Real-world Repository-level Code Completion Abilities of Code Large Language Models
di: Deng, Ken, et al.
Pubblicazione: (2024)
di: Deng, Ken, et al.
Pubblicazione: (2024)
Instruction Adherence in Coding Agent Configuration Files: A Factorial Study of Four File-Structure Variables
di: McMillan, Damon
Pubblicazione: (2026)
di: McMillan, Damon
Pubblicazione: (2026)
Iterative Self-Training for Code Generation via Reinforced Re-Ranking
di: Sorokin, Nikita, et al.
Pubblicazione: (2025)
di: Sorokin, Nikita, et al.
Pubblicazione: (2025)
Scalable Defect Detection via Traversal on Code Graph
di: Liu, Zhengyao, et al.
Pubblicazione: (2024)
di: Liu, Zhengyao, et al.
Pubblicazione: (2024)
Python Symbolic Execution with LLM-powered Code Generation
di: Wang, Wenhan, et al.
Pubblicazione: (2024)
di: Wang, Wenhan, et al.
Pubblicazione: (2024)
CodePod: A Language-Agnostic Hierarchical Scoping System for Interactive Development
di: Li, Hebi, et al.
Pubblicazione: (2023)
di: Li, Hebi, et al.
Pubblicazione: (2023)
Documenti analoghi
-
SWE-Pruner: Self-Adaptive Context Pruning for Coding Agents
di: Wang, Yuhang, et al.
Pubblicazione: (2026) -
ECO: Enhanced Code Optimization via Performance-Aware Prompting for Code-LLMs
di: Kim, Su-Hyeon, et al.
Pubblicazione: (2025) -
Automatically Benchmarking LLM Code Agents through Agent-Driven Annotation and Evaluation
di: Fu, Lingyue, et al.
Pubblicazione: (2025) -
AgentPack: A Dataset of Code Changes, Co-Authored by Agents and Humans
di: Zi, Yangtian, et al.
Pubblicazione: (2025) -
Is Vibe Coding Safe? Benchmarking Vulnerability of Agent-Generated Code in Real-World Tasks
di: Zhao, Songwen, et al.
Pubblicazione: (2025)