Solution-oriented Agent-based Models Generation with Verifier-assisted Iterative In-context Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Niu, Tong, Zhang, Weihao, Zhao, Rong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Survey on Code Generation with LLM-based Agents
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
APIGen: Automated Pipeline for Generating Verifiable and Diverse Function-Calling Datasets
von: Liu, Zuxin, et al.
Veröffentlicht: (2024)
von: Liu, Zuxin, et al.
Veröffentlicht: (2024)
CodeVisionary: An Agent-based Framework for Evaluating Large Language Models in Code Generation
von: Wang, Xinchen, et al.
Veröffentlicht: (2025)
von: Wang, Xinchen, et al.
Veröffentlicht: (2025)
The Art of Repair: Optimizing Iterative Program Repair with Instruction-Tuned Models
von: Ruiz, Fernando Vallecillos, et al.
Veröffentlicht: (2025)
von: Ruiz, Fernando Vallecillos, et al.
Veröffentlicht: (2025)
Scoring Verifiers: Evaluating Synthetic Verification for Code and Reasoning
von: Ficek, Aleksander, et al.
Veröffentlicht: (2025)
von: Ficek, Aleksander, et al.
Veröffentlicht: (2025)
Agentless: Demystifying LLM-based Software Engineering Agents
von: Xia, Chunqiu Steven, et al.
Veröffentlicht: (2024)
von: Xia, Chunqiu Steven, et al.
Veröffentlicht: (2024)
DocAgent: A Multi-Agent System for Automated Code Documentation Generation
von: Yang, Dayu, et al.
Veröffentlicht: (2025)
von: Yang, Dayu, et al.
Veröffentlicht: (2025)
Experiential Co-Learning of Software-Developing Agents
von: Qian, Chen, et al.
Veröffentlicht: (2023)
von: Qian, Chen, et al.
Veröffentlicht: (2023)
SWE-Protégé: Learning to Selectively Collaborate With an Expert Unlocks Small Language Models as Software Engineering Agents
von: Kon, Patrick Tser Jern, et al.
Veröffentlicht: (2026)
von: Kon, Patrick Tser Jern, et al.
Veröffentlicht: (2026)
SELA: Tree-Search Enhanced LLM Agents for Automated Machine Learning
von: Chi, Yizhou, et al.
Veröffentlicht: (2024)
von: Chi, Yizhou, et al.
Veröffentlicht: (2024)
BenchEvolver: Frontier Task Synthesis via Solution-Centric Evolution
von: Wu, Yangzhen, et al.
Veröffentlicht: (2026)
von: Wu, Yangzhen, et al.
Veröffentlicht: (2026)
Dive into Claude Code: The Design Space of Today's and Future AI Agent Systems
von: Liu, Jiacheng, et al.
Veröffentlicht: (2026)
von: Liu, Jiacheng, et al.
Veröffentlicht: (2026)
Confucius Code Agent: Scalable Agent Scaffolding for Real-World Codebases
von: Wong, Sherman, et al.
Veröffentlicht: (2025)
von: Wong, Sherman, et al.
Veröffentlicht: (2025)
VERINA: Benchmarking Verifiable Code Generation
von: Ye, Zhe, et al.
Veröffentlicht: (2025)
von: Ye, Zhe, et al.
Veröffentlicht: (2025)
CLEVER: A Curated Benchmark for Formally Verified Code Generation
von: Thakur, Amitayush, et al.
Veröffentlicht: (2025)
von: Thakur, Amitayush, et al.
Veröffentlicht: (2025)
AlgoTune: Can Language Models Speed Up General-Purpose Numerical Programs?
von: Press, Ori, et al.
Veröffentlicht: (2025)
von: Press, Ori, et al.
Veröffentlicht: (2025)
Learning to Generate Unit Tests for Automated Debugging
von: Prasad, Archiki, et al.
Veröffentlicht: (2025)
von: Prasad, Archiki, et al.
Veröffentlicht: (2025)
Does Few-Shot Learning Help LLM Performance in Code Synthesis?
von: Xu, Derek, et al.
Veröffentlicht: (2024)
von: Xu, Derek, et al.
Veröffentlicht: (2024)
From I/O to Code with Discovery Agent
von: Dong, Yihong, et al.
Veröffentlicht: (2026)
von: Dong, Yihong, et al.
Veröffentlicht: (2026)
Toward Training Superintelligent Software Agents through Self-Play SWE-RL
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
Live-SWE-agent: Can Software Engineering Agents Self-Evolve on the Fly?
von: Xia, Chunqiu Steven, et al.
Veröffentlicht: (2025)
von: Xia, Chunqiu Steven, et al.
Veröffentlicht: (2025)
Selective Prompt Anchoring for Code Generation
von: Tian, Yuan, et al.
Veröffentlicht: (2024)
von: Tian, Yuan, et al.
Veröffentlicht: (2024)
Diversity Empowers Intelligence: Integrating Expertise of Software Engineering Agents
von: Zhang, Kexun, et al.
Veröffentlicht: (2024)
von: Zhang, Kexun, et al.
Veröffentlicht: (2024)
CP-Agent: Agentic Constraint Programming
von: Szeider, Stefan
Veröffentlicht: (2025)
von: Szeider, Stefan
Veröffentlicht: (2025)
Security Degradation in Iterative AI Code Generation -- A Systematic Analysis of the Paradox
von: Shukla, Shivani, et al.
Veröffentlicht: (2025)
von: Shukla, Shivani, et al.
Veröffentlicht: (2025)
Utilizing Deep Learning to Optimize Software Development Processes
von: Li, Keqin, et al.
Veröffentlicht: (2024)
von: Li, Keqin, et al.
Veröffentlicht: (2024)
Bridging Online and Offline RL: Contextual Bandit Learning for Multi-Turn Code Generation
von: Chen, Ziru, et al.
Veröffentlicht: (2026)
von: Chen, Ziru, et al.
Veröffentlicht: (2026)
Learner-Tailored Program Repair: A Solution Generator with Iterative Edit-Driven Retrieval Enhancement
von: Dai, Zhenlong, et al.
Veröffentlicht: (2026)
von: Dai, Zhenlong, et al.
Veröffentlicht: (2026)
HGAdapter: Hypergraph-based Adapters in Language Models for Code Summarization and Clone Detection
von: Yang, Guang, et al.
Veröffentlicht: (2025)
von: Yang, Guang, et al.
Veröffentlicht: (2025)
DDPT: Diffusion-Driven Prompt Tuning for Large Language Model Code Generation
von: Li, Jinyang, et al.
Veröffentlicht: (2025)
von: Li, Jinyang, et al.
Veröffentlicht: (2025)
On Problems of Implicit Context Compression for Software Engineering Agents
von: Gelvan, Kirill, et al.
Veröffentlicht: (2026)
von: Gelvan, Kirill, et al.
Veröffentlicht: (2026)
$\textbf{Only-IF}$:Revealing the Decisive Effect of Instruction Diversity on Generalization
von: Zhang, Dylan, et al.
Veröffentlicht: (2024)
von: Zhang, Dylan, et al.
Veröffentlicht: (2024)
GSO: Challenging Software Optimization Tasks for Evaluating SWE-Agents
von: Shetty, Manish, et al.
Veröffentlicht: (2025)
von: Shetty, Manish, et al.
Veröffentlicht: (2025)
Maestro: Joint Graph & Config Optimization for Reliable AI Agents
von: Wang, Wenxiao, et al.
Veröffentlicht: (2025)
von: Wang, Wenxiao, et al.
Veröffentlicht: (2025)
ReCode: Reinforcing Code Generation with Reasoning-Process Rewards
von: Fan, Lishui, et al.
Veröffentlicht: (2025)
von: Fan, Lishui, et al.
Veröffentlicht: (2025)
KOCO-BENCH: Can Large Language Models Leverage Domain Knowledge in Software Development?
von: Jiang, Xue, et al.
Veröffentlicht: (2026)
von: Jiang, Xue, et al.
Veröffentlicht: (2026)
ToolFactory: Automating Tool Generation by Leveraging LLM to Understand REST API Documentations
von: Ni, Xinyi, et al.
Veröffentlicht: (2025)
von: Ni, Xinyi, et al.
Veröffentlicht: (2025)
AFlow: Automating Agentic Workflow Generation
von: Zhang, Jiayi, et al.
Veröffentlicht: (2024)
von: Zhang, Jiayi, et al.
Veröffentlicht: (2024)
GRAPH-GRPO-LEX: Contract Graph Modeling and Reinforcement Learning with Group Relative Policy Optimization
von: Dechtiar, Moriya, et al.
Veröffentlicht: (2025)
von: Dechtiar, Moriya, et al.
Veröffentlicht: (2025)
DeepCRCEval: Revisiting the Evaluation of Code Review Comment Generation
von: Lu, Junyi, et al.
Veröffentlicht: (2024)
von: Lu, Junyi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Survey on Code Generation with LLM-based Agents
von: Dong, Yihong, et al.
Veröffentlicht: (2025) -
APIGen: Automated Pipeline for Generating Verifiable and Diverse Function-Calling Datasets
von: Liu, Zuxin, et al.
Veröffentlicht: (2024) -
CodeVisionary: An Agent-based Framework for Evaluating Large Language Models in Code Generation
von: Wang, Xinchen, et al.
Veröffentlicht: (2025) -
The Art of Repair: Optimizing Iterative Program Repair with Instruction-Tuned Models
von: Ruiz, Fernando Vallecillos, et al.
Veröffentlicht: (2025) -
Scoring Verifiers: Evaluating Synthetic Verification for Code and Reasoning
von: Ficek, Aleksander, et al.
Veröffentlicht: (2025)