Saved in:
| Main Authors: | Ren, Yukun, Yu, Siwei, Chen, Kai, Ma, Jianwei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2512.14429 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Empowering smart app development with SolidGPT: an edge-cloud hybrid AI agent framework
by: Hu, Liao, et al.
Published: (2025)
by: Hu, Liao, et al.
Published: (2025)
Teaching LLMs to Learn Tool Trialing and Execution through Environment Interaction
by: Gao, Xingjie, et al.
Published: (2026)
by: Gao, Xingjie, et al.
Published: (2026)
A microservice architecture for real-time IoT data processing: A reusable Web of things approach for smart ports
by: Ortiz, Guadalupe, et al.
Published: (2024)
by: Ortiz, Guadalupe, et al.
Published: (2024)
A Problem-Oriented Perspective and Anchor Verification for Code Optimization
by: Ye, Tong, et al.
Published: (2024)
by: Ye, Tong, et al.
Published: (2024)
LLM4EFFI: Leveraging Large Language Models to Enhance Code Efficiency and Correctness
by: Ye, Tong, et al.
Published: (2025)
by: Ye, Tong, et al.
Published: (2025)
COAST: Enhancing the Code Debugging Ability of LLMs through Communicative Agent Based Data Synthesis
by: Yang, Weiqing, et al.
Published: (2024)
by: Yang, Weiqing, et al.
Published: (2024)
Multi-line AI-assisted Code Authoring
by: Dunay, Omer, et al.
Published: (2024)
by: Dunay, Omer, et al.
Published: (2024)
Exploring the Challenges and Opportunities of AI-assisted Codebase Generation
by: Eibl, Philipp, et al.
Published: (2025)
by: Eibl, Philipp, et al.
Published: (2025)
RefAgent: A Multi-agent LLM-based Framework for Automatic Software Refactoring
by: Oueslati, Khouloud, et al.
Published: (2025)
by: Oueslati, Khouloud, et al.
Published: (2025)
Cooperative Multi-agent Approach for Automated Computer Game Testing
by: Shirzadeh-hajimahmood, Samira, et al.
Published: (2024)
by: Shirzadeh-hajimahmood, Samira, et al.
Published: (2024)
Automated QoR improvement in OpenROAD with coding agents
by: Ghose, Amur, et al.
Published: (2026)
by: Ghose, Amur, et al.
Published: (2026)
Architectural Constraints Alignment in AI-assisted, Platform-based Service Development
by: Irion, Julius, et al.
Published: (2026)
by: Irion, Julius, et al.
Published: (2026)
An end-to-end agentic pipeline for smart contract translation and quality evaluation
by: Goel, Abhinav, et al.
Published: (2026)
by: Goel, Abhinav, et al.
Published: (2026)
Understanding Conversational Patterns in Multi-agent Programming: A Case Study on Fibonacci Game Development
by: Basu, Srijita, et al.
Published: (2026)
by: Basu, Srijita, et al.
Published: (2026)
CrashFixer: A crash resolution agent for the Linux kernel
by: Mathai, Alex, et al.
Published: (2025)
by: Mathai, Alex, et al.
Published: (2025)
Taming Scylla: Understanding the multi-headed agentic daemon of the coding seas
by: Villmow, Micah
Published: (2026)
by: Villmow, Micah
Published: (2026)
VeriStruct: AI-assisted Automated Verification of Data-Structure Modules in Verus
by: Sun, Chuyue, et al.
Published: (2025)
by: Sun, Chuyue, et al.
Published: (2025)
A Multi-agent AI System for Deep Learning Model Migration from TensorFlow to JAX
by: Nikolov, Stoyan, et al.
Published: (2026)
by: Nikolov, Stoyan, et al.
Published: (2026)
LLM-based agents for automating the enhancement of user story quality: An early report
by: Zhang, Zheying, et al.
Published: (2024)
by: Zhang, Zheying, et al.
Published: (2024)
Green My LLM: Studying the key factors affecting the energy consumption of code assistants
by: Coignion, Tristan, et al.
Published: (2024)
by: Coignion, Tristan, et al.
Published: (2024)
AI-assisted Code Authoring at Scale: Fine-tuning, deploying, and mixed methods evaluation
by: Murali, Vijayaraghavan, et al.
Published: (2023)
by: Murali, Vijayaraghavan, et al.
Published: (2023)
SpecAgent: A Speculative Retrieval and Forecasting Agent for Code Completion
by: Ma, George, et al.
Published: (2025)
by: Ma, George, et al.
Published: (2025)
Coherence Collapse: Diagnosing Why Code Agents Fail After Reaching the Right Code
by: Kim, Myeongsoo, et al.
Published: (2026)
by: Kim, Myeongsoo, et al.
Published: (2026)
Beyond Final Code: A Process-Oriented Error Analysis of Software Development Agents in Real-World GitHub Scenarios
by: Chen, Zhi, et al.
Published: (2025)
by: Chen, Zhi, et al.
Published: (2025)
LLM assisted web application functional requirements generation: A case study of four popular LLMs over a Mess Management System
by: Gupta, Rashmi, et al.
Published: (2025)
by: Gupta, Rashmi, et al.
Published: (2025)
Beyond Autonomy: A Dynamic Tiered AgentRunner Framework for Governable and Resilient Enterprise AI Execution
by: Pan, Kai, et al.
Published: (2026)
by: Pan, Kai, et al.
Published: (2026)
Real Money, Fake Models: Deceptive Model Claims in Shadow APIs
by: Zhang, Yage, et al.
Published: (2026)
by: Zhang, Yage, et al.
Published: (2026)
SEW: Self-Evolving Agentic Workflows for Automated Code Generation
by: Liu, Siwei, et al.
Published: (2025)
by: Liu, Siwei, et al.
Published: (2025)
ATime-Consistent Benchmark for Repository-Level Software Engineering Evaluation
by: Xianpeng, et al.
Published: (2026)
by: Xianpeng, et al.
Published: (2026)
Faster Configuration Performance Bug Testing with Neural Dual-level Prioritization
by: Ma, Youpeng, et al.
Published: (2025)
by: Ma, Youpeng, et al.
Published: (2025)
Learning to Debug: LLM-Organized Knowledge Trees for Solving RTL Assertion Failures
by: Bai, Yunsheng, et al.
Published: (2025)
by: Bai, Yunsheng, et al.
Published: (2025)
An LLM-based Quantitative Framework for Evaluating High-Stealthy Backdoor Risks in OSS Supply Chains
by: Yan, Zihe, et al.
Published: (2025)
by: Yan, Zihe, et al.
Published: (2025)
AI-Driven Self-Evolving Software: A Promising Path Toward Software Automation
by: Cai, Liyi, et al.
Published: (2025)
by: Cai, Liyi, et al.
Published: (2025)
InteractScience: Programmatic and Visually-Grounded Evaluation of Interactive Scientific Demonstration Code Generation
by: Chen, Qiaosheng, et al.
Published: (2025)
by: Chen, Qiaosheng, et al.
Published: (2025)
SWE-Next: Scalable Real-World Software Engineering Tasks for Agents
by: Liang, Jiarong, et al.
Published: (2026)
by: Liang, Jiarong, et al.
Published: (2026)
The Foundation Cracks: A Comprehensive Study on Bugs and Testing Practices in LLM Libraries
by: Jiang, Weipeng, et al.
Published: (2025)
by: Jiang, Weipeng, et al.
Published: (2025)
LogiDebrief: A Signal-Temporal Logic based Automated Debriefing Approach with Large Language Models Integration
by: Chen, Zirong, et al.
Published: (2025)
by: Chen, Zirong, et al.
Published: (2025)
Same Signal, Different Semantics: A Cross-Framework Behavioral Analysis of Software Engineering Agents
by: Ma, Wei, et al.
Published: (2026)
by: Ma, Wei, et al.
Published: (2026)
Model-Enhanced LLM-Driven VUI Testing of VPA Apps
by: Li, Suwan, et al.
Published: (2024)
by: Li, Suwan, et al.
Published: (2024)
On the Challenges of Fuzzing Techniques via Large Language Models
by: Huang, Linghan, et al.
Published: (2024)
by: Huang, Linghan, et al.
Published: (2024)
Similar Items
-
Empowering smart app development with SolidGPT: an edge-cloud hybrid AI agent framework
by: Hu, Liao, et al.
Published: (2025) -
Teaching LLMs to Learn Tool Trialing and Execution through Environment Interaction
by: Gao, Xingjie, et al.
Published: (2026) -
A microservice architecture for real-time IoT data processing: A reusable Web of things approach for smart ports
by: Ortiz, Guadalupe, et al.
Published: (2024) -
A Problem-Oriented Perspective and Anchor Verification for Code Optimization
by: Ye, Tong, et al.
Published: (2024) -
LLM4EFFI: Leveraging Large Language Models to Enhance Code Efficiency and Correctness
by: Ye, Tong, et al.
Published: (2025)