AutoDev: Automated AI-Driven Development
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tufano, Michele, Agarwal, Anisha, Jang, Jinu, Moghaddam, Roshanak Zilouchian, Sundaresan, Neel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Copilot Evaluation Harness: Evaluating LLM-Guided Software Programming
von: Agarwal, Anisha, et al.
Veröffentlicht: (2024)
von: Agarwal, Anisha, et al.
Veröffentlicht: (2024)
RAPGen: An Approach for Fixing Code Inefficiencies in Zero-Shot
von: Garg, Spandan, et al.
Veröffentlicht: (2023)
von: Garg, Spandan, et al.
Veröffentlicht: (2023)
PerfBench: Can Agents Resolve Real-World Performance Bugs?
von: Garg, Spandan, et al.
Veröffentlicht: (2025)
von: Garg, Spandan, et al.
Veröffentlicht: (2025)
SetupBench: Assessing Software Engineering Agents' Ability to Bootstrap Development Environments
von: Arora, Avi, et al.
Veröffentlicht: (2025)
von: Arora, Avi, et al.
Veröffentlicht: (2025)
RefactorBench: Evaluating Stateful Reasoning in Language Agents Through Code
von: Gautam, Dhruv, et al.
Veröffentlicht: (2025)
von: Gautam, Dhruv, et al.
Veröffentlicht: (2025)
The SWE-Bench Illusion: When State-of-the-Art LLMs Remember Instead of Reason
von: Liang, Shanchao, et al.
Veröffentlicht: (2025)
von: Liang, Shanchao, et al.
Veröffentlicht: (2025)
FaultLine: Automated Proof-of-Vulnerability Generation Using LLM Agents
von: Nitin, Vikram, et al.
Veröffentlicht: (2025)
von: Nitin, Vikram, et al.
Veröffentlicht: (2025)
Reinforcement Learning from Automatic Feedback for High-Quality Unit Test Generation
von: Steenhoek, Benjamin, et al.
Veröffentlicht: (2024)
von: Steenhoek, Benjamin, et al.
Veröffentlicht: (2024)
Reinforcement Learning from Automatic Feedback for High-Quality Unit Test Generation
von: Steenhoek, Benjamin, et al.
Veröffentlicht: (2023)
von: Steenhoek, Benjamin, et al.
Veröffentlicht: (2023)
ProjDevBench: Benchmarking AI Coding Agents on End-to-End Project Development
von: Lu, Pengrui, et al.
Veröffentlicht: (2026)
von: Lu, Pengrui, et al.
Veröffentlicht: (2026)
Agentic Bug Reproduction for Effective Automated Program Repair at Google
von: Cheng, Runxiang, et al.
Veröffentlicht: (2025)
von: Cheng, Runxiang, et al.
Veröffentlicht: (2025)
EvoDev: An Iterative Feature-Driven Framework for End-to-End Software Development with LLM-based Agents
von: Liu, Junwei, et al.
Veröffentlicht: (2025)
von: Liu, Junwei, et al.
Veröffentlicht: (2025)
AI for DevSecOps: A Landscape and Future Opportunities
von: Fu, Michael, et al.
Veröffentlicht: (2024)
von: Fu, Michael, et al.
Veröffentlicht: (2024)
DevOps-Gym: Benchmarking AI Agents in Software DevOps Cycle
von: Tang, Yuheng, et al.
Veröffentlicht: (2026)
von: Tang, Yuheng, et al.
Veröffentlicht: (2026)
WebDevJudge: Evaluating (M)LLMs as Critiques for Web Development Quality
von: Li, Chunyang, et al.
Veröffentlicht: (2025)
von: Li, Chunyang, et al.
Veröffentlicht: (2025)
ML-Dev-Bench: Comparative Analysis of AI Agents on ML development workflows
von: Padigela, Harshith, et al.
Veröffentlicht: (2025)
von: Padigela, Harshith, et al.
Veröffentlicht: (2025)
DevLicOps: A Framework for Mitigating Licensing Risks in AI-Generated Code
von: Sharma, Pratyush Nidhi, et al.
Veröffentlicht: (2025)
von: Sharma, Pratyush Nidhi, et al.
Veröffentlicht: (2025)
Designing Adaptive Digital Nudging Systems with LLM-Driven Reasoning
von: Santilli, Tiziano, et al.
Veröffentlicht: (2026)
von: Santilli, Tiziano, et al.
Veröffentlicht: (2026)
Breaking Barriers in Software Testing: The Power of AI-Driven Automation
von: Naqvi, Saba, et al.
Veröffentlicht: (2025)
von: Naqvi, Saba, et al.
Veröffentlicht: (2025)
Auto-SPT: Automating Semantic Preserving Transformations for Code
von: Hooda, Ashish, et al.
Veröffentlicht: (2025)
von: Hooda, Ashish, et al.
Veröffentlicht: (2025)
AutoDroid: LLM-powered Task Automation in Android
von: Wen, Hao, et al.
Veröffentlicht: (2023)
von: Wen, Hao, et al.
Veröffentlicht: (2023)
Closing the Gap: A User Study on the Real-world Usefulness of AI-powered Vulnerability Detection & Repair in the IDE
von: Steenhoek, Benjamin, et al.
Veröffentlicht: (2024)
von: Steenhoek, Benjamin, et al.
Veröffentlicht: (2024)
Evaluating Agent-based Program Repair at Google
von: Rondon, Pat, et al.
Veröffentlicht: (2025)
von: Rondon, Pat, et al.
Veröffentlicht: (2025)
Dynamic Cogeneration of Bug Reproduction Test in Agentic Program Repair
von: Cheng, Runxiang, et al.
Veröffentlicht: (2026)
von: Cheng, Runxiang, et al.
Veröffentlicht: (2026)
LADs: Leveraging LLMs for AI-Driven DevOps
von: Khan, Ahmad Faraz, et al.
Veröffentlicht: (2025)
von: Khan, Ahmad Faraz, et al.
Veröffentlicht: (2025)
GameDevBench: Evaluating Agentic Capabilities Through Game Development
von: Chi, Wayne, et al.
Veröffentlicht: (2026)
von: Chi, Wayne, et al.
Veröffentlicht: (2026)
daVinci-Dev: Agent-native Mid-training for Software Engineering
von: Zeng, Ji, et al.
Veröffentlicht: (2026)
von: Zeng, Ji, et al.
Veröffentlicht: (2026)
AutoIOT: LLM-Driven Automated Natural Language Programming for AIoT Applications
von: Shen, Leming, et al.
Veröffentlicht: (2025)
von: Shen, Leming, et al.
Veröffentlicht: (2025)
DevBench: A Realistic, Developer-Informed Benchmark for Code Generation Models
von: Kumarappan, Adarsh, et al.
Veröffentlicht: (2026)
von: Kumarappan, Adarsh, et al.
Veröffentlicht: (2026)
AI-Driven Development of a Publishing Imprint: Xynapse Traces
von: Zimmerman, Fred
Veröffentlicht: (2025)
von: Zimmerman, Fred
Veröffentlicht: (2025)
AI-Driven Self-Evolving Software: A Promising Path Toward Software Automation
von: Cai, Liyi, et al.
Veröffentlicht: (2025)
von: Cai, Liyi, et al.
Veröffentlicht: (2025)
AutoEmpirical: LLM-Based Automated Research for Empirical Software Fault Analysis
von: Yu, Jiongchi, et al.
Veröffentlicht: (2025)
von: Yu, Jiongchi, et al.
Veröffentlicht: (2025)
Enhancing Educational Efficiency: Generative AI Chatbots and DevOps in Education 4.0
von: Mekić, Edis, et al.
Veröffentlicht: (2024)
von: Mekić, Edis, et al.
Veröffentlicht: (2024)
AI-Generated Smells: An Analysis of Code and Architecture in LLM and Agent-Driven Development
von: Zhu, Yuecai, et al.
Veröffentlicht: (2026)
von: Zhu, Yuecai, et al.
Veröffentlicht: (2026)
Spec-Driven Development:From Code to Contract in the Age of AI Coding Assistants
von: Piskala, Deepak Babu
Veröffentlicht: (2026)
von: Piskala, Deepak Babu
Veröffentlicht: (2026)
AgentMesh: A Cooperative Multi-Agent Generative AI Framework for Software Development Automation
von: Khanzadeh, Sourena
Veröffentlicht: (2025)
von: Khanzadeh, Sourena
Veröffentlicht: (2025)
DoVer: Intervention-Driven Auto Debugging for LLM Multi-Agent Systems
von: Ma, Ming, et al.
Veröffentlicht: (2025)
von: Ma, Ming, et al.
Veröffentlicht: (2025)
AutoRestTest: A Tool for Automated REST API Testing Using LLMs and MARL
von: Stennett, Tyler, et al.
Veröffentlicht: (2025)
von: Stennett, Tyler, et al.
Veröffentlicht: (2025)
RAD-AI: Rethinking Architecture Documentation for AI-Augmented Ecosystems
von: Larsen, Oliver Aleksander, et al.
Veröffentlicht: (2026)
von: Larsen, Oliver Aleksander, et al.
Veröffentlicht: (2026)
When Developer Aid Becomes Security Debt: A Systematic Analysis of Insecure Behaviors in LLM Coding Agents
von: Kozak, Matous, et al.
Veröffentlicht: (2025)
von: Kozak, Matous, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Copilot Evaluation Harness: Evaluating LLM-Guided Software Programming
von: Agarwal, Anisha, et al.
Veröffentlicht: (2024) -
RAPGen: An Approach for Fixing Code Inefficiencies in Zero-Shot
von: Garg, Spandan, et al.
Veröffentlicht: (2023) -
PerfBench: Can Agents Resolve Real-World Performance Bugs?
von: Garg, Spandan, et al.
Veröffentlicht: (2025) -
SetupBench: Assessing Software Engineering Agents' Ability to Bootstrap Development Environments
von: Arora, Avi, et al.
Veröffentlicht: (2025) -
RefactorBench: Evaluating Stateful Reasoning in Language Agents Through Code
von: Gautam, Dhruv, et al.
Veröffentlicht: (2025)