AppForge: From Assistant to Independent Developer -- Are GPTs Ready for Software Development?
Fuente:
arXiv
Saved in:
| Main Authors: | Ran, Dezhi, Cao, Yuan, Wu, Mengzhou, Chen, Simin, Guo, Yuzhe, Ren, Jun, Song, Zihe, Yu, Hao, Wei, Jialei, Li, Linyi, Yang, Wei, Ray, Baishakhi, Xie, Tao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Foundation Model Engineering: Engineering Foundation Models Just as Engineering Software
by: Ran, Dezhi, et al.
Published: (2024)
by: Ran, Dezhi, et al.
Published: (2024)
An Infrastructure Software Perspective Toward Computation Offloading between Executable Specifications and Foundation Models
by: Ran, Dezhi, et al.
Published: (2025)
by: Ran, Dezhi, et al.
Published: (2025)
Trustworthy AI Software Engineers
by: Aleti, Aldeida, et al.
Published: (2026)
by: Aleti, Aldeida, et al.
Published: (2026)
An Empirical Study and Theoretical Explanation on Task-Level Model-Merging Collapse
by: Cao, Yuan, et al.
Published: (2026)
by: Cao, Yuan, et al.
Published: (2026)
Dynamic Benchmarking of Reasoning Capabilities in Code Large Language Models Under Data Contamination
by: Chen, Simin, et al.
Published: (2025)
by: Chen, Simin, et al.
Published: (2025)
Skill-Adpative Imitation Learning for UI Test Reuse
by: Wu, Mengzhou, et al.
Published: (2024)
by: Wu, Mengzhou, et al.
Published: (2024)
Red Teaming Program Repair Agents: When Correct Patches can Hide Vulnerabilities
by: Chen, Simin, et al.
Published: (2025)
by: Chen, Simin, et al.
Published: (2025)
From User Interface to Agent Interface: Efficiency Optimization of UI Representations for LLM Agents
by: Ran, Dezhi, et al.
Published: (2025)
by: Ran, Dezhi, et al.
Published: (2025)
Agentic AI Software Engineers: Programming with Trust
by: Roychoudhury, Abhik, et al.
Published: (2025)
by: Roychoudhury, Abhik, et al.
Published: (2025)
REFINE: Enhancing Program Repair Agents through Context-Aware Patch Refinement
by: Pabba, Anvith, et al.
Published: (2025)
by: Pabba, Anvith, et al.
Published: (2025)
Towards a Readiness Model for Usable‐Software Development in Organizations
by: Haifa Al‐Shammare, et al.
Published: (2024)
by: Haifa Al‐Shammare, et al.
Published: (2024)
CodeSense: a Real-World Benchmark and Dataset for Code Semantic Reasoning
by: Roy, Monoshi Kumar, et al.
Published: (2025)
by: Roy, Monoshi Kumar, et al.
Published: (2025)
SpecTra: Enhancing the Code Translation Ability of Language Models by Generating Multi-Modal Specifications
by: Nitin, Vikram, et al.
Published: (2024)
by: Nitin, Vikram, et al.
Published: (2024)
Harnessing the Potential of Gen-AI Coding Assistants in Public Sector Software Development
by: Ng, Kevin KB, et al.
Published: (2024)
by: Ng, Kevin KB, et al.
Published: (2024)
On the Need to Rethink Trust in AI Assistants for Software Development: A Critical Review
by: Baltes, Sebastian, et al.
Published: (2025)
by: Baltes, Sebastian, et al.
Published: (2025)
Your Compiler is Backdooring Your Model: Understanding and Exploiting Compilation Inconsistency Vulnerabilities in Deep Learning Compilers
by: Chen, Simin, et al.
Published: (2025)
by: Chen, Simin, et al.
Published: (2025)
An Empirical Study of Proactive Coding Assistants in Real-World Software Development
by: Li, Lehui, et al.
Published: (2026)
by: Li, Lehui, et al.
Published: (2026)
Benchmarking and Studying the LLM-based Agent System in End-to-End Software Development
by: Zeng, Zhengran, et al.
Published: (2025)
by: Zeng, Zhengran, et al.
Published: (2025)
FaultLine: Automated Proof-of-Vulnerability Generation Using LLM Agents
by: Nitin, Vikram, et al.
Published: (2025)
by: Nitin, Vikram, et al.
Published: (2025)
Code Reasoning for Software Engineering Tasks: A Survey and A Call to Action
by: Pujar, Saurabh, et al.
Published: (2025)
by: Pujar, Saurabh, et al.
Published: (2025)
Yuga: Automatically Detecting Lifetime Annotation Bugs in the Rust Language
by: Nitin, Vikram, et al.
Published: (2023)
by: Nitin, Vikram, et al.
Published: (2023)
The Evolution of Information Seeking in Software Development: Understanding the Role and Impact of AI Assistants
by: Haque, Ebtesam Al, et al.
Published: (2024)
by: Haque, Ebtesam Al, et al.
Published: (2024)
Using AI Assistants in Software Development: A Qualitative Study on Security Practices and Concerns
by: Klemmer, Jan H., et al.
Published: (2024)
by: Klemmer, Jan H., et al.
Published: (2024)
Beyond Pass or Fail: Multi-Dimensional Benchmarking of Foundation Models for Goal-based Mobile UI Navigation
by: Ran, Dezhi, et al.
Published: (2025)
by: Ran, Dezhi, et al.
Published: (2025)
An Empirical Study on the Security Vulnerabilities of GPTs
by: Wu, Tong, et al.
Published: (2025)
by: Wu, Tong, et al.
Published: (2025)
MedForge: Building Medical Foundation Models Like Open Source Software Development
by: Tan, Zheling, et al.
Published: (2025)
by: Tan, Zheling, et al.
Published: (2025)
Agentic Pipelines in Embedded Software Engineering: Emerging Practices and Challenges
by: Sun, Simin, et al.
Published: (2026)
by: Sun, Simin, et al.
Published: (2026)
Disrupting Test Development with AI Assistants
by: Joshi, Vijay, et al.
Published: (2024)
by: Joshi, Vijay, et al.
Published: (2024)
GPTZoo: A Large-scale Dataset of GPTs for the Research Community
by: Hou, Xinyi, et al.
Published: (2024)
by: Hou, Xinyi, et al.
Published: (2024)
What Is an App Store? The Software Engineering Perspective
by: Zhu, Wenhan, et al.
Published: (2024)
by: Zhu, Wenhan, et al.
Published: (2024)
C2SaferRust: Transforming C Projects into Safer Rust with NeuroSymbolic Techniques
by: Nitin, Vikram, et al.
Published: (2025)
by: Nitin, Vikram, et al.
Published: (2025)
The Impact of LLM-Assistants on Software Developer Productivity: A Systematic Review and Mapping Study
by: Mohamed, Amr, et al.
Published: (2025)
by: Mohamed, Amr, et al.
Published: (2025)
Polymer: Development Workflows as Software
by: Parthasarathy, Dhasarathy, et al.
Published: (2025)
by: Parthasarathy, Dhasarathy, et al.
Published: (2025)
Beyond the Commit: Developer Perspectives on Productivity with AI Coding Assistants
by: Chen, Valerie, et al.
Published: (2026)
by: Chen, Valerie, et al.
Published: (2026)
Characterizing Real-World Bugs in Tile Programs for Automated Bug Detection
by: Rathnasuriya, Ravishka, et al.
Published: (2026)
by: Rathnasuriya, Ravishka, et al.
Published: (2026)
Prioritizing App Reviews for Developer Responses on Google Play
by: Jafari, Mohsen, et al.
Published: (2025)
by: Jafari, Mohsen, et al.
Published: (2025)
WebApp1K: A Practical Code-Generation Benchmark for Web App Development
by: Cui, Yi
Published: (2024)
by: Cui, Yi
Published: (2024)
Uncovering Non-native Speakers' Experiences in Global Software Development Teams -- A Bourdieusian Perspective
by: Wang, Yi, et al.
Published: (2025)
by: Wang, Yi, et al.
Published: (2025)
AI Tools in Software Development: Developer Perceptions and Usage Patterns
by: Looi, Mark
Published: (2026)
by: Looi, Mark
Published: (2026)
Rapid Mobile App Development for Generative AI Agents on MIT App Inventor
by: Gao, Jaida, et al.
Published: (2024)
by: Gao, Jaida, et al.
Published: (2024)
Similar Items
-
Foundation Model Engineering: Engineering Foundation Models Just as Engineering Software
by: Ran, Dezhi, et al.
Published: (2024) -
An Infrastructure Software Perspective Toward Computation Offloading between Executable Specifications and Foundation Models
by: Ran, Dezhi, et al.
Published: (2025) -
Trustworthy AI Software Engineers
by: Aleti, Aldeida, et al.
Published: (2026) -
An Empirical Study and Theoretical Explanation on Task-Level Model-Merging Collapse
by: Cao, Yuan, et al.
Published: (2026) -
Dynamic Benchmarking of Reasoning Capabilities in Code Large Language Models Under Data Contamination
by: Chen, Simin, et al.
Published: (2025)