From Runnable to Shippable: Multi-Agent Test-Driven Development for Generating Full-Stack Web Applications from Requirements
Fuente:
arXiv
Saved in:
| Main Authors: | Wan, Yuxuan, Liang, Tingshuo, Xu, Jiakai, Xiao, Jingyu, Huo, Yintong, Lyu, Michael R |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Automatically Generating Web Applications from Requirements Via Multi-Agent Test-Driven Development
by: Wan, Yuxuan, et al.
Published: (2025)
by: Wan, Yuxuan, et al.
Published: (2025)
Envisioning Future Interactive Web Development: Editing Webpage with Natural Language
by: Dang, Truong Hai, et al.
Published: (2025)
by: Dang, Truong Hai, et al.
Published: (2025)
MRWeb: An Exploration of Generating Multi-Page Resource-Aware Web Code from UI Designs
by: Wan, Yuxuan, et al.
Published: (2024)
by: Wan, Yuxuan, et al.
Published: (2024)
UIBenchKit: A unified toolkit for design-to-code model evaluation
by: Le, Chinh T., et al.
Published: (2026)
by: Le, Chinh T., et al.
Published: (2026)
EfficientUICoder: Efficient MLLM-based UI Code Generation via Input and Output Token Compression
by: Xiao, Jingyu, et al.
Published: (2025)
by: Xiao, Jingyu, et al.
Published: (2025)
ARC: Compiling Large Multi-Modal Requirement Documents into Runnable Software Systems
by: Kong, Weiyu, et al.
Published: (2026)
by: Kong, Weiyu, et al.
Published: (2026)
DesignBench: A Comprehensive Benchmark for MLLM-based Front-end Code Generation
by: Xiao, Jingyu, et al.
Published: (2025)
by: Xiao, Jingyu, et al.
Published: (2025)
End-to-End Automated Logging via Multi-Agent Framework
by: Zhong, Renyi, et al.
Published: (2025)
by: Zhong, Renyi, et al.
Published: (2025)
ComUICoder: Component-based Reusable UI Code Generation for Complex Websites via Semantic Segmentation and Element-wise Feedback
by: Xiao, Jingyu, et al.
Published: (2026)
by: Xiao, Jingyu, et al.
Published: (2026)
Runnable Directories: The Solution to the Monorepo vs. Multi-repo Debate
by: Ghasemnezhad, Shayan, et al.
Published: (2025)
by: Ghasemnezhad, Shayan, et al.
Published: (2025)
FullStack-Agent: Enhancing Agentic Full-Stack Web Coding via Development-Oriented Testing and Repository Back-Translation
by: Lu, Zimu, et al.
Published: (2026)
by: Lu, Zimu, et al.
Published: (2026)
Exploring Autonomous Agents: A Closer Look at Why They Fail When Completing Tasks
by: Lu, Ruofan, et al.
Published: (2025)
by: Lu, Ruofan, et al.
Published: (2025)
Automatically Generating UI Code from Screenshot: A Divide-and-Conquer-Based Approach
by: Wan, Yuxuan, et al.
Published: (2024)
by: Wan, Yuxuan, et al.
Published: (2024)
WebMAC: A Multi-Agent Collaborative Framework for Scenario Testing of Web Systems
by: Wan, Zhenyu, et al.
Published: (2026)
by: Wan, Zhenyu, et al.
Published: (2026)
iReDev: A Knowledge-Driven Multi-Agent Framework for Intelligent Requirements Development
by: Jin, Dongming, et al.
Published: (2025)
by: Jin, Dongming, et al.
Published: (2025)
Interaction2Code: Benchmarking MLLM-based Interactive Webpage Code Generation from Interactive Prototyping
by: Xiao, Jingyu, et al.
Published: (2024)
by: Xiao, Jingyu, et al.
Published: (2024)
Small is Beautiful: A Practical and Efficient Log Parsing Framework
by: Wang, Minxing, et al.
Published: (2026)
by: Wang, Minxing, et al.
Published: (2026)
Enhancing LLM-Based Coding Tools through Native Integration of IDE-Derived Static Context
by: Li, Yichen, et al.
Published: (2024)
by: Li, Yichen, et al.
Published: (2024)
LogUpdater: Automated Detection and Repair of Specific Defects in Logging Statements
by: Zhong, Renyi, et al.
Published: (2024)
by: Zhong, Renyi, et al.
Published: (2024)
Single-Language Evidence Is Insufficient for Automated Logging: A Multilingual Benchmark and Empirical Study with LLMs
by: Zhong, Renyi, et al.
Published: (2026)
by: Zhong, Renyi, et al.
Published: (2026)
Self-Organizing Multi-Agent Systems for Continuous Software Development
by: Lyu, Wenhan, et al.
Published: (2026)
by: Lyu, Wenhan, et al.
Published: (2026)
90% Faster, 100% Code-Free: MLLM-Driven Zero-Code 3D Game Development
by: Yang, Runxin, et al.
Published: (2025)
by: Yang, Runxin, et al.
Published: (2025)
Primary Breadth-First Development (PBFD): An Approach to Full Stack Software Development
by: Liu, Dong
Published: (2025)
by: Liu, Dong
Published: (2025)
Next Edit Prediction: Learning to Predict Code Edits from Context and Interaction History
by: Lu, Ruofan, et al.
Published: (2025)
by: Lu, Ruofan, et al.
Published: (2025)
FullStack Bench: Evaluating LLMs as Full Stack Coders
by: Bytedance-Seed-Foundation-Code-Team, et al.
Published: (2024)
by: Bytedance-Seed-Foundation-Code-Team, et al.
Published: (2024)
Temac: Multi-Agent Collaboration for Automated Web GUI Testing
by: Liu, Chenxu, et al.
Published: (2025)
by: Liu, Chenxu, et al.
Published: (2025)
Larger Is Not Always Better: Exploring Small Open-source Language Models in Logging Statement Generation
by: Zhong, Renyi, et al.
Published: (2025)
by: Zhong, Renyi, et al.
Published: (2025)
TrajAudit: Automated Failure Diagnosis for Agentic Coding Systems
by: Wang, Minxing, et al.
Published: (2026)
by: Wang, Minxing, et al.
Published: (2026)
Prompting for Automatic Log Template Extraction
by: Xu, Junjielong, et al.
Published: (2023)
by: Xu, Junjielong, et al.
Published: (2023)
Requirements Development and Formalization for Reliable Code Generation: A Multi-Agent Vision
by: Lu, Xu, et al.
Published: (2025)
by: Lu, Xu, et al.
Published: (2025)
TREAT: A Code LLMs Trustworthiness / Reliability Evaluation and Testing Framework
by: Gao, Shuzheng, et al.
Published: (2025)
by: Gao, Shuzheng, et al.
Published: (2025)
Requirements-Driven Automated Software Testing: A Systematic Review
by: Wang, Fanyu, et al.
Published: (2025)
by: Wang, Fanyu, et al.
Published: (2025)
Exploring the Effectiveness of LLMs in Automated Logging Generation: An Empirical Study
by: Li, Yichen, et al.
Published: (2023)
by: Li, Yichen, et al.
Published: (2023)
APITestGenie: Generating Web API Tests from Requirements and API Specifications with LLMs
by: Pereira, André, et al.
Published: (2026)
by: Pereira, André, et al.
Published: (2026)
AUITestAgent: Automatic Requirements Oriented GUI Function Testing
by: Hu, Yongxiang, et al.
Published: (2024)
by: Hu, Yongxiang, et al.
Published: (2024)
Distinguishability-guided Test Program Generation for WebAssembly Runtime Performance Testing
by: Jiang, Shuyao, et al.
Published: (2024)
by: Jiang, Shuyao, et al.
Published: (2024)
REAgent: Requirement-Driven LLM Agents for Software Issue Resolution
by: Kuang, Shiqi, et al.
Published: (2026)
by: Kuang, Shiqi, et al.
Published: (2026)
LogFold: Compressing Logs with Structured Tokens and Hybrid Encoding
by: Shan, Shiwen, et al.
Published: (2026)
by: Shan, Shiwen, et al.
Published: (2026)
CelerLog: Fast Log Parsing via Dynamic Routing
by: Shan, Shiwen, et al.
Published: (2026)
by: Shan, Shiwen, et al.
Published: (2026)
ConfLogger: Enhance Systems' Configuration Diagnosability through Configuration Logging
by: Shan, Shiwen, et al.
Published: (2025)
by: Shan, Shiwen, et al.
Published: (2025)
Similar Items
-
Automatically Generating Web Applications from Requirements Via Multi-Agent Test-Driven Development
by: Wan, Yuxuan, et al.
Published: (2025) -
Envisioning Future Interactive Web Development: Editing Webpage with Natural Language
by: Dang, Truong Hai, et al.
Published: (2025) -
MRWeb: An Exploration of Generating Multi-Page Resource-Aware Web Code from UI Designs
by: Wan, Yuxuan, et al.
Published: (2024) -
UIBenchKit: A unified toolkit for design-to-code model evaluation
by: Le, Chinh T., et al.
Published: (2026) -
EfficientUICoder: Efficient MLLM-based UI Code Generation via Input and Output Token Compression
by: Xiao, Jingyu, et al.
Published: (2025)