FullStack-Agent: Enhancing Agentic Full-Stack Web Coding via Development-Oriented Testing and Repository Back-Translation
Fuente:
arXiv
Saved in:
| Main Authors: | Lu, Zimu, Ren, Houxing, Yang, Yunqiao, Wang, Ke, Zong, Zhuofan, Zhan, Mingjie, Li, Hongsheng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FullStack Bench: Evaluating LLMs as Full Stack Coders
by: Bytedance-Seed-Foundation-Code-Team, et al.
Published: (2024)
by: Bytedance-Seed-Foundation-Code-Team, et al.
Published: (2024)
WebGen-Agent: Enhancing Interactive Website Generation with Multi-Level Feedback and Step-Level Reinforcement Learning
by: Lu, Zimu, et al.
Published: (2025)
by: Lu, Zimu, et al.
Published: (2025)
SlidesGen-Bench: Evaluating Slides Generation via Computational and Quantitative Metrics
by: Yang, Yunqiao, et al.
Published: (2026)
by: Yang, Yunqiao, et al.
Published: (2026)
Towards Robust Real-World Spreadsheet Understanding with Multi-Agent Multi-Format Reasoning
by: Ren, Houxing, et al.
Published: (2026)
by: Ren, Houxing, et al.
Published: (2026)
VoiceAssistant-Eval: Benchmarking AI Assistants across Listening, Speaking, and Viewing
by: Wang, Ke, et al.
Published: (2025)
by: Wang, Ke, et al.
Published: (2025)
Edit-Based Refinement for Parallel Masked Diffusion Language Models
by: Ren, Houxing, et al.
Published: (2026)
by: Ren, Houxing, et al.
Published: (2026)
WebGen-Bench: Evaluating LLMs on Generating Interactive and Functional Websites from Scratch
by: Lu, Zimu, et al.
Published: (2025)
by: Lu, Zimu, et al.
Published: (2025)
Alignment with Fill-In-the-Middle for Enhancing Code Generation
by: Ren, Houxing, et al.
Published: (2025)
by: Ren, Houxing, et al.
Published: (2025)
From Runnable to Shippable: Multi-Agent Test-Driven Development for Generating Full-Stack Web Applications from Requirements
by: Wan, Yuxuan, et al.
Published: (2026)
by: Wan, Yuxuan, et al.
Published: (2026)
Probability-Consistent Preference Optimization for Enhanced LLM Reasoning
by: Yang, Yunqiao, et al.
Published: (2025)
by: Yang, Yunqiao, et al.
Published: (2025)
MathCoder-VL: Bridging Vision and Code for Enhanced Multimodal Mathematical Reasoning
by: Wang, Ke, et al.
Published: (2025)
by: Wang, Ke, et al.
Published: (2025)
MathGenie: Generating Synthetic Data with Question Back-translation for Enhancing Mathematical Reasoning of LLMs
by: Lu, Zimu, et al.
Published: (2024)
by: Lu, Zimu, et al.
Published: (2024)
MathCoder2: Better Math Reasoning from Continued Pretraining on Model-translated Mathematical Code
by: Lu, Zimu, et al.
Published: (2024)
by: Lu, Zimu, et al.
Published: (2024)
EnStack: An Ensemble Stacking Framework of Large Language Models for Enhanced Vulnerability Detection in Source Code
by: Ridoy, Shahriyar Zaman, et al.
Published: (2024)
by: Ridoy, Shahriyar Zaman, et al.
Published: (2024)
Primary Breadth-First Development (PBFD): An Approach to Full Stack Software Development
by: Liu, Dong
Published: (2025)
by: Liu, Dong
Published: (2025)
Empowering Character-level Text Infilling by Eliminating Sub-Tokens
by: Ren, Houxing, et al.
Published: (2024)
by: Ren, Houxing, et al.
Published: (2024)
StackEval: Benchmarking LLMs in Coding Assistance
by: Shah, Nidhish, et al.
Published: (2024)
by: Shah, Nidhish, et al.
Published: (2024)
Step-Controlled DPO: Leveraging Stepwise Error for Enhanced Mathematical Reasoning
by: Lu, Zimu, et al.
Published: (2024)
by: Lu, Zimu, et al.
Published: (2024)
CodeInsight: A Curated Dataset of Practical Coding Solutions from Stack Overflow
by: Beau, Nathanaël, et al.
Published: (2024)
by: Beau, Nathanaël, et al.
Published: (2024)
From Solver to Tutor: Evaluating the Pedagogical Intelligence of LLMs with KMP-Bench
by: Shi, Weikang, et al.
Published: (2026)
by: Shi, Weikang, et al.
Published: (2026)
ToolRosella: Translating Code Repositories into Standardized Tools for Scientific Agents
by: Di, Shimin, et al.
Published: (2026)
by: Di, Shimin, et al.
Published: (2026)
ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation
by: Ren, Houxing, et al.
Published: (2024)
by: Ren, Houxing, et al.
Published: (2024)
RepoST: Scalable Repository-Level Coding Environment Construction with Sandbox Testing
by: Xie, Yiqing, et al.
Published: (2025)
by: Xie, Yiqing, et al.
Published: (2025)
Full-Stack Quantum Software in Practice: Ecosystem, Stakeholders and Challenges
by: Stirbu, Vlad, et al.
Published: (2023)
by: Stirbu, Vlad, et al.
Published: (2023)
Formally and Empirically Verified Methodologies for Scalable Hierarchical Full-Stack Systems
by: Liu, Dong
Published: (2025)
by: Liu, Dong
Published: (2025)
Improving Code Localization with Repository Memory
by: Wang, Boshi, et al.
Published: (2025)
by: Wang, Boshi, et al.
Published: (2025)
Remember Your Trace: Memory-Guided Long-Horizon Agentic Framework for Consistent and Hierarchical Repository-Level Code Documentation
by: Bae, Suyoung, et al.
Published: (2026)
by: Bae, Suyoung, et al.
Published: (2026)
WEBSERV: A Full-Stack and RL-Ready Web Environment for Training Web Agents at Scale
by: Lu, Yuxuan, et al.
Published: (2025)
by: Lu, Yuxuan, et al.
Published: (2025)
SafeCoop: Unravelling Full Stack Safety in Agentic Collaborative Driving
by: Gao, Xiangbo, et al.
Published: (2025)
by: Gao, Xiangbo, et al.
Published: (2025)
EffiSkill: Agent Skill Based Automated Code Efficiency Optimization
by: Wang, Zimu, et al.
Published: (2026)
by: Wang, Zimu, et al.
Published: (2026)
Operationalizing Ethics for AI Agents: How Developers Encode Values into Repository Context Files
by: Treude, Christoph, et al.
Published: (2026)
by: Treude, Christoph, et al.
Published: (2026)
Advancing Quantum Software Engineering: A Vision of Hybrid Full-Stack Iterative Model
by: Khan, Arif Ali, et al.
Published: (2024)
by: Khan, Arif Ali, et al.
Published: (2024)
Non-Fungible Programs: Private Full-Stack Applications for Web3
by: Regalia, Blake, et al.
Published: (2024)
by: Regalia, Blake, et al.
Published: (2024)
RepoTransBench: A Real-World Multilingual Benchmark for Repository-Level Code Translation
by: Wang, Yanli, et al.
Published: (2024)
by: Wang, Yanli, et al.
Published: (2024)
Agent-Oriented Visual Programming for the Web of Things
by: Burattini, Samuele, et al.
Published: (2025)
by: Burattini, Samuele, et al.
Published: (2025)
Repoformer: Selective Retrieval for Repository-Level Code Completion
by: Wu, Di, et al.
Published: (2024)
by: Wu, Di, et al.
Published: (2024)
A Tale of Two Communities: Exploring Academic References on Stack Overflow
by: Huang, Run, et al.
Published: (2024)
by: Huang, Run, et al.
Published: (2024)
Software for Information Storage and Retrieval Tested, Evaluated and Compared. Part IV--Indexing and Full Text Retrieval Programs.
by: Sieverts, Eric G., et al.
Published: (1992)
by: Sieverts, Eric G., et al.
Published: (1992)
M2rc-Eval: Massively Multilingual Repository-level Code Completion Evaluation
by: Liu, Jiaheng, et al.
Published: (2024)
by: Liu, Jiaheng, et al.
Published: (2024)
Prompting Large Language Models to Tackle the Full Software Development Lifecycle: A Case Study
by: Li, Bowen, et al.
Published: (2024)
by: Li, Bowen, et al.
Published: (2024)
Similar Items
-
FullStack Bench: Evaluating LLMs as Full Stack Coders
by: Bytedance-Seed-Foundation-Code-Team, et al.
Published: (2024) -
WebGen-Agent: Enhancing Interactive Website Generation with Multi-Level Feedback and Step-Level Reinforcement Learning
by: Lu, Zimu, et al.
Published: (2025) -
SlidesGen-Bench: Evaluating Slides Generation via Computational and Quantitative Metrics
by: Yang, Yunqiao, et al.
Published: (2026) -
Towards Robust Real-World Spreadsheet Understanding with Multi-Agent Multi-Format Reasoning
by: Ren, Houxing, et al.
Published: (2026) -
VoiceAssistant-Eval: Benchmarking AI Assistants across Listening, Speaking, and Viewing
by: Wang, Ke, et al.
Published: (2025)