Dive into Claude Code: The Design Space of Today's and Future AI Agent Systems
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Liu, Jiacheng, Zhao, Xiaohan, Shang, Xinyi, Shen, Zhiqiang |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
DocAgent: A Multi-Agent System for Automated Code Documentation Generation
par: Yang, Dayu, et autres
Publié: (2025)
par: Yang, Dayu, et autres
Publié: (2025)
Next-Gen CAPTCHAs: Leveraging the Cognitive Gap for Scalable and Diverse GUI-Agent Defense
par: Liu, Jiacheng, et autres
Publié: (2026)
par: Liu, Jiacheng, et autres
Publié: (2026)
Confucius Code Agent: Scalable Agent Scaffolding for Real-World Codebases
par: Wong, Sherman, et autres
Publié: (2025)
par: Wong, Sherman, et autres
Publié: (2025)
From I/O to Code with Discovery Agent
par: Dong, Yihong, et autres
Publié: (2026)
par: Dong, Yihong, et autres
Publié: (2026)
CodeVisionary: An Agent-based Framework for Evaluating Large Language Models in Code Generation
par: Wang, Xinchen, et autres
Publié: (2025)
par: Wang, Xinchen, et autres
Publié: (2025)
A Survey on Code Generation with LLM-based Agents
par: Dong, Yihong, et autres
Publié: (2025)
par: Dong, Yihong, et autres
Publié: (2025)
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
par: Zheng, Zihan, et autres
Publié: (2025)
par: Zheng, Zihan, et autres
Publié: (2025)
Code to Think, Think to Code: A Survey on Code-Enhanced Reasoning and Reasoning-Driven Code Intelligence in LLMs
par: Yang, Dayu, et autres
Publié: (2025)
par: Yang, Dayu, et autres
Publié: (2025)
Diversity Empowers Intelligence: Integrating Expertise of Software Engineering Agents
par: Zhang, Kexun, et autres
Publié: (2024)
par: Zhang, Kexun, et autres
Publié: (2024)
AppWorld: A Controllable World of Apps and People for Benchmarking Interactive Coding Agents
par: Trivedi, Harsh, et autres
Publié: (2024)
par: Trivedi, Harsh, et autres
Publié: (2024)
ReCode: Reinforcing Code Generation with Reasoning-Process Rewards
par: Fan, Lishui, et autres
Publié: (2025)
par: Fan, Lishui, et autres
Publié: (2025)
Maestro: Joint Graph & Config Optimization for Reliable AI Agents
par: Wang, Wenxiao, et autres
Publié: (2025)
par: Wang, Wenxiao, et autres
Publié: (2025)
Automated Cloud Infrastructure-as-Code Reconciliation with AI Agents
par: Yang, Zhenning, et autres
Publié: (2025)
par: Yang, Zhenning, et autres
Publié: (2025)
Lita: Light Agent Uncovers the Agentic Coding Capabilities of LLMs
par: Dai, Hankun, et autres
Publié: (2025)
par: Dai, Hankun, et autres
Publié: (2025)
CodeEditorBench: Evaluating Code Editing Capability of Large Language Models
par: Guo, Jiawei, et autres
Publié: (2024)
par: Guo, Jiawei, et autres
Publié: (2024)
Encoding architecture algebra
par: Bersier, Stephane, et autres
Publié: (2024)
par: Bersier, Stephane, et autres
Publié: (2024)
CODEMENV: Benchmarking Large Language Models on Code Migration
par: Cheng, Keyuan, et autres
Publié: (2025)
par: Cheng, Keyuan, et autres
Publié: (2025)
Rethinking Repetition Problems of LLMs in Code Generation
par: Dong, Yihong, et autres
Publié: (2025)
par: Dong, Yihong, et autres
Publié: (2025)
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3
par: Sadik, Ahmed R., et autres
Publié: (2025)
par: Sadik, Ahmed R., et autres
Publié: (2025)
Let the Code LLM Edit Itself When You Edit the Code
par: He, Zhenyu, et autres
Publié: (2024)
par: He, Zhenyu, et autres
Publié: (2024)
Solution-oriented Agent-based Models Generation with Verifier-assisted Iterative In-context Learning
par: Niu, Tong, et autres
Publié: (2024)
par: Niu, Tong, et autres
Publié: (2024)
MonoCoder: Domain-Specific Code Language Model for HPC Codes and Tasks
par: Kadosh, Tal, et autres
Publié: (2023)
par: Kadosh, Tal, et autres
Publié: (2023)
ToolFactory: Automating Tool Generation by Leveraging LLM to Understand REST API Documentations
par: Ni, Xinyi, et autres
Publié: (2025)
par: Ni, Xinyi, et autres
Publié: (2025)
Do Large Code Models Understand Programming Concepts? Counterfactual Analysis for Code Predicates
par: Hooda, Ashish, et autres
Publié: (2024)
par: Hooda, Ashish, et autres
Publié: (2024)
Vibe Checker: Aligning Code Evaluation with Human Preference
par: Zhong, Ming, et autres
Publié: (2025)
par: Zhong, Ming, et autres
Publié: (2025)
Experiential Co-Learning of Software-Developing Agents
par: Qian, Chen, et autres
Publié: (2023)
par: Qian, Chen, et autres
Publié: (2023)
Can Coding Agents Be General Agents?
par: Ivanov, Maksim, et autres
Publié: (2026)
par: Ivanov, Maksim, et autres
Publié: (2026)
XFT: Unlocking the Power of Code Instruction Tuning by Simply Merging Upcycled Mixture-of-Experts
par: Ding, Yifeng, et autres
Publié: (2024)
par: Ding, Yifeng, et autres
Publié: (2024)
How Do Your Code LLMs Perform? Empowering Code Instruction Tuning with High-Quality Data
par: Wang, Yejie, et autres
Publié: (2024)
par: Wang, Yejie, et autres
Publié: (2024)
Selective Prompt Anchoring for Code Generation
par: Tian, Yuan, et autres
Publié: (2024)
par: Tian, Yuan, et autres
Publié: (2024)
Large Language Models for Code Summarization
par: Szalontai, Balázs, et autres
Publié: (2024)
par: Szalontai, Balázs, et autres
Publié: (2024)
Your Simulation Runs but Solves the Wrong Physics: PDE-Grounded Intent Verification for LLM-Generated Multiphysics Simulation Code
par: Song, Zhenghan, et autres
Publié: (2026)
par: Song, Zhenghan, et autres
Publié: (2026)
CodeTool: Enhancing Programmatic Tool Invocation of LLMs via Process Supervision
par: Lu, Yifei, et autres
Publié: (2025)
par: Lu, Yifei, et autres
Publié: (2025)
RovoDev Code Reviewer: A Large-Scale Online Evaluation of LLM-based Code Review Automation at Atlassian
par: Tantithamthavorn, Kla, et autres
Publié: (2026)
par: Tantithamthavorn, Kla, et autres
Publié: (2026)
GSO: Challenging Software Optimization Tasks for Evaluating SWE-Agents
par: Shetty, Manish, et autres
Publié: (2025)
par: Shetty, Manish, et autres
Publié: (2025)
VibeTensor: System Software for Deep Learning, Fully Generated by AI Agents
par: Xu, Bing, et autres
Publié: (2026)
par: Xu, Bing, et autres
Publié: (2026)
Scaling Test-Time Compute for Agentic Coding
par: Kim, Joongwon, et autres
Publié: (2026)
par: Kim, Joongwon, et autres
Publié: (2026)
SELA: Tree-Search Enhanced LLM Agents for Automated Machine Learning
par: Chi, Yizhou, et autres
Publié: (2024)
par: Chi, Yizhou, et autres
Publié: (2024)
Scoring Verifiers: Evaluating Synthetic Verification for Code and Reasoning
par: Ficek, Aleksander, et autres
Publié: (2025)
par: Ficek, Aleksander, et autres
Publié: (2025)
AuPair: Golden Example Pairs for Code Repair
par: Mavalankar, Aditi, et autres
Publié: (2025)
par: Mavalankar, Aditi, et autres
Publié: (2025)
Documents similaires
-
DocAgent: A Multi-Agent System for Automated Code Documentation Generation
par: Yang, Dayu, et autres
Publié: (2025) -
Next-Gen CAPTCHAs: Leveraging the Cognitive Gap for Scalable and Diverse GUI-Agent Defense
par: Liu, Jiacheng, et autres
Publié: (2026) -
Confucius Code Agent: Scalable Agent Scaffolding for Real-World Codebases
par: Wong, Sherman, et autres
Publié: (2025) -
From I/O to Code with Discovery Agent
par: Dong, Yihong, et autres
Publié: (2026) -
CodeVisionary: An Agent-based Framework for Evaluating Large Language Models in Code Generation
par: Wang, Xinchen, et autres
Publié: (2025)