Scaling Coding Agents via Atomic Skills
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ma, Yingwei, Liu, Yue, Yang, Xinlong, Li, Yanhao, Fu, Kelin, Miao, Yibo, Xie, Yuchong, Wang, Zhexu, Cheung, Shing-Chi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Kimi-Dev: Agentless Training as Skill Prior for SWE-Agents
von: Yang, Zonghan, et al.
Veröffentlicht: (2025)
von: Yang, Zonghan, et al.
Veröffentlicht: (2025)
RulER: Automated Rule-Based Semantic Error Localization and Repair for Code Translation
von: Jin, Shuo, et al.
Veröffentlicht: (2025)
von: Jin, Shuo, et al.
Veröffentlicht: (2025)
Multi-Docker-Eval: A `Shovel of the Gold Rush' Benchmark on Automatic Environment Building for Software Engineering
von: Fu, Kelin, et al.
Veröffentlicht: (2025)
von: Fu, Kelin, et al.
Veröffentlicht: (2025)
What Builds Effective In-Context Examples for Code Generation?
von: Li, Dongze, et al.
Veröffentlicht: (2025)
von: Li, Dongze, et al.
Veröffentlicht: (2025)
Concerned with Data Contamination? Assessing Countermeasures in Code Language Model
von: Cao, Jialun, et al.
Veröffentlicht: (2024)
von: Cao, Jialun, et al.
Veröffentlicht: (2024)
Automatic Build Repair for Test Cases using Incompatible Java Versions
von: Mak, Ching Hang, et al.
Veröffentlicht: (2024)
von: Mak, Ching Hang, et al.
Veröffentlicht: (2024)
Word Closure-Based Metamorphic Testing for Machine Translation
von: Xie, Xiaoyuan, et al.
Veröffentlicht: (2023)
von: Xie, Xiaoyuan, et al.
Veröffentlicht: (2023)
When LLMs Meet API Documentation: Can Retrieval Augmentation Aid Code Generation Just as It Helps Developers?
von: Chen, Jingyi, et al.
Veröffentlicht: (2025)
von: Chen, Jingyi, et al.
Veröffentlicht: (2025)
Multi-Agent Systems for Dataset Adaptation in Software Engineering: Capabilities, Limitations, and Future Directions
von: Chen, Jingyi, et al.
Veröffentlicht: (2025)
von: Chen, Jingyi, et al.
Veröffentlicht: (2025)
Can Emulating Semantic Translation Help LLMs with Code Translation? A Study Based on Pseudocode
von: Chen, Songqiang, et al.
Veröffentlicht: (2025)
von: Chen, Songqiang, et al.
Veröffentlicht: (2025)
CAM: A Causality-based Analysis Framework for Multi-Agent Code Generation Systems
von: Lyu, Zongyi, et al.
Veröffentlicht: (2026)
von: Lyu, Zongyi, et al.
Veröffentlicht: (2026)
ReuseDroid: A VLM-empowered Android UI Test Migrator Boosted by Active Feedback
von: Li, Xiaolei, et al.
Veröffentlicht: (2025)
von: Li, Xiaolei, et al.
Veröffentlicht: (2025)
Isolating Language-Coding from Problem-Solving: Benchmarking LLMs with PseudoEval
von: Wu, Jiarong, et al.
Veröffentlicht: (2025)
von: Wu, Jiarong, et al.
Veröffentlicht: (2025)
Understanding and Bridging the Planner-Coder Gap: A Systematic Study on the Robustness of Multi-Agent Systems for Code Generation
von: Lyu, Zongyi, et al.
Veröffentlicht: (2025)
von: Lyu, Zongyi, et al.
Veröffentlicht: (2025)
Towards Understanding the Bugs in Solidity Compiler
von: Ma, Haoyang, et al.
Veröffentlicht: (2024)
von: Ma, Haoyang, et al.
Veröffentlicht: (2024)
A Tale of Two DL Cities: When Library Tests Meet Compiler
von: Shen, Qingchao, et al.
Veröffentlicht: (2024)
von: Shen, Qingchao, et al.
Veröffentlicht: (2024)
Can Large Language Models Model Programs Formally?
von: Chen, Zhiyong, et al.
Veröffentlicht: (2026)
von: Chen, Zhiyong, et al.
Veröffentlicht: (2026)
MR-Scout: Automated Synthesis of Metamorphic Relations from Existing Test Cases
von: Xu, Congying, et al.
Veröffentlicht: (2023)
von: Xu, Congying, et al.
Veröffentlicht: (2023)
Understanding and Characterizing Mock Assertions in Unit Tests
von: Zhu, Hengcheng, et al.
Veröffentlicht: (2025)
von: Zhu, Hengcheng, et al.
Veröffentlicht: (2025)
Across Programming Language Silos: A Study on Cross-Lingual Retrieval-augmented Code Generation
von: Zhu, Qiming, et al.
Veröffentlicht: (2025)
von: Zhu, Qiming, et al.
Veröffentlicht: (2025)
DOMAINEVAL: An Auto-Constructed Benchmark for Multi-Domain Code Generation
von: Zhu, Qiming, et al.
Veröffentlicht: (2024)
von: Zhu, Qiming, et al.
Veröffentlicht: (2024)
Thinking Longer, Not Larger: Enhancing Software Engineering Agents via Scaling Test-Time Compute
von: Ma, Yingwei, et al.
Veröffentlicht: (2025)
von: Ma, Yingwei, et al.
Veröffentlicht: (2025)
How Far are App Secrets from Being Stolen? A Case Study on Android
von: Wei, Lili, et al.
Veröffentlicht: (2025)
von: Wei, Lili, et al.
Veröffentlicht: (2025)
Alibaba LingmaAgent: Improving Automated Issue Resolution via Comprehensive Repository Exploration
von: Ma, Yingwei, et al.
Veröffentlicht: (2024)
von: Ma, Yingwei, et al.
Veröffentlicht: (2024)
EffiSkill: Agent Skill Based Automated Code Efficiency Optimization
von: Wang, Zimu, et al.
Veröffentlicht: (2026)
von: Wang, Zimu, et al.
Veröffentlicht: (2026)
Combinatorial Synthesis: Scaling Code RLVR via Atomic Decomposition and Recombination
von: Zheng, Jiasheng, et al.
Veröffentlicht: (2026)
von: Zheng, Jiasheng, et al.
Veröffentlicht: (2026)
SkillReducer: Optimizing LLM Agent Skills for Token Efficiency
von: Gao, Yudong, et al.
Veröffentlicht: (2026)
von: Gao, Yudong, et al.
Veröffentlicht: (2026)
MR-Coupler: Automated Metamorphic Test Generation via Functional Coupling Analysis
von: Xu, Congying, et al.
Veröffentlicht: (2026)
von: Xu, Congying, et al.
Veröffentlicht: (2026)
Optimization-Aware Test Generation for Deep Learning Compilers
von: Shen, Qingchao, et al.
Veröffentlicht: (2025)
von: Shen, Qingchao, et al.
Veröffentlicht: (2025)
RedCodeAgent: Automatic Red-teaming Agent against Diverse Code Agents
von: Guo, Chengquan, et al.
Veröffentlicht: (2025)
von: Guo, Chengquan, et al.
Veröffentlicht: (2025)
SkillClone: Multi-Modal Clone Detection and Clone Propagation Analysis in the Agent Skill Ecosystem
von: Zhu, Jiaying, et al.
Veröffentlicht: (2026)
von: Zhu, Jiaying, et al.
Veröffentlicht: (2026)
MR-Adopt: Automatic Deduction of Input Transformation Function for Metamorphic Testing
von: Xu, Congying, et al.
Veröffentlicht: (2024)
von: Xu, Congying, et al.
Veröffentlicht: (2024)
ZTaint-Havoc: From Havoc Mode to Zero-Execution Fuzzing-Driven Taint Inference
von: Xie, Yuchong, et al.
Veröffentlicht: (2025)
von: Xie, Yuchong, et al.
Veröffentlicht: (2025)
EmbedAgent: Benchmarking Large Language Models in Embedded System Development
von: Xu, Ruiyang, et al.
Veröffentlicht: (2025)
von: Xu, Ruiyang, et al.
Veröffentlicht: (2025)
Lingma SWE-GPT: An Open Development-Process-Centric Language Model for Automated Software Improvement
von: Ma, Yingwei, et al.
Veröffentlicht: (2024)
von: Ma, Yingwei, et al.
Veröffentlicht: (2024)
ClarEval: A Benchmark for Evaluating Clarification Skills of Code Agents under Ambiguous Instructions
von: Li, Jialin, et al.
Veröffentlicht: (2026)
von: Li, Jialin, et al.
Veröffentlicht: (2026)
CODECLEANER: Elevating Standards with A Robust Data Contamination Mitigation Toolkit
von: Cao, Jialun, et al.
Veröffentlicht: (2024)
von: Cao, Jialun, et al.
Veröffentlicht: (2024)
CodeChemist: Functional Knowledge Transfer for Low-Resource Code Generation via Test-Time Scaling
von: Wang, Kaixin, et al.
Veröffentlicht: (2025)
von: Wang, Kaixin, et al.
Veröffentlicht: (2025)
Immersion in the GitHub Universe: Scaling Coding Agents to Mastery
von: Zhao, Jiale, et al.
Veröffentlicht: (2026)
von: Zhao, Jiale, et al.
Veröffentlicht: (2026)
LLM-Powered Detection of Price Manipulation in DeFi
von: Liu, Lu, et al.
Veröffentlicht: (2025)
von: Liu, Lu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Kimi-Dev: Agentless Training as Skill Prior for SWE-Agents
von: Yang, Zonghan, et al.
Veröffentlicht: (2025) -
RulER: Automated Rule-Based Semantic Error Localization and Repair for Code Translation
von: Jin, Shuo, et al.
Veröffentlicht: (2025) -
Multi-Docker-Eval: A `Shovel of the Gold Rush' Benchmark on Automatic Environment Building for Software Engineering
von: Fu, Kelin, et al.
Veröffentlicht: (2025) -
What Builds Effective In-Context Examples for Code Generation?
von: Li, Dongze, et al.
Veröffentlicht: (2025) -
Concerned with Data Contamination? Assessing Countermeasures in Code Language Model
von: Cao, Jialun, et al.
Veröffentlicht: (2024)