EGSS: Entropy-guided Stepwise Scaling for Reliable Software Engineering
Fuente:
arXiv
Salvato in:
| Autori principali: | Mao, Chenhui, Lei, Yuanting, Wei, Zhixiang, Liang, Ming, Wang, Zhixiang, Xu, Jingxuan, Chen, Dajun, Jiang, Wei, Li, Yong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Learning Adaptive Parallel Execution for Efficient Code Localization
di: Xu, Ke, et al.
Pubblicazione: (2026)
di: Xu, Ke, et al.
Pubblicazione: (2026)
An Empirical Study of Interaction Bugs in ROS-based Software
di: Chen, Zhixiang, et al.
Pubblicazione: (2025)
di: Chen, Zhixiang, et al.
Pubblicazione: (2025)
Unified Software Engineering Agent as AI Software Engineer
di: Applis, Leonhard, et al.
Pubblicazione: (2025)
di: Applis, Leonhard, et al.
Pubblicazione: (2025)
SVRepair: Structured Visual Reasoning for Automated Program Repair
di: Tang, Xiaoxuan, et al.
Pubblicazione: (2026)
di: Tang, Xiaoxuan, et al.
Pubblicazione: (2026)
An Empirical Study on Embodied Artificial Intelligence Robot (EAIR) Software Bugs
di: Liao, Zeqin, et al.
Pubblicazione: (2025)
di: Liao, Zeqin, et al.
Pubblicazione: (2025)
Same Signal, Different Semantics: A Cross-Framework Behavioral Analysis of Software Engineering Agents
di: Ma, Wei, et al.
Pubblicazione: (2026)
di: Ma, Wei, et al.
Pubblicazione: (2026)
Programming with Data: Test-Driven Data Engineering for Self-Improving LLMs from Raw Corpora
di: Pan, Chenkai, et al.
Pubblicazione: (2026)
di: Pan, Chenkai, et al.
Pubblicazione: (2026)
Thinking Longer, Not Larger: Enhancing Software Engineering Agents via Scaling Test-Time Compute
di: Ma, Yingwei, et al.
Pubblicazione: (2025)
di: Ma, Yingwei, et al.
Pubblicazione: (2025)
Satori-SWE: Evolutionary Test-Time Scaling for Sample-Efficient Software Engineering
di: Zeng, Guangtao, et al.
Pubblicazione: (2025)
di: Zeng, Guangtao, et al.
Pubblicazione: (2025)
DiagEval: Trajectory-Conditioned Diagnosis for Reliable Software Evaluation with GUI Agents
di: Hong, Sirui, et al.
Pubblicazione: (2026)
di: Hong, Sirui, et al.
Pubblicazione: (2026)
PEACE: Towards Efficient Project-Level Efficiency Optimization via Hybrid Code Editing
di: Ren, Xiaoxue, et al.
Pubblicazione: (2025)
di: Ren, Xiaoxue, et al.
Pubblicazione: (2025)
Towards AI-Native Software Engineering (SE 3.0): A Vision and a Challenge Roadmap
di: Hassan, Ahmed E., et al.
Pubblicazione: (2024)
di: Hassan, Ahmed E., et al.
Pubblicazione: (2024)
Trae Agent: An LLM-based Agent for Software Engineering with Test-time Scaling
di: Trae Research Team, et al.
Pubblicazione: (2025)
di: Trae Research Team, et al.
Pubblicazione: (2025)
Lifecycle-Aware code generation: Leveraging Software Engineering Phases in LLMs
di: Xing, Xing, et al.
Pubblicazione: (2025)
di: Xing, Xing, et al.
Pubblicazione: (2025)
Beyond Final Code: A Process-Oriented Error Analysis of Software Development Agents in Real-World GitHub Scenarios
di: Chen, Zhi, et al.
Pubblicazione: (2025)
di: Chen, Zhi, et al.
Pubblicazione: (2025)
SWE-Next: Scalable Real-World Software Engineering Tasks for Agents
di: Liang, Jiarong, et al.
Pubblicazione: (2026)
di: Liang, Jiarong, et al.
Pubblicazione: (2026)
Rethinking Software Engineering in the Foundation Model Era: From Task-Driven AI Copilots to Goal-Driven AI Pair Programmers
di: Hassan, Ahmed E., et al.
Pubblicazione: (2024)
di: Hassan, Ahmed E., et al.
Pubblicazione: (2024)
Software Performance Engineering for Foundation Model-Powered Software
di: Zhang, Haoxiang, et al.
Pubblicazione: (2024)
di: Zhang, Haoxiang, et al.
Pubblicazione: (2024)
Loosely-Structured Software: Engineering Context, Structure, and Evolution Entropy in Runtime-Rewired Multi-Agent Systems
di: Zhang, Weihao, et al.
Pubblicazione: (2026)
di: Zhang, Weihao, et al.
Pubblicazione: (2026)
Foundation Model Engineering: Engineering Foundation Models Just as Engineering Software
di: Ran, Dezhi, et al.
Pubblicazione: (2024)
di: Ran, Dezhi, et al.
Pubblicazione: (2024)
SWE-Skills-Bench: Do Agent Skills Actually Help in Real-World Software Engineering?
di: Han, Tingxu, et al.
Pubblicazione: (2026)
di: Han, Tingxu, et al.
Pubblicazione: (2026)
HyperAgent: Generalist Software Engineering Agents to Solve Coding Tasks at Scale
di: Phan, Huy Nhat, et al.
Pubblicazione: (2024)
di: Phan, Huy Nhat, et al.
Pubblicazione: (2024)
Rethinking the Value of Agent-Generated Tests for LLM-Based Software Engineering Agents
di: Chen, Zhi, et al.
Pubblicazione: (2026)
di: Chen, Zhi, et al.
Pubblicazione: (2026)
A Comprehensive Study on the Use of Word Embedding Models in Software Engineering Domain
di: Chen, Xiaohan, et al.
Pubblicazione: (2025)
di: Chen, Xiaohan, et al.
Pubblicazione: (2025)
Can GPT-4 Replicate Empirical Software Engineering Research?
di: Liang, Jenny T., et al.
Pubblicazione: (2023)
di: Liang, Jenny T., et al.
Pubblicazione: (2023)
LLMs: A Game-Changer for Software Engineers?
di: Haque, Md Asraful
Pubblicazione: (2024)
di: Haque, Md Asraful
Pubblicazione: (2024)
What Software Engineering Looks Like to AI Agents? -- An Empirical Study of AI-Only Technical Discourse on MoltBook
di: Huo, Junyu, et al.
Pubblicazione: (2026)
di: Huo, Junyu, et al.
Pubblicazione: (2026)
A Systematic Literature Review on Explainability for Machine/Deep Learning-based Software Engineering Research
di: Cao, Sicong, et al.
Pubblicazione: (2024)
di: Cao, Sicong, et al.
Pubblicazione: (2024)
SWE-smith: Scaling Data for Software Engineering Agents
di: Yang, John, et al.
Pubblicazione: (2025)
di: Yang, John, et al.
Pubblicazione: (2025)
Rethinking Software Engineering in the Foundation Model Era: A Curated Catalogue of Challenges in the Development of Trustworthy FMware
di: Hassan, Ahmed E., et al.
Pubblicazione: (2024)
di: Hassan, Ahmed E., et al.
Pubblicazione: (2024)
Beyond the 'Diff': Addressing Agentic Entropy in Agentic Software Development
di: Casserini, Matteo, et al.
Pubblicazione: (2026)
di: Casserini, Matteo, et al.
Pubblicazione: (2026)
AI-Tutoring in Software Engineering Education
di: Frankford, Eduard, et al.
Pubblicazione: (2024)
di: Frankford, Eduard, et al.
Pubblicazione: (2024)
Large Language Model-Based Agents for Software Engineering: A Survey
di: Liu, Junwei, et al.
Pubblicazione: (2024)
di: Liu, Junwei, et al.
Pubblicazione: (2024)
Towards Structured, State-Aware, and Execution-Grounded Reasoning for Software Engineering Agents
di: Tse-Hsun, et al.
Pubblicazione: (2026)
di: Tse-Hsun, et al.
Pubblicazione: (2026)
Can LLMs Replace Human Evaluators? An Empirical Study of LLM-as-a-Judge in Software Engineering
di: Wang, Ruiqi, et al.
Pubblicazione: (2025)
di: Wang, Ruiqi, et al.
Pubblicazione: (2025)
Invisible Load: Uncovering the Challenges of Neurodivergent Women in Software Engineering
di: Zaib, Munazza, et al.
Pubblicazione: (2025)
di: Zaib, Munazza, et al.
Pubblicazione: (2025)
TOM-SWE: User Mental Modeling For Software Engineering Agents
di: Zhou, Xuhui, et al.
Pubblicazione: (2025)
di: Zhou, Xuhui, et al.
Pubblicazione: (2025)
ATime-Consistent Benchmark for Repository-Level Software Engineering Evaluation
di: Xianpeng, et al.
Pubblicazione: (2026)
di: Xianpeng, et al.
Pubblicazione: (2026)
Agentic Software Engineering: Foundational Pillars and a Research Roadmap
di: Hassan, Ahmed E., et al.
Pubblicazione: (2025)
di: Hassan, Ahmed E., et al.
Pubblicazione: (2025)
PaperFit: Vision-in-the-Loop Typesetting Optimization for Scientific Documents
di: Yu, Bihui, et al.
Pubblicazione: (2026)
di: Yu, Bihui, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Learning Adaptive Parallel Execution for Efficient Code Localization
di: Xu, Ke, et al.
Pubblicazione: (2026) -
An Empirical Study of Interaction Bugs in ROS-based Software
di: Chen, Zhixiang, et al.
Pubblicazione: (2025) -
Unified Software Engineering Agent as AI Software Engineer
di: Applis, Leonhard, et al.
Pubblicazione: (2025) -
SVRepair: Structured Visual Reasoning for Automated Program Repair
di: Tang, Xiaoxuan, et al.
Pubblicazione: (2026) -
An Empirical Study on Embodied Artificial Intelligence Robot (EAIR) Software Bugs
di: Liao, Zeqin, et al.
Pubblicazione: (2025)