Saved in:
| Main Authors: | Masera, Marco, Leone, Alessandro, Köster, Johannes, Molineris, Ivan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2505.02841 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ML-Dev-Bench: Comparative Analysis of AI Agents on ML development workflows
by: Padigela, Harshith, et al.
Published: (2025)
by: Padigela, Harshith, et al.
Published: (2025)
GeoAnalystBench: A GeoAI benchmark for assessing large language models for spatial analysis workflow and code generation
by: Zhang, Qianheng, et al.
Published: (2025)
by: Zhang, Qianheng, et al.
Published: (2025)
Post-hoc LLM-Supported Debugging of Distributed Processes
by: Schiese, Dennis, et al.
Published: (2025)
by: Schiese, Dennis, et al.
Published: (2025)
Do AI models help produce verified bug fixes?
by: Huang, Li, et al.
Published: (2025)
by: Huang, Li, et al.
Published: (2025)
A self-adaptive system of systems architecture to enable its ad-hoc scalability: Unmanned Vehicle Fleet -- Mission Control Center Case study
by: Sadik, Ahmed R., et al.
Published: (2024)
by: Sadik, Ahmed R., et al.
Published: (2024)
A systematic review of generative AI usage for IT project management
by: Anghel, Ionut, et al.
Published: (2026)
by: Anghel, Ionut, et al.
Published: (2026)
Automating the Correctness Assessment of AI-generated Code for Security Contexts
by: Cotroneo, Domenico, et al.
Published: (2023)
by: Cotroneo, Domenico, et al.
Published: (2023)
GenAI-based test case generation and execution in SDV platform
by: Zyberaj, Denesa, et al.
Published: (2025)
by: Zyberaj, Denesa, et al.
Published: (2025)
PEFA-AI: Advancing Open-source LLMs for RTL generation using Progressive Error Feedback Agentic-AI
by: Narayanan, Athma, et al.
Published: (2025)
by: Narayanan, Athma, et al.
Published: (2025)
Digital Twins & ZeroConf AI: Structuring Automated Intelligent Pipelines for Industrial Applications
by: Picone, Marco, et al.
Published: (2026)
by: Picone, Marco, et al.
Published: (2026)
Evaluating AI-generated code for C++, Fortran, Go, Java, Julia, Matlab, Python, R, and Rust
by: Diehl, Patrick, et al.
Published: (2024)
by: Diehl, Patrick, et al.
Published: (2024)
Meta-Engineering Harnesses for AI-Native Software Production: A Contract-Driven Adversarial Verification Architecture with Early Deployment Report
by: Sengupta, Satadru, et al.
Published: (2026)
by: Sengupta, Satadru, et al.
Published: (2026)
Guarded Repair for Harm-Aware Post-hoc Replacement of LLM Mathematical Reasoning
by: Xia, Haizhou
Published: (2026)
by: Xia, Haizhou
Published: (2026)
Towards knowledge-based workflows: a semantic approach to atomistic simulations for mechanical and thermodynamic properties
by: Guzman, Abril Azocar, et al.
Published: (2026)
by: Guzman, Abril Azocar, et al.
Published: (2026)
A Comparison of Conversational Models and Humans in Answering Technical Questions: the Firefox Case
by: Correia, Joao, et al.
Published: (2025)
by: Correia, Joao, et al.
Published: (2025)
Modeling Resilience of Collaborative AI Systems
by: Rimawi, Diaeddin, et al.
Published: (2024)
by: Rimawi, Diaeddin, et al.
Published: (2024)
AI Techniques in the Microservices Life-Cycle: A Systematic Mapping Study
by: Moreschini, Sergio, et al.
Published: (2023)
by: Moreschini, Sergio, et al.
Published: (2023)
AI-Assisted Assessment of Coding Practices in Modern Code Review
by: Vijayvergiya, Manushree, et al.
Published: (2024)
by: Vijayvergiya, Manushree, et al.
Published: (2024)
The importance of visual modelling languages in generative software engineering
by: Rossi, Roberto
Published: (2024)
by: Rossi, Roberto
Published: (2024)
Greening AI-enabled Systems with Software Engineering: A Research Agenda for Environmentally Sustainable AI Practices
by: Cruz, Luís, et al.
Published: (2025)
by: Cruz, Luís, et al.
Published: (2025)
Poisoning Programs by Un-Repairing Code: Security Concerns of AI-generated Code
by: Improta, Cristina
Published: (2024)
by: Improta, Cristina
Published: (2024)
PELLI: Framework to effectively integrate LLMs for quality software generation
by: Krebs, Rasmus, et al.
Published: (2026)
by: Krebs, Rasmus, et al.
Published: (2026)
Large Language Models in Code Co-generation for Safe Autonomous Vehicles
by: Nouri, Ali, et al.
Published: (2025)
by: Nouri, Ali, et al.
Published: (2025)
Assessing LLM code generation quality through path planning tasks
by: Chen, Wanyi, et al.
Published: (2025)
by: Chen, Wanyi, et al.
Published: (2025)
Lifecycle-Aware code generation: Leveraging Software Engineering Phases in LLMs
by: Xing, Xing, et al.
Published: (2025)
by: Xing, Xing, et al.
Published: (2025)
Evaluating perturbation robustness of generative systems that use COBOL code inputs
by: Ackerman, Samuel, et al.
Published: (2025)
by: Ackerman, Samuel, et al.
Published: (2025)
Large Process Models: A Vision for Business Process Management in the Age of Generative AI
by: Kampik, Timotheus, et al.
Published: (2023)
by: Kampik, Timotheus, et al.
Published: (2023)
LlamaRestTest: Effective REST API Testing with Small Language Models
by: Kim, Myeongsoo, et al.
Published: (2025)
by: Kim, Myeongsoo, et al.
Published: (2025)
Beyond the 'Diff': Addressing Agentic Entropy in Agentic Software Development
by: Casserini, Matteo, et al.
Published: (2026)
by: Casserini, Matteo, et al.
Published: (2026)
The Future of Generative AI in Software Engineering: A Vision from Industry and Academia in the European GENIUS Project
by: Gröpler, Robin, et al.
Published: (2025)
by: Gröpler, Robin, et al.
Published: (2025)
Quality Assurance of LLM-generated Code: Addressing Non-Functional Quality Characteristics
by: Sun, Xin, et al.
Published: (2025)
by: Sun, Xin, et al.
Published: (2025)
Canonical Intermediate Representation for LLM-based optimization problem formulation and code generation
by: Lyu, Zhongyuan, et al.
Published: (2026)
by: Lyu, Zhongyuan, et al.
Published: (2026)
Evaluating SAP Joule for Code Generation
by: Heisler, Joshua, et al.
Published: (2025)
by: Heisler, Joshua, et al.
Published: (2025)
Towards LLM-generated explanations for Component-based Knowledge Graph Question Answering Systems
by: Schiese, Dennis, et al.
Published: (2025)
by: Schiese, Dennis, et al.
Published: (2025)
Leveraging LLMs for User Stories in AI Systems: UStAI Dataset
by: Yamani, Asma, et al.
Published: (2025)
by: Yamani, Asma, et al.
Published: (2025)
Echoes of AI: Investigating the Downstream Effects of AI Assistants on Software Maintainability
by: Borg, Markus, et al.
Published: (2025)
by: Borg, Markus, et al.
Published: (2025)
Assessing AI Detectors in Identifying AI-Generated Code: Implications for Education
by: Pan, Wei Hung, et al.
Published: (2024)
by: Pan, Wei Hung, et al.
Published: (2024)
AI Assurance: A Comprehensive Testing Strategy for Enterprise AI Systems
by: Badagi, Chitra, et al.
Published: (2026)
by: Badagi, Chitra, et al.
Published: (2026)
PACIFIC: a framework for generating benchmarks to check Precise Automatically Checked Instruction Following In Code
by: Dreyfuss, Itay, et al.
Published: (2025)
by: Dreyfuss, Itay, et al.
Published: (2025)
Workflow for Safe-AI
by: Veljanovska, Suzana, et al.
Published: (2025)
by: Veljanovska, Suzana, et al.
Published: (2025)
Similar Items
-
ML-Dev-Bench: Comparative Analysis of AI Agents on ML development workflows
by: Padigela, Harshith, et al.
Published: (2025) -
GeoAnalystBench: A GeoAI benchmark for assessing large language models for spatial analysis workflow and code generation
by: Zhang, Qianheng, et al.
Published: (2025) -
Post-hoc LLM-Supported Debugging of Distributed Processes
by: Schiese, Dennis, et al.
Published: (2025) -
Do AI models help produce verified bug fixes?
by: Huang, Li, et al.
Published: (2025) -
A self-adaptive system of systems architecture to enable its ad-hoc scalability: Unmanned Vehicle Fleet -- Mission Control Center Case study
by: Sadik, Ahmed R., et al.
Published: (2024)