Rethinking Autonomy: Preventing Failures in AI-Driven Software Engineering
Fuente:
arXiv
Guardado en:
| Autores principales: | Navneet, Satyam Kumar, Chandra, Joydeep |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Policy-Driven AI in Dataspaces: Taxonomy, Explainability, and Pathways for Compliant Innovation
por: Chandra, Joydeep, et al.
Publicado: (2025)
por: Chandra, Joydeep, et al.
Publicado: (2025)
Investigating Detection and Obfuscation of Prompt Injection Attacks Against Software Reverse Engineering AI Agents
por: Crawford, Brian, et al.
Publicado: (2026)
por: Crawford, Brian, et al.
Publicado: (2026)
Data and Context Matter: Towards Generalizing AI-based Software Vulnerability Detection
por: Safdar, Rijha, et al.
Publicado: (2025)
por: Safdar, Rijha, et al.
Publicado: (2025)
DevOps-Gym: Benchmarking AI Agents in Software DevOps Cycle
por: Tang, Yuheng, et al.
Publicado: (2026)
por: Tang, Yuheng, et al.
Publicado: (2026)
LinuxArena: A Control Setting for AI Agents in Live Production Software Environments
por: Tracy, Tyler, et al.
Publicado: (2026)
por: Tracy, Tyler, et al.
Publicado: (2026)
SOK: Exploring Hallucinations and Security Risks in AI-Assisted Software Development with Insights for LLM Deployment
por: Haque, Ariful, et al.
Publicado: (2025)
por: Haque, Ariful, et al.
Publicado: (2025)
Securing the Future of IVR: AI-Driven Innovation with Agile Security, Data Regulation, and Ethical AI Integration
por: Shaikh, Khushbu Mehboob, et al.
Publicado: (2025)
por: Shaikh, Khushbu Mehboob, et al.
Publicado: (2025)
Safety and Performance, Why Not Both? Bi-Objective Optimized Model Compression against Heterogeneous Attacks Toward AI Software Deployment
por: Zhu, Jie, et al.
Publicado: (2024)
por: Zhu, Jie, et al.
Publicado: (2024)
From Detection to Prevention: Explaining Security-Critical Code to Avoid Vulnerabilities
por: Krishnamurthy, Ranjith, et al.
Publicado: (2026)
por: Krishnamurthy, Ranjith, et al.
Publicado: (2026)
Rethinking and Exploring String-Based Malware Family Classification in the Era of LLMs and RAG
por: Chen, Yufan, et al.
Publicado: (2025)
por: Chen, Yufan, et al.
Publicado: (2025)
AdaptiveGuard: Towards Adaptive Runtime Safety for LLM-Powered Software
por: Yang, Rui, et al.
Publicado: (2025)
por: Yang, Rui, et al.
Publicado: (2025)
Bridging Semantics & Structure for Software Vulnerability Detection using Hybrid Network Models
por: Gajjar, Jugal, et al.
Publicado: (2025)
por: Gajjar, Jugal, et al.
Publicado: (2025)
Automatically Generating Rules of Malicious Software Packages via Large Language Model
por: Zhang, XiangRui, et al.
Publicado: (2025)
por: Zhang, XiangRui, et al.
Publicado: (2025)
Co-PatcheR: Collaborative Software Patching with Component(s)-specific Small Reasoning Models
por: Tang, Yuheng, et al.
Publicado: (2025)
por: Tang, Yuheng, et al.
Publicado: (2025)
Harnessing Large Language Models for Software Vulnerability Detection: A Comprehensive Benchmarking Study
por: Tamberg, Karl, et al.
Publicado: (2024)
por: Tamberg, Karl, et al.
Publicado: (2024)
A Qualitative Study on Using ChatGPT for Software Security: Perception vs. Practicality
por: Kholoosi, M. Mehdi, et al.
Publicado: (2024)
por: Kholoosi, M. Mehdi, et al.
Publicado: (2024)
EaTVul: ChatGPT-based Evasion Attack Against Software Vulnerability Detection
por: Liu, Shigang, et al.
Publicado: (2024)
por: Liu, Shigang, et al.
Publicado: (2024)
Reflection-Driven Control for Trustworthy Code Agents
por: Wang, Bin, et al.
Publicado: (2025)
por: Wang, Bin, et al.
Publicado: (2025)
Benchmarking Prompt Engineering Techniques for Secure Code Generation with GPT Models
por: Bruni, Marc, et al.
Publicado: (2025)
por: Bruni, Marc, et al.
Publicado: (2025)
Options, Not Clicks: Lattice Refinement for Consent-Driven MCP Authorization
por: Li, Ying, et al.
Publicado: (2026)
por: Li, Ying, et al.
Publicado: (2026)
SKILLS: Structured Knowledge Injection for LLM-Driven Telecommunications Operations
por: Brett, Ivo
Publicado: (2026)
por: Brett, Ivo
Publicado: (2026)
A sketch of an AI control safety case
por: Korbak, Tomek, et al.
Publicado: (2025)
por: Korbak, Tomek, et al.
Publicado: (2025)
AI security and cyber risk in IoT systems
por: Radanliev, Petar, et al.
Publicado: (2024)
por: Radanliev, Petar, et al.
Publicado: (2024)
AI Code Generators for Security: Friend or Foe?
por: Natella, Roberto, et al.
Publicado: (2024)
por: Natella, Roberto, et al.
Publicado: (2024)
Semantic-Aware Fuzzing: An Empirical Framework for LLM-Guided, Reasoning-Driven Input Mutation
por: Lu, Mengdi, et al.
Publicado: (2025)
por: Lu, Mengdi, et al.
Publicado: (2025)
Containment Verification: AI Safety Guarantees Independent of Alignment
por: Moon, Royce, et al.
Publicado: (2026)
por: Moon, Royce, et al.
Publicado: (2026)
Risks of ignoring uncertainty propagation in AI-augmented security pipelines
por: Mezzi, Emanuele, et al.
Publicado: (2024)
por: Mezzi, Emanuele, et al.
Publicado: (2024)
Identifying the Supply Chain of AI for Trustworthiness and Risk Management in Critical Applications
por: Sheh, Raymond K., et al.
Publicado: (2025)
por: Sheh, Raymond K., et al.
Publicado: (2025)
Securing the AI Frontier: Urgent Ethical and Regulatory Imperatives for AI-Driven Cybersecurity
por: Kulothungan, Vikram
Publicado: (2025)
por: Kulothungan, Vikram
Publicado: (2025)
Poisoning Programs by Un-Repairing Code: Security Concerns of AI-generated Code
por: Improta, Cristina
Publicado: (2024)
por: Improta, Cristina
Publicado: (2024)
Testing Storage-System Correctness: Challenges, Fuzzing Limitations, and AI-Augmented Opportunities
por: Wang, Ying, et al.
Publicado: (2026)
por: Wang, Ying, et al.
Publicado: (2026)
Block MedCare: Advancing healthcare through blockchain integration with AI and IoT
por: Simonoski, Oliver, et al.
Publicado: (2024)
por: Simonoski, Oliver, et al.
Publicado: (2024)
AIBoMGen: Generating an AI Bill of Materials for Secure, Transparent, and Compliant Model Training
por: Vandendriessche, Wiebe, et al.
Publicado: (2026)
por: Vandendriessche, Wiebe, et al.
Publicado: (2026)
Broken by Default: A Formal Verification Study of Security Vulnerabilities in AI-Generated Code
por: Blain, Dominik, et al.
Publicado: (2026)
por: Blain, Dominik, et al.
Publicado: (2026)
Zer0n: An AI-Assisted Vulnerability Discovery and Blockchain-Backed Integrity Framework
por: Parmar, Harshil, et al.
Publicado: (2026)
por: Parmar, Harshil, et al.
Publicado: (2026)
MalCodeAI: Autonomous Vulnerability Detection and Remediation via Language Agnostic Code Reasoning
por: Gajjar, Jugal, et al.
Publicado: (2025)
por: Gajjar, Jugal, et al.
Publicado: (2025)
Tracking Software Security Topics
por: Vu, Phong Minh, et al.
Publicado: (2024)
por: Vu, Phong Minh, et al.
Publicado: (2024)
Cyber Threat Detection and Vulnerability Assessment System using Generative AI and Large Language Model
por: M, Keerthi Kumar., et al.
Publicado: (2026)
por: M, Keerthi Kumar., et al.
Publicado: (2026)
SCoPE: Evaluating LLMs for Software Vulnerability Detection
por: Gonçalves, José, et al.
Publicado: (2024)
por: Gonçalves, José, et al.
Publicado: (2024)
Evaluating LLaMA 3.2 for Software Vulnerability Detection
por: Gonçalves, José, et al.
Publicado: (2025)
por: Gonçalves, José, et al.
Publicado: (2025)
Ejemplares similares
-
Policy-Driven AI in Dataspaces: Taxonomy, Explainability, and Pathways for Compliant Innovation
por: Chandra, Joydeep, et al.
Publicado: (2025) -
Investigating Detection and Obfuscation of Prompt Injection Attacks Against Software Reverse Engineering AI Agents
por: Crawford, Brian, et al.
Publicado: (2026) -
Data and Context Matter: Towards Generalizing AI-based Software Vulnerability Detection
por: Safdar, Rijha, et al.
Publicado: (2025) -
DevOps-Gym: Benchmarking AI Agents in Software DevOps Cycle
por: Tang, Yuheng, et al.
Publicado: (2026) -
LinuxArena: A Control Setting for AI Agents in Live Production Software Environments
por: Tracy, Tyler, et al.
Publicado: (2026)