PatchPilot: A Cost-Efficient Software Engineering Agent with Early Attempts on Formal Verification
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Hongwei, Tang, Yuheng, Wang, Shiqi, Guo, Wenbo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Co-PatcheR: Collaborative Software Patching with Component(s)-specific Small Reasoning Models
di: Tang, Yuheng, et al.
Pubblicazione: (2025)
di: Tang, Yuheng, et al.
Pubblicazione: (2025)
WaveVerif: Acoustic Side-Channel based Verification of Robotic Workflows
di: Erdogan, Zeynep Yasemin, et al.
Pubblicazione: (2025)
di: Erdogan, Zeynep Yasemin, et al.
Pubblicazione: (2025)
DevOps-Gym: Benchmarking AI Agents in Software DevOps Cycle
di: Tang, Yuheng, et al.
Pubblicazione: (2026)
di: Tang, Yuheng, et al.
Pubblicazione: (2026)
SafeAgentBench: A Benchmark for Safe Task Planning of Embodied LLM Agents
di: Yin, Sheng, et al.
Pubblicazione: (2024)
di: Yin, Sheng, et al.
Pubblicazione: (2024)
Progent: Securing AI Agents with Privilege Control
di: Shi, Tianneng, et al.
Pubblicazione: (2025)
di: Shi, Tianneng, et al.
Pubblicazione: (2025)
ARACNE: An LLM-Based Autonomous Shell Pentesting Agent
di: Nieponice, Tomas, et al.
Pubblicazione: (2025)
di: Nieponice, Tomas, et al.
Pubblicazione: (2025)
VulnLLM-R: Specialized Reasoning LLM with Agent Scaffold for Vulnerability Detection
di: Nie, Yuzhou, et al.
Pubblicazione: (2025)
di: Nie, Yuzhou, et al.
Pubblicazione: (2025)
Agent Security Bench (ASB): Formalizing and Benchmarking Attacks and Defenses in LLM-based Agents
di: Zhang, Hanrong, et al.
Pubblicazione: (2024)
di: Zhang, Hanrong, et al.
Pubblicazione: (2024)
SoK: On the Semantic AI Security in Autonomous Driving
di: Shen, Junjie, et al.
Pubblicazione: (2022)
di: Shen, Junjie, et al.
Pubblicazione: (2022)
Towards Robust and Secure Embodied AI: A Survey on Vulnerabilities and Attacks
di: Xing, Wenpeng, et al.
Pubblicazione: (2025)
di: Xing, Wenpeng, et al.
Pubblicazione: (2025)
Automatically Attacking Software Reverse Engineering AI Agents
di: Crawford, Brian, et al.
Pubblicazione: (2026)
di: Crawford, Brian, et al.
Pubblicazione: (2026)
SilentDrift: Exploiting Action Chunking for Stealthy Backdoor Attacks on Vision-Language-Action Models
di: Xu, Bingxin, et al.
Pubblicazione: (2026)
di: Xu, Bingxin, et al.
Pubblicazione: (2026)
DropVLA: An Action-Level Backdoor Attack on Vision-Language-Action Models
di: Xu, Zonghuan, et al.
Pubblicazione: (2025)
di: Xu, Zonghuan, et al.
Pubblicazione: (2025)
A Formal Security Framework for MCP-Based AI Agents: Threat Taxonomy, Verification Models, and Defense Mechanisms
di: Acharya, Nirajan, et al.
Pubblicazione: (2026)
di: Acharya, Nirajan, et al.
Pubblicazione: (2026)
Formal Verification of Robustness and Resilience of Learning-Enabled State Estimation Systems
di: Huang, Wei, et al.
Pubblicazione: (2020)
di: Huang, Wei, et al.
Pubblicazione: (2020)
Time-Constrained Intelligent Adversaries for Automation Vulnerability Testing: A Multi-Robot Patrol Case Study
di: Ward, James C., et al.
Pubblicazione: (2025)
di: Ward, James C., et al.
Pubblicazione: (2025)
From Prompt to Physical Actuation: Holistic Threat Modeling of LLM-Enabled Robotic Systems
di: Nagaraja, Neha, et al.
Pubblicazione: (2026)
di: Nagaraja, Neha, et al.
Pubblicazione: (2026)
Not What You Asked For: Typographic Attacks in Household Robot Manipulation
di: Iranmanesh, Ali, et al.
Pubblicazione: (2026)
di: Iranmanesh, Ali, et al.
Pubblicazione: (2026)
Aportes para el cumplimiento del Reglamento (UE) 2024/1689 en robótica y sistemas autónomos
di: Lera, Francisco J. Rodríguez, et al.
Pubblicazione: (2025)
di: Lera, Francisco J. Rodríguez, et al.
Pubblicazione: (2025)
Cyber-Physical Steganography in Robotic Motion Control
di: Chang, Ching-Chun, et al.
Pubblicazione: (2025)
di: Chang, Ching-Chun, et al.
Pubblicazione: (2025)
Enhancing Privacy and Security of Autonomous UAV Navigation
di: Aggarwal, Vatsal, et al.
Pubblicazione: (2024)
di: Aggarwal, Vatsal, et al.
Pubblicazione: (2024)
Drones that Think on their Feet: Sudden Landing Decisions with Embodied AI
di: Barbosa, Diego Ortiz, et al.
Pubblicazione: (2025)
di: Barbosa, Diego Ortiz, et al.
Pubblicazione: (2025)
SeCodePLT: A Unified Platform for Evaluating the Security of Code GenAI
di: Nie, Yuzhou, et al.
Pubblicazione: (2024)
di: Nie, Yuzhou, et al.
Pubblicazione: (2024)
A Framework for Formalizing LLM Agent Security
di: Siu, Vincent, et al.
Pubblicazione: (2026)
di: Siu, Vincent, et al.
Pubblicazione: (2026)
RAT: Adversarial Attacks on Deep Reinforcement Agents for Targeted Behaviors
di: Bai, Fengshuo, et al.
Pubblicazione: (2024)
di: Bai, Fengshuo, et al.
Pubblicazione: (2024)
MELON: Provable Defense Against Indirect Prompt Injection Attacks in AI Agents
di: Zhu, Kaijie, et al.
Pubblicazione: (2025)
di: Zhu, Kaijie, et al.
Pubblicazione: (2025)
AgentVigil: Generic Black-Box Red-teaming for Indirect Prompt Injection against LLM Agents
di: Wang, Zhun, et al.
Pubblicazione: (2025)
di: Wang, Zhun, et al.
Pubblicazione: (2025)
Just-in-Time Detection of Silent Security Patches
di: Tang, Xunzhu, et al.
Pubblicazione: (2023)
di: Tang, Xunzhu, et al.
Pubblicazione: (2023)
Secure and Efficient Access Control for Computer-Use Agents via Context Space
di: Gong, Haochen, et al.
Pubblicazione: (2025)
di: Gong, Haochen, et al.
Pubblicazione: (2025)
ChainCaps: Composition-Safe Tool-Using Agents via Monotonic Capability Attenuation
di: Jiang, Xiaochong, et al.
Pubblicazione: (2026)
di: Jiang, Xiaochong, et al.
Pubblicazione: (2026)
CryptoFormalEval: Integrating LLMs and Formal Verification for Automated Cryptographic Protocol Vulnerability Detection
di: Curaba, Cristian, et al.
Pubblicazione: (2024)
di: Curaba, Cristian, et al.
Pubblicazione: (2024)
Securing LLM-Generated Embedded Firmware through AI Agent-Driven Validation and Patching
di: Abtahi, Seyed Moein, et al.
Pubblicazione: (2025)
di: Abtahi, Seyed Moein, et al.
Pubblicazione: (2025)
AttackPilot: Autonomous Inference Attacks Against ML Services With LLM-Based Agents
di: Wu, Yixin, et al.
Pubblicazione: (2025)
di: Wu, Yixin, et al.
Pubblicazione: (2025)
Backdoor Attack Against Vision Transformers via Attention Gradient-Based Image Erosion
di: Guo, Ji, et al.
Pubblicazione: (2024)
di: Guo, Ji, et al.
Pubblicazione: (2024)
Temporal UI State Inconsistency in Desktop GUI Agents: Formalizing and Defending Against TOCTOU Attacks on Computer-Use Agents
di: Xu, Wenpeng
Pubblicazione: (2026)
di: Xu, Wenpeng
Pubblicazione: (2026)
Red-Teaming Agent Execution Contexts: Open-World Security Evaluation on OpenClaw
di: Yao, Hongwei, et al.
Pubblicazione: (2026)
di: Yao, Hongwei, et al.
Pubblicazione: (2026)
The Ripple Effect: On Unforeseen Complications of Backdoor Attacks
di: Zhang, Rui, et al.
Pubblicazione: (2025)
di: Zhang, Rui, et al.
Pubblicazione: (2025)
OpenSage: Self-programming Agent Generation Engine
di: Li, Hongwei, et al.
Pubblicazione: (2026)
di: Li, Hongwei, et al.
Pubblicazione: (2026)
Behavioral Integrity Verification for AI Agent Skills
di: Wu, Yuhao, et al.
Pubblicazione: (2026)
di: Wu, Yuhao, et al.
Pubblicazione: (2026)
Onyx: Cost-Efficient Disk-Oblivious ANN Search
di: Rathee, Deevashwer, et al.
Pubblicazione: (2026)
di: Rathee, Deevashwer, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Co-PatcheR: Collaborative Software Patching with Component(s)-specific Small Reasoning Models
di: Tang, Yuheng, et al.
Pubblicazione: (2025) -
WaveVerif: Acoustic Side-Channel based Verification of Robotic Workflows
di: Erdogan, Zeynep Yasemin, et al.
Pubblicazione: (2025) -
DevOps-Gym: Benchmarking AI Agents in Software DevOps Cycle
di: Tang, Yuheng, et al.
Pubblicazione: (2026) -
SafeAgentBench: A Benchmark for Safe Task Planning of Embodied LLM Agents
di: Yin, Sheng, et al.
Pubblicazione: (2024) -
Progent: Securing AI Agents with Privilege Control
di: Shi, Tianneng, et al.
Pubblicazione: (2025)