Optimizing Agent Planning for Security and Autonomy
Fuente:
arXiv
Salvato in:
| Autori principali: | Kolluri, Aashish, Sharma, Rishi, Costa, Manuel, Köpf, Boris, Nießen, Tobias, Russinovich, Mark, Tople, Shruti, Zanella-Béguelin, Santiago |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Securing AI Agents with Information-Flow Control
di: Costa, Manuel, et al.
Pubblicazione: (2025)
di: Costa, Manuel, et al.
Pubblicazione: (2025)
Closed-Form Bounds for DP-SGD against Record-level Inference
di: Cherubin, Giovanni, et al.
Pubblicazione: (2024)
di: Cherubin, Giovanni, et al.
Pubblicazione: (2024)
The Canary's Echo: Auditing Privacy Risks of LLM-Generated Synthetic Text
di: Meeus, Matthieu, et al.
Pubblicazione: (2025)
di: Meeus, Matthieu, et al.
Pubblicazione: (2025)
Multi-Objective Optimization for Synthetic-to-Real Style Transfer
di: Chigot, Estelle, et al.
Pubblicazione: (2026)
di: Chigot, Estelle, et al.
Pubblicazione: (2026)
AI-Driven Security Alert Screening and Alert Fatigue Mitigation in Security Operations Centers: A Survey
di: Ndichu, Samuel, et al.
Pubblicazione: (2026)
di: Ndichu, Samuel, et al.
Pubblicazione: (2026)
VulCPE: Context-Aware Cybersecurity Vulnerability Retrieval and Management
di: Jiang, Yuning, et al.
Pubblicazione: (2025)
di: Jiang, Yuning, et al.
Pubblicazione: (2025)
Towards Optimal Agentic Architectures for Offensive Security Tasks
di: David, Isaac, et al.
Pubblicazione: (2026)
di: David, Isaac, et al.
Pubblicazione: (2026)
Incentivizing Secure Software Development: the Role of Voluntary Audit and Liability Waiver
di: Huang, Ziyuan, et al.
Pubblicazione: (2024)
di: Huang, Ziyuan, et al.
Pubblicazione: (2024)
Hidden in Memory: Sleeper Memory Poisoning in LLM Agents
di: Pulipaka, Sidharth, et al.
Pubblicazione: (2026)
di: Pulipaka, Sidharth, et al.
Pubblicazione: (2026)
Inverting Cryptographic Hash Functions via Cube-and-Conquer
di: Zaikin, Oleg
Pubblicazione: (2022)
di: Zaikin, Oleg
Pubblicazione: (2022)
PhenoAuth: A Novel PUF-Phenotype-based Authentication Protocol for IoT Devices
di: Fei, Hongming, et al.
Pubblicazione: (2024)
di: Fei, Hongming, et al.
Pubblicazione: (2024)
Quantifying Return on Security Controls in LLM Systems
di: Moulton, Richard Helder, et al.
Pubblicazione: (2025)
di: Moulton, Richard Helder, et al.
Pubblicazione: (2025)
Predicting Known Vulnerabilities from Attack Descriptions Using Sentence Transformers
di: Othman, Refat
Pubblicazione: (2026)
di: Othman, Refat
Pubblicazione: (2026)
Owner-Harm: A Missing Threat Model for AI Agent Safety
di: Zhang, Dongcheng, et al.
Pubblicazione: (2026)
di: Zhang, Dongcheng, et al.
Pubblicazione: (2026)
Towards Agentic Investigation of Security Alerts
di: Eilertsen, Even, et al.
Pubblicazione: (2026)
di: Eilertsen, Even, et al.
Pubblicazione: (2026)
Threat-Oriented Digital Twinning for Security Evaluation of Autonomous Platforms
di: Neubert, Thomas J., et al.
Pubblicazione: (2026)
di: Neubert, Thomas J., et al.
Pubblicazione: (2026)
A Survey on the Security of Long-Term Memory in LLM Agents: Toward Mnemonic Sovereignty
di: Lin, Zehao, et al.
Pubblicazione: (2026)
di: Lin, Zehao, et al.
Pubblicazione: (2026)
MEMSAD: Gradient-Coupled Anomaly Detection for Memory Poisoning in Retrieval-Augmented Agents
di: Gowda, Ishrith
Pubblicazione: (2026)
di: Gowda, Ishrith
Pubblicazione: (2026)
MathLedger: A Verifiable Learning Substrate with Ledger-Attested Feedback
di: Abdullah, Ismail Ahmad
Pubblicazione: (2025)
di: Abdullah, Ismail Ahmad
Pubblicazione: (2025)
GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing
di: Zhang, Peiyan, et al.
Pubblicazione: (2025)
di: Zhang, Peiyan, et al.
Pubblicazione: (2025)
OpCode-Based Malware Classification Using Machine Learning and Deep Learning Techniques
di: Saini, Varij, et al.
Pubblicazione: (2025)
di: Saini, Varij, et al.
Pubblicazione: (2025)
KGMark: A Diffusion Watermark for Knowledge Graphs
di: Peng, Hongrui, et al.
Pubblicazione: (2025)
di: Peng, Hongrui, et al.
Pubblicazione: (2025)
David vs. Goliath: Verifiable Agent-to-Agent Jailbreaking via Reinforcement Learning
di: Nellessen, Samuel, et al.
Pubblicazione: (2026)
di: Nellessen, Samuel, et al.
Pubblicazione: (2026)
Governance Architecture for Autonomous Agent Systems: Threats, Framework, and Engineering Practice
di: Ge, Yuxu
Pubblicazione: (2026)
di: Ge, Yuxu
Pubblicazione: (2026)
IDFace: Face Template Protection for Efficient and Secure Identification
di: Kim, Sunpill, et al.
Pubblicazione: (2025)
di: Kim, Sunpill, et al.
Pubblicazione: (2025)
A Practical and Secure Byzantine Robust Aggregator
di: Lee, De Zhang, et al.
Pubblicazione: (2025)
di: Lee, De Zhang, et al.
Pubblicazione: (2025)
Securing Agentic AI Systems -- A Multilayer Security Framework
di: Arora, Sunil, et al.
Pubblicazione: (2025)
di: Arora, Sunil, et al.
Pubblicazione: (2025)
CLUE-MARK: Watermarking Diffusion Models using CLWE
di: Shehata, Kareem, et al.
Pubblicazione: (2024)
di: Shehata, Kareem, et al.
Pubblicazione: (2024)
Poison in the Well: Feature Embedding Disruption in Backdoor Attacks
di: Feng, Zhou, et al.
Pubblicazione: (2025)
di: Feng, Zhou, et al.
Pubblicazione: (2025)
AgentSentry: Mitigating Indirect Prompt Injection in LLM Agents via Temporal Causal Diagnostics and Context Purification
di: Zhang, Tian, et al.
Pubblicazione: (2026)
di: Zhang, Tian, et al.
Pubblicazione: (2026)
Security Considerations for Multi-agent Systems
di: Nguyen, Tam, et al.
Pubblicazione: (2026)
di: Nguyen, Tam, et al.
Pubblicazione: (2026)
Refusal Evaluation in Coding LLMs and Code Agents: A Systematic Review of Thirteen Malicious-Code Prompt Corpora (2023-2025)
di: Young, Richard J., et al.
Pubblicazione: (2026)
di: Young, Richard J., et al.
Pubblicazione: (2026)
Attacking Delay-based PUFs with Minimal Adversary Model
di: Fei, Hongming, et al.
Pubblicazione: (2024)
di: Fei, Hongming, et al.
Pubblicazione: (2024)
ROI: A method for identifying organizations receiving personal data
di: Rodriguez, David, et al.
Pubblicazione: (2022)
di: Rodriguez, David, et al.
Pubblicazione: (2022)
Detecting Prompt Injection Attacks Against Application Using Classifiers
di: Shaheer, Safwan, et al.
Pubblicazione: (2025)
di: Shaheer, Safwan, et al.
Pubblicazione: (2025)
Beyond the Benchmark: Innovative Defenses Against Prompt Injection Attacks
di: Shaheer, Safwan, et al.
Pubblicazione: (2025)
di: Shaheer, Safwan, et al.
Pubblicazione: (2025)
Provable Repair of Deep Neural Network Defects by Preimage Synthesis and Property Refinement
di: Ma, Jianan, et al.
Pubblicazione: (2025)
di: Ma, Jianan, et al.
Pubblicazione: (2025)
Toward Intelligent and Secure Cloud: Large Language Model Empowered Proactive Defense
di: Zhou, Yuyang, et al.
Pubblicazione: (2024)
di: Zhou, Yuyang, et al.
Pubblicazione: (2024)
Beyond Static Sandboxing: Learned Capability Governance for Autonomous AI Agents
di: Sidik, Bronislav, et al.
Pubblicazione: (2026)
di: Sidik, Bronislav, et al.
Pubblicazione: (2026)
Differentiability in infinite dimension and the Malliavin calculus
di: Bignamini, Davide A., et al.
Pubblicazione: (2023)
di: Bignamini, Davide A., et al.
Pubblicazione: (2023)
Documenti analoghi
-
Securing AI Agents with Information-Flow Control
di: Costa, Manuel, et al.
Pubblicazione: (2025) -
Closed-Form Bounds for DP-SGD against Record-level Inference
di: Cherubin, Giovanni, et al.
Pubblicazione: (2024) -
The Canary's Echo: Auditing Privacy Risks of LLM-Generated Synthetic Text
di: Meeus, Matthieu, et al.
Pubblicazione: (2025) -
Multi-Objective Optimization for Synthetic-to-Real Style Transfer
di: Chigot, Estelle, et al.
Pubblicazione: (2026) -
AI-Driven Security Alert Screening and Alert Fatigue Mitigation in Security Operations Centers: A Survey
di: Ndichu, Samuel, et al.
Pubblicazione: (2026)