Prompt Flow Integrity to Prevent Privilege Escalation in LLM Agents
Fuente:
arXiv
Salvato in:
| Autori principali: | Kim, Juhee, Choi, Woohyuk, Lee, Byoungyoung |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Prompt Infection: LLM-to-LLM Prompt Injection within Multi-Agent Systems
di: Lee, Donghyun, et al.
Pubblicazione: (2024)
di: Lee, Donghyun, et al.
Pubblicazione: (2024)
Ambient Persuasion in a Deployed AI Agent: Unauthorized Escalation Following Routine Non-Adversarial Content Exposure
di: Cuadros, Diego F., et al.
Pubblicazione: (2026)
di: Cuadros, Diego F., et al.
Pubblicazione: (2026)
A Vision for Access Control in LLM-based Agent Systems
di: Li, Xinfeng, et al.
Pubblicazione: (2025)
di: Li, Xinfeng, et al.
Pubblicazione: (2025)
Web Fraud Attacks Against LLM-Driven Multi-Agent Systems
di: Kong, Dezhang, et al.
Pubblicazione: (2025)
di: Kong, Dezhang, et al.
Pubblicazione: (2025)
CTFExplorer: Evaluating LLM Offensive Agents Through Multi-Target Web CTF Benchmarking
di: Rani, Nanda, et al.
Pubblicazione: (2026)
di: Rani, Nanda, et al.
Pubblicazione: (2026)
Attack the Messages, Not the Agents: A Multi-round Adaptive Stealthy Tampering Framework for LLM-MAS
di: Yan, Bingyu, et al.
Pubblicazione: (2025)
di: Yan, Bingyu, et al.
Pubblicazione: (2025)
GAMMAF: A Common Framework for Graph-Based Anomaly Monitoring Benchmarking in LLM Multi-Agent Systems
di: Mateo-Torrejón, Pablo, et al.
Pubblicazione: (2026)
di: Mateo-Torrejón, Pablo, et al.
Pubblicazione: (2026)
Explainable and Fine-Grained Safeguarding of LLM Multi-Agent Systems via Bi-Level Graph Anomaly Detection
di: Pan, Junjun, et al.
Pubblicazione: (2025)
di: Pan, Junjun, et al.
Pubblicazione: (2025)
Protecting Context and Prompts: Deterministic Security for Non-Deterministic AI
di: Rajagopalan, Mohan, et al.
Pubblicazione: (2026)
di: Rajagopalan, Mohan, et al.
Pubblicazione: (2026)
Beyond Single-Agent Alignment: Preventing Context-Fragmented Violations in Multi-Agent Systems
di: Wu, Jie, et al.
Pubblicazione: (2026)
di: Wu, Jie, et al.
Pubblicazione: (2026)
Who Owns This Agent? Tracing AI Agents Back to Their Owners
di: Chocron, Ruben, et al.
Pubblicazione: (2026)
di: Chocron, Ruben, et al.
Pubblicazione: (2026)
Agents for Agents: An Interrogator-Based Secure Framework for Autonomous Internet of Underwater Things
di: Akarma, Ali, et al.
Pubblicazione: (2026)
di: Akarma, Ali, et al.
Pubblicazione: (2026)
Open Challenges in Multi-Agent Security: Towards Secure Systems of Interacting AI Agents
di: de Witt, Christian Schroeder, et al.
Pubblicazione: (2025)
di: de Witt, Christian Schroeder, et al.
Pubblicazione: (2025)
Trusted AI Agents in the Cloud
di: Bodea, Teofil, et al.
Pubblicazione: (2025)
di: Bodea, Teofil, et al.
Pubblicazione: (2025)
SentinelAgent: Intent-Verified Delegation Chains for Securing Federal Multi-Agent AI Systems
di: Patil, KrishnaSaiReddy
Pubblicazione: (2026)
di: Patil, KrishnaSaiReddy
Pubblicazione: (2026)
Agent Capability Negotiation and Binding Protocol (ACNBP)
di: Huang, Ken, et al.
Pubblicazione: (2025)
di: Huang, Ken, et al.
Pubblicazione: (2025)
The Art of Building Verifiers for Computer Use Agents
di: Rosset, Corby, et al.
Pubblicazione: (2026)
di: Rosset, Corby, et al.
Pubblicazione: (2026)
Agent Name Service (ANS): A Proof-of-Concept Trust Layer for Secure AI Agent Discovery, Identity, and Governance in Kubernetes
di: Mittal, Akshay, et al.
Pubblicazione: (2026)
di: Mittal, Akshay, et al.
Pubblicazione: (2026)
Chronology of Multi-Agent Interactions for Provenance of Evolving Information
di: Chang, Ching-Chun, et al.
Pubblicazione: (2025)
di: Chang, Ching-Chun, et al.
Pubblicazione: (2025)
Multi-Agent Actor-Critics in Autonomous Cyber Defense
di: Wang, Mingjun, et al.
Pubblicazione: (2024)
di: Wang, Mingjun, et al.
Pubblicazione: (2024)
Towards Log Analysis with AI Agents: Cowrie Case Study
di: Karaarslan, Enis, et al.
Pubblicazione: (2025)
di: Karaarslan, Enis, et al.
Pubblicazione: (2025)
MAS-Shield: A Defense Framework for Secure and Efficient LLM MAS
di: Wang, Kaixiang, et al.
Pubblicazione: (2025)
di: Wang, Kaixiang, et al.
Pubblicazione: (2025)
LLM-Agent-UMF: LLM-based Agent Unified Modeling Framework for Seamless Design of Multi Active/Passive Core-Agent Architectures
di: Hassouna, Amine Ben, et al.
Pubblicazione: (2024)
di: Hassouna, Amine Ben, et al.
Pubblicazione: (2024)
The Aegis Protocol: A Foundational Security Framework for Autonomous AI Agents
di: Adapala, Sai Teja Reddy, et al.
Pubblicazione: (2025)
di: Adapala, Sai Teja Reddy, et al.
Pubblicazione: (2025)
Heartbeat-Bound Hierarchical Credentials: Cryptographic Revocation for AI Agent Swarms
di: Deochake, Saurabh
Pubblicazione: (2026)
di: Deochake, Saurabh
Pubblicazione: (2026)
LegalSim: Multi-Agent Simulation of Legal Systems for Discovering Procedural Exploits
di: Badhe, Sanket
Pubblicazione: (2025)
di: Badhe, Sanket
Pubblicazione: (2025)
From Cloud-Native to Trust-Native: A Protocol for Verifiable Multi-Agent Systems
di: Li, Muyang
Pubblicazione: (2025)
di: Li, Muyang
Pubblicazione: (2025)
AutoRISE: Agent-Driven Strategy Evolution for Red-Teaming Large Language Models
di: Gautam, Tanmay, et al.
Pubblicazione: (2026)
di: Gautam, Tanmay, et al.
Pubblicazione: (2026)
Digital Identity for Agentic Systems: Toward a Portable Authorization Standard for Autonomous Agents
di: Madhira, Partha
Pubblicazione: (2026)
di: Madhira, Partha
Pubblicazione: (2026)
OrchJail: Jailbreaking Tool-Calling Text-to-Image Agents by Orchestration-Guided Fuzzing
di: Chen, Jianming, et al.
Pubblicazione: (2026)
di: Chen, Jianming, et al.
Pubblicazione: (2026)
SV-LLM: An Agentic Approach for SoC Security Verification using Large Language Models
di: Saha, Dipayan, et al.
Pubblicazione: (2025)
di: Saha, Dipayan, et al.
Pubblicazione: (2025)
Formal Policy Enforcement for Real-World Agentic Systems
di: Palumbo, Nils, et al.
Pubblicazione: (2026)
di: Palumbo, Nils, et al.
Pubblicazione: (2026)
Depending on yourself when you should: Mentoring LLM with RL agents to become the master in cybersecurity games
di: Yan, Yikuan, et al.
Pubblicazione: (2024)
di: Yan, Yikuan, et al.
Pubblicazione: (2024)
CRAKEN: Cybersecurity LLM Agent with Knowledge-Based Execution
di: Shao, Minghao, et al.
Pubblicazione: (2025)
di: Shao, Minghao, et al.
Pubblicazione: (2025)
Aegis: Towards Governance, Integrity, and Security of AI Voice Agents
di: Li, Xiang, et al.
Pubblicazione: (2026)
di: Li, Xiang, et al.
Pubblicazione: (2026)
Wolfpack Adversarial Attack for Robust Multi-Agent Reinforcement Learning
di: Lee, Sunwoo, et al.
Pubblicazione: (2025)
di: Lee, Sunwoo, et al.
Pubblicazione: (2025)
Practical challenges of control monitoring in frontier AI deployments
di: Lindner, David, et al.
Pubblicazione: (2025)
di: Lindner, David, et al.
Pubblicazione: (2025)
Formalizing the Safety, Security, and Functional Properties of Agentic AI Systems
di: Allegrini, Edoardo, et al.
Pubblicazione: (2025)
di: Allegrini, Edoardo, et al.
Pubblicazione: (2025)
A Novel Zero-Trust Identity Framework for Agentic AI: Decentralized Authentication and Fine-Grained Access Control
di: Huang, Ken, et al.
Pubblicazione: (2025)
di: Huang, Ken, et al.
Pubblicazione: (2025)
Privacy-Utility-Fairness: A Balanced Approach to Vehicular-Traffic Management System
di: Sengupta, Poushali, et al.
Pubblicazione: (2025)
di: Sengupta, Poushali, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Prompt Infection: LLM-to-LLM Prompt Injection within Multi-Agent Systems
di: Lee, Donghyun, et al.
Pubblicazione: (2024) -
Ambient Persuasion in a Deployed AI Agent: Unauthorized Escalation Following Routine Non-Adversarial Content Exposure
di: Cuadros, Diego F., et al.
Pubblicazione: (2026) -
A Vision for Access Control in LLM-based Agent Systems
di: Li, Xinfeng, et al.
Pubblicazione: (2025) -
Web Fraud Attacks Against LLM-Driven Multi-Agent Systems
di: Kong, Dezhang, et al.
Pubblicazione: (2025) -
CTFExplorer: Evaluating LLM Offensive Agents Through Multi-Target Web CTF Benchmarking
di: Rani, Nanda, et al.
Pubblicazione: (2026)