Securing AI Agents with Information-Flow Control
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Costa, Manuel, Köpf, Boris, Kolluri, Aashish, Paverd, Andrew, Russinovich, Mark, Salem, Ahmed, Tople, Shruti, Wutschitz, Lukas, Zanella-Béguelin, Santiago |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Closed-Form Bounds for DP-SGD against Record-level Inference
von: Cherubin, Giovanni, et al.
Veröffentlicht: (2024)
von: Cherubin, Giovanni, et al.
Veröffentlicht: (2024)
Optimizing Agent Planning for Security and Autonomy
von: Kolluri, Aashish, et al.
Veröffentlicht: (2026)
von: Kolluri, Aashish, et al.
Veröffentlicht: (2026)
The Canary's Echo: Auditing Privacy Risks of LLM-Generated Synthetic Text
von: Meeus, Matthieu, et al.
Veröffentlicht: (2025)
von: Meeus, Matthieu, et al.
Veröffentlicht: (2025)
MetaBackdoor: Exploiting Positional Encoding as a Backdoor Attack Surface in LLMs
von: Wen, Rui, et al.
Veröffentlicht: (2026)
von: Wen, Rui, et al.
Veröffentlicht: (2026)
Jailbreaking is (Mostly) Simpler Than You Think
von: Russinovich, Mark, et al.
Veröffentlicht: (2025)
von: Russinovich, Mark, et al.
Veröffentlicht: (2025)
Hey, That's My Model! Introducing Chain & Hash, An LLM Fingerprinting Technique
von: Russinovich, Mark, et al.
Veröffentlicht: (2024)
von: Russinovich, Mark, et al.
Veröffentlicht: (2024)
Permissive Information-Flow Analysis for Large Language Models
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2024)
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2024)
Obliviate: Efficient Unmemorization for Protecting Intellectual Property in Large Language Models
von: Russinovich, Mark, et al.
Veröffentlicht: (2025)
von: Russinovich, Mark, et al.
Veröffentlicht: (2025)
Great, Now Write an Article About That: The Crescendo Multi-Turn LLM Jailbreak Attack
von: Russinovich, Mark, et al.
Veröffentlicht: (2024)
von: Russinovich, Mark, et al.
Veröffentlicht: (2024)
Stateless Yet Not Forgetful: Implicit Memory as a Hidden Channel in LLMs
von: Salem, Ahmed, et al.
Veröffentlicht: (2026)
von: Salem, Ahmed, et al.
Veröffentlicht: (2026)
A Practical and Secure Byzantine Robust Aggregator
von: Lee, De Zhang, et al.
Veröffentlicht: (2025)
von: Lee, De Zhang, et al.
Veröffentlicht: (2025)
CLUE-MARK: Watermarking Diffusion Models using CLWE
von: Shehata, Kareem, et al.
Veröffentlicht: (2024)
von: Shehata, Kareem, et al.
Veröffentlicht: (2024)
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks
von: Bullwinkel, Blake, et al.
Veröffentlicht: (2025)
von: Bullwinkel, Blake, et al.
Veröffentlicht: (2025)
A Systematization of Security Vulnerabilities in Computer Use Agents
von: Jones, Daniel, et al.
Veröffentlicht: (2025)
von: Jones, Daniel, et al.
Veröffentlicht: (2025)
Attacking Byzantine Robust Aggregation in High Dimensions
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2023)
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2023)
Beyond Membership: Limitations of Add/Remove Adjacency in Differential Privacy
von: Pradhan, Gauri, et al.
Veröffentlicht: (2025)
von: Pradhan, Gauri, et al.
Veröffentlicht: (2025)
Invariant Aggregator for Defending against Federated Backdoor Attacks
von: Wang, Xiaoyang, et al.
Veröffentlicht: (2022)
von: Wang, Xiaoyang, et al.
Veröffentlicht: (2022)
Get my drift? Catching LLM Task Drift with Activation Deltas
von: Abdelnabi, Sahar, et al.
Veröffentlicht: (2024)
von: Abdelnabi, Sahar, et al.
Veröffentlicht: (2024)
Investigating the Effect of Misalignment on Membership Privacy in the White-box Setting
von: Cretu, Ana-Maria, et al.
Veröffentlicht: (2023)
von: Cretu, Ana-Maria, et al.
Veröffentlicht: (2023)
Toward Securing AI Agents Like Operating Systems
von: Pirch, Lukas, et al.
Veröffentlicht: (2026)
von: Pirch, Lukas, et al.
Veröffentlicht: (2026)
CI-Work: Benchmarking Contextual Integrity in Enterprise LLM Agents
von: Fu, Wenjie, et al.
Veröffentlicht: (2026)
von: Fu, Wenjie, et al.
Veröffentlicht: (2026)
Design Patterns for Securing LLM Agents against Prompt Injections
von: Beurer-Kellner, Luca, et al.
Veröffentlicht: (2025)
von: Beurer-Kellner, Luca, et al.
Veröffentlicht: (2025)
A Systematic Review of Algorithmic Red Teaming Methodologies for Assurance and Security of AI Applications
von: Srivastava, Shruti, et al.
Veröffentlicht: (2026)
von: Srivastava, Shruti, et al.
Veröffentlicht: (2026)
Hidden Risks: The Centralization of NFT Metadata and What It Means for the Market
von: Salem, Hamza, et al.
Veröffentlicht: (2024)
von: Salem, Hamza, et al.
Veröffentlicht: (2024)
AgentCrypt: Advancing Privacy and (Secure) Computation in AI Agent Collaboration
von: Karthikeyan, Harish, et al.
Veröffentlicht: (2025)
von: Karthikeyan, Harish, et al.
Veröffentlicht: (2025)
Security of AI Agents
von: He, Yifeng, et al.
Veröffentlicht: (2024)
von: He, Yifeng, et al.
Veröffentlicht: (2024)
IBAC Mathematics and Mechanics: The Case for 'Integer Based Access Control' of Data Security in the Age of AI and AI Automation
von: Stocks, Mark
Veröffentlicht: (2024)
von: Stocks, Mark
Veröffentlicht: (2024)
ConVerse: Benchmarking Contextual Safety in Agent-to-Agent Conversations
von: Gomaa, Amr, et al.
Veröffentlicht: (2025)
von: Gomaa, Amr, et al.
Veröffentlicht: (2025)
Cybersecurity AI: The World's Top AI Agent for Security Capture-the-Flag (CTF)
von: Mayoral-Vilches, Víctor, et al.
Veröffentlicht: (2025)
von: Mayoral-Vilches, Víctor, et al.
Veröffentlicht: (2025)
CSUM: A Novel Mechanism for Updating CubeSat while Preserving Authenticity and Integrity
von: Gangwal, Ankit, et al.
Veröffentlicht: (2024)
von: Gangwal, Ankit, et al.
Veröffentlicht: (2024)
VerifiableFL: Verifiable Claims for Federated Learning using Exclaves
von: Guo, Jinnan, et al.
Veröffentlicht: (2024)
von: Guo, Jinnan, et al.
Veröffentlicht: (2024)
Prevalence of Security and Privacy Risk-Inducing Usage of AI-based Conversational Agents
von: Grosse, Kathrin, et al.
Veröffentlicht: (2025)
von: Grosse, Kathrin, et al.
Veröffentlicht: (2025)
Implementation of Entropically Secure Encryption: Securing Personal Health Data
von: Temel, Mehmet Hüseyin, et al.
Veröffentlicht: (2024)
von: Temel, Mehmet Hüseyin, et al.
Veröffentlicht: (2024)
Position: Mind the Gap-AI Security and the Limits of Current Reporting Standards
von: Bieringer, Lukas, et al.
Veröffentlicht: (2024)
von: Bieringer, Lukas, et al.
Veröffentlicht: (2024)
De-authentication using Ambient Light Sensor
von: Gangwal, Ankit, et al.
Veröffentlicht: (2023)
von: Gangwal, Ankit, et al.
Veröffentlicht: (2023)
DoomArena: A framework for Testing AI Agents Against Evolving Security Threats
von: Boisvert, Leo, et al.
Veröffentlicht: (2025)
von: Boisvert, Leo, et al.
Veröffentlicht: (2025)
LLMail-Inject: A Dataset from a Realistic Adaptive Prompt Injection Challenge
von: Abdelnabi, Sahar, et al.
Veröffentlicht: (2025)
von: Abdelnabi, Sahar, et al.
Veröffentlicht: (2025)
Progent: Securing AI Agents with Privilege Control
von: Shi, Tianneng, et al.
Veröffentlicht: (2025)
von: Shi, Tianneng, et al.
Veröffentlicht: (2025)
Secure Goal-Oriented Communication: Defending against Eavesdropping Timing Attacks
von: Mason, Federico, et al.
Veröffentlicht: (2025)
von: Mason, Federico, et al.
Veröffentlicht: (2025)
The AI Security Zugzwang
von: Alevizos, Lampis
Veröffentlicht: (2025)
von: Alevizos, Lampis
Veröffentlicht: (2025)
Ähnliche Einträge
-
Closed-Form Bounds for DP-SGD against Record-level Inference
von: Cherubin, Giovanni, et al.
Veröffentlicht: (2024) -
Optimizing Agent Planning for Security and Autonomy
von: Kolluri, Aashish, et al.
Veröffentlicht: (2026) -
The Canary's Echo: Auditing Privacy Risks of LLM-Generated Synthetic Text
von: Meeus, Matthieu, et al.
Veröffentlicht: (2025) -
MetaBackdoor: Exploiting Positional Encoding as a Backdoor Attack Surface in LLMs
von: Wen, Rui, et al.
Veröffentlicht: (2026) -
Jailbreaking is (Mostly) Simpler Than You Think
von: Russinovich, Mark, et al.
Veröffentlicht: (2025)