Adversarial Feeds Steer LLM Agent Decisions Against Their Defaults
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Usman, Rana Muhammad |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Research on Security Enhancement Methods for Adversarial Robust Large Language Model Intelligent Agents for Medical Decision-Making Tasks
von: Hu, Saisai
Veröffentlicht: (2026)
von: Hu, Saisai
Veröffentlicht: (2026)
Governance Architecture for Autonomous Agent Systems: Threats, Framework, and Engineering Practice
von: Ge, Yuxu
Veröffentlicht: (2026)
von: Ge, Yuxu
Veröffentlicht: (2026)
LLM Scalability Risk for Agentic-AI and Model Supply Chain Security
von: Ahi, Kiarash, et al.
Veröffentlicht: (2026)
von: Ahi, Kiarash, et al.
Veröffentlicht: (2026)
Security Considerations for Multi-agent Systems
von: Nguyen, Tam, et al.
Veröffentlicht: (2026)
von: Nguyen, Tam, et al.
Veröffentlicht: (2026)
Beyond Benchmark Islands: Toward Representative Trustworthiness Evaluation for Agentic AI
von: Qi, Jinhu, et al.
Veröffentlicht: (2026)
von: Qi, Jinhu, et al.
Veröffentlicht: (2026)
Systematic Capability Benchmarking of Frontier Large Language Models for Offensive Cyber Tasks
von: Merves, Tyler H., et al.
Veröffentlicht: (2026)
von: Merves, Tyler H., et al.
Veröffentlicht: (2026)
Security Attack and Defense Strategies for Autonomous Agent Frameworks: A Layered Review with OpenClaw as a Case Study
von: Xu, Luyao, et al.
Veröffentlicht: (2026)
von: Xu, Luyao, et al.
Veröffentlicht: (2026)
Design Principles for the Construction of a Benchmark Evaluating Security Operation Capabilities of Multi-agent AI Systems
von: Cai, Yicheng, et al.
Veröffentlicht: (2026)
von: Cai, Yicheng, et al.
Veröffentlicht: (2026)
Towards trustworthy agentic AI: a comprehensive survey of safety, robustness, privacy, and system security
von: Qi, Jinhu, et al.
Veröffentlicht: (2026)
von: Qi, Jinhu, et al.
Veröffentlicht: (2026)
From Multi-Agent Systems and the Semantic Web to Agentic AI: A Unified Narrative of the Web of Agents
von: Petrova, Tatiana, et al.
Veröffentlicht: (2025)
von: Petrova, Tatiana, et al.
Veröffentlicht: (2025)
Depth-Dependent Indirect Prompt Injection in Tool-Calling ReAct Agents: Injection Depth, Payload Framing, and Turn-Budget Sensitivity
von: Rashidi, Mohammadreza
Veröffentlicht: (2026)
von: Rashidi, Mohammadreza
Veröffentlicht: (2026)
Hidden in Memory: Sleeper Memory Poisoning in LLM Agents
von: Pulipaka, Sidharth, et al.
Veröffentlicht: (2026)
von: Pulipaka, Sidharth, et al.
Veröffentlicht: (2026)
Refute-or-Promote: An Adversarial Stage-Gated Multi-Agent Review Methodology for High-Precision LLM-Assisted Defect Discovery
von: Agarwal, Abhinav
Veröffentlicht: (2026)
von: Agarwal, Abhinav
Veröffentlicht: (2026)
Benchmarking Large Language Models for IoC Recovery under Adversarial Code Obfuscation and Encryption
von: Morales, Jaime, et al.
Veröffentlicht: (2026)
von: Morales, Jaime, et al.
Veröffentlicht: (2026)
An Agentic Multi-Agent Architecture for Cybersecurity Risk Management
von: Gupta, Ravish, et al.
Veröffentlicht: (2026)
von: Gupta, Ravish, et al.
Veröffentlicht: (2026)
SALLIE: Safeguarding Against Latent Language & Image Exploits
von: Azov, Guy, et al.
Veröffentlicht: (2026)
von: Azov, Guy, et al.
Veröffentlicht: (2026)
Terrarium: Revisiting the Blackboard for Multi-Agent Safety, Privacy, and Security Studies
von: Nakamura, Mason, et al.
Veröffentlicht: (2025)
von: Nakamura, Mason, et al.
Veröffentlicht: (2025)
Measuring Harmfulness of Computer-Using Agents
von: Tian, Aaron Xuxiang, et al.
Veröffentlicht: (2025)
von: Tian, Aaron Xuxiang, et al.
Veröffentlicht: (2025)
Send to which account? Evaluation of an LLM-based Scambaiting System
von: Siadati, Hossein, et al.
Veröffentlicht: (2025)
von: Siadati, Hossein, et al.
Veröffentlicht: (2025)
MASH: Evading Black-Box AI-Generated Text Detectors via Style Humanization
von: Gu, Yongtong, et al.
Veröffentlicht: (2026)
von: Gu, Yongtong, et al.
Veröffentlicht: (2026)
ILION: Deterministic Pre-Execution Safety Gates for Agentic AI Systems
von: Chitan, Florin Adrian
Veröffentlicht: (2026)
von: Chitan, Florin Adrian
Veröffentlicht: (2026)
Session Risk Memory (SRM): Temporal Authorization for Deterministic Pre-Execution Safety Gates
von: Chitan, Florin Adrian
Veröffentlicht: (2026)
von: Chitan, Florin Adrian
Veröffentlicht: (2026)
Cyber Defense Benchmark: Agentic Threat Hunting Evaluation for LLMs in SecOps
von: Chona, Alankrit, et al.
Veröffentlicht: (2026)
von: Chona, Alankrit, et al.
Veröffentlicht: (2026)
Countermind: A Multi-Layered Security Architecture for Large Language Models
von: Schwarz, Dominik
Veröffentlicht: (2025)
von: Schwarz, Dominik
Veröffentlicht: (2025)
Safe and Policy-Compliant Multi-Agent Orchestration for Enterprise AI
von: Pasupuleti, Vinil, et al.
Veröffentlicht: (2026)
von: Pasupuleti, Vinil, et al.
Veröffentlicht: (2026)
HBEE: Human Behavioral Entropy Engine -- Pre-Registered Multi-Agent LLM Simulation of Peer-Suspicion-Based Detection Inversion
von: Ferrel, Vickson
Veröffentlicht: (2026)
von: Ferrel, Vickson
Veröffentlicht: (2026)
When the Agent Is the Adversary: Architectural Requirements for Agentic AI Containment After the April 2026 Frontier Model Escape
von: Mitchell, Richard Joseph
Veröffentlicht: (2026)
von: Mitchell, Richard Joseph
Veröffentlicht: (2026)
AgentSentry: Mitigating Indirect Prompt Injection in LLM Agents via Temporal Causal Diagnostics and Context Purification
von: Zhang, Tian, et al.
Veröffentlicht: (2026)
von: Zhang, Tian, et al.
Veröffentlicht: (2026)
An Organization-Scoped LLM Agent Runtime Architecture for Regulated Cybersecurity Operations
von: Fatouros, George, et al.
Veröffentlicht: (2026)
von: Fatouros, George, et al.
Veröffentlicht: (2026)
Right to History: A Sovereignty Kernel for Verifiable AI Agent Execution
von: Zhang, Jing
Veröffentlicht: (2026)
von: Zhang, Jing
Veröffentlicht: (2026)
Token-Level Generalization in LoRA Adapter Backdoors: Attack Characterization and Behavioral Detection
von: Lelle, Travis
Veröffentlicht: (2026)
von: Lelle, Travis
Veröffentlicht: (2026)
The Importance of Out-of-Band Metadata for Safe Autonomous Agents: The Redpanda Agentic Data Plane
von: Akidau, Tyler, et al.
Veröffentlicht: (2026)
von: Akidau, Tyler, et al.
Veröffentlicht: (2026)
Kill-Chain Canaries: Stage-Level Tracking of Prompt Injection Across Attack Surfaces and Model Safety Tiers
von: Wang, Haochuan Kevin, et al.
Veröffentlicht: (2026)
von: Wang, Haochuan Kevin, et al.
Veröffentlicht: (2026)
Council Mode: A Heterogeneous Multi-Agent Consensus Framework for Reducing LLM Hallucination and Bias
von: Wu, Shuai, et al.
Veröffentlicht: (2026)
von: Wu, Shuai, et al.
Veröffentlicht: (2026)
LLM Detectors Still Fall Short of Real World: Case of LLM-Generated Short News-Like Posts
von: Gameiro, Henrique Da Silva, et al.
Veröffentlicht: (2024)
von: Gameiro, Henrique Da Silva, et al.
Veröffentlicht: (2024)
Agent Capsules: Quality-Gated Granularity Control for Multi-Agent LLM Pipelines
von: Ray, Aninda
Veröffentlicht: (2026)
von: Ray, Aninda
Veröffentlicht: (2026)
Seeing the Forest through the Trees: Data Leakage from Partial Transformer Gradients
von: Li, Weijun, et al.
Veröffentlicht: (2024)
von: Li, Weijun, et al.
Veröffentlicht: (2024)
Predicting Known Vulnerabilities from Attack Descriptions Using Sentence Transformers
von: Othman, Refat
Veröffentlicht: (2026)
von: Othman, Refat
Veröffentlicht: (2026)
JPS: Jailbreak Multimodal Large Language Models with Collaborative Visual Perturbation and Textual Steering
von: Chen, Renmiao, et al.
Veröffentlicht: (2025)
von: Chen, Renmiao, et al.
Veröffentlicht: (2025)
$δ$-STEAL: LLM Stealing Attack with Local Differential Privacy
von: Dang, Kieu, et al.
Veröffentlicht: (2025)
von: Dang, Kieu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Research on Security Enhancement Methods for Adversarial Robust Large Language Model Intelligent Agents for Medical Decision-Making Tasks
von: Hu, Saisai
Veröffentlicht: (2026) -
Governance Architecture for Autonomous Agent Systems: Threats, Framework, and Engineering Practice
von: Ge, Yuxu
Veröffentlicht: (2026) -
LLM Scalability Risk for Agentic-AI and Model Supply Chain Security
von: Ahi, Kiarash, et al.
Veröffentlicht: (2026) -
Security Considerations for Multi-agent Systems
von: Nguyen, Tam, et al.
Veröffentlicht: (2026) -
Beyond Benchmark Islands: Toward Representative Trustworthiness Evaluation for Agentic AI
von: Qi, Jinhu, et al.
Veröffentlicht: (2026)