Owner-Harm: A Missing Threat Model for AI Agent Safety
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Dongcheng, Jiang, Yiqing |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hidden in Memory: Sleeper Memory Poisoning in LLM Agents
by: Pulipaka, Sidharth, et al.
Published: (2026)
by: Pulipaka, Sidharth, et al.
Published: (2026)
HDP: A Lightweight Cryptographic Protocol for Human Delegation Provenance in Agentic AI Systems
by: Dalugoda, Asiri
Published: (2026)
by: Dalugoda, Asiri
Published: (2026)
Towards Optimal Agentic Architectures for Offensive Security Tasks
by: David, Isaac, et al.
Published: (2026)
by: David, Isaac, et al.
Published: (2026)
Governance Architecture for Autonomous Agent Systems: Threats, Framework, and Engineering Practice
by: Ge, Yuxu
Published: (2026)
by: Ge, Yuxu
Published: (2026)
Security Considerations for Multi-agent Systems
by: Nguyen, Tam, et al.
Published: (2026)
by: Nguyen, Tam, et al.
Published: (2026)
ILION: Deterministic Pre-Execution Safety Gates for Agentic AI Systems
by: Chitan, Florin Adrian
Published: (2026)
by: Chitan, Florin Adrian
Published: (2026)
Session Risk Memory (SRM): Temporal Authorization for Deterministic Pre-Execution Safety Gates
by: Chitan, Florin Adrian
Published: (2026)
by: Chitan, Florin Adrian
Published: (2026)
David vs. Goliath: Verifiable Agent-to-Agent Jailbreaking via Reinforcement Learning
by: Nellessen, Samuel, et al.
Published: (2026)
by: Nellessen, Samuel, et al.
Published: (2026)
Right to History: A Sovereignty Kernel for Verifiable AI Agent Execution
by: Zhang, Jing
Published: (2026)
by: Zhang, Jing
Published: (2026)
ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree
by: Koc, Vincent, et al.
Published: (2026)
by: Koc, Vincent, et al.
Published: (2026)
Securing Agentic AI Systems -- A Multilayer Security Framework
by: Arora, Sunil, et al.
Published: (2025)
by: Arora, Sunil, et al.
Published: (2025)
HBEE: Human Behavioral Entropy Engine -- Pre-Registered Multi-Agent LLM Simulation of Peer-Suspicion-Based Detection Inversion
by: Ferrel, Vickson
Published: (2026)
by: Ferrel, Vickson
Published: (2026)
Ontology-Constrained Neural Reasoning in Enterprise Agentic Systems: A Neurosymbolic Architecture for Domain-Grounded AI Agents
by: Tuan, Thanh Luong, et al.
Published: (2026)
by: Tuan, Thanh Luong, et al.
Published: (2026)
When the Agent Is the Adversary: Architectural Requirements for Agentic AI Containment After the April 2026 Frontier Model Escape
by: Mitchell, Richard Joseph
Published: (2026)
by: Mitchell, Richard Joseph
Published: (2026)
$\mathsf{OPA}$: One-shot Private Aggregation with Single Client Interaction and its Applications to Federated Learning
by: Karthikeyan, Harish, et al.
Published: (2024)
by: Karthikeyan, Harish, et al.
Published: (2024)
Advancing Multimodal Agent Reasoning with Long-Term Neuro-Symbolic Memory
by: Jiang, Rongjie, et al.
Published: (2026)
by: Jiang, Rongjie, et al.
Published: (2026)
From Multi-Agent Systems and the Semantic Web to Agentic AI: A Unified Narrative of the Web of Agents
by: Petrova, Tatiana, et al.
Published: (2025)
by: Petrova, Tatiana, et al.
Published: (2025)
AIP: Agent Identity Protocol for Verifiable Delegation Across MCP and A2A
by: Prakash, Sunil
Published: (2026)
by: Prakash, Sunil
Published: (2026)
Harnessing Embodied Agents: Runtime Governance for Policy-Constrained Execution
by: Qin, Xue, et al.
Published: (2026)
by: Qin, Xue, et al.
Published: (2026)
Who Governs the Machine? A Machine Identity Governance Taxonomy (MIGT) for AI Systems Operating Across Enterprise and Geopolitical Boundaries
by: Kurtz, Andrew, et al.
Published: (2026)
by: Kurtz, Andrew, et al.
Published: (2026)
Formal Analysis and Supply Chain Security for Agentic AI Skills
by: Bhardwaj, Varun Pratap
Published: (2026)
by: Bhardwaj, Varun Pratap
Published: (2026)
Knowledge Equivalence in Digital Twins of Intelligent Systems
by: Zhang, Nan, et al.
Published: (2022)
by: Zhang, Nan, et al.
Published: (2022)
AI Bill of Materials and Beyond: Systematizing Security Assurance through the AI Risk Scanning (AIRS) Framework
by: Nathanson, Samuel, et al.
Published: (2025)
by: Nathanson, Samuel, et al.
Published: (2025)
Unveiling Hidden Threats: Using Fractal Triggers to Boost Stealthiness of Distributed Backdoor Attacks in Federated Learning
by: Wang, Jian, et al.
Published: (2025)
by: Wang, Jian, et al.
Published: (2025)
Faramesh: A Protocol-Agnostic Execution Control Plane for Autonomous Agent Systems
by: Fatmi, Amjad
Published: (2026)
by: Fatmi, Amjad
Published: (2026)
FSFM: A Biologically-Inspired Framework for Selective Forgetting of Agent Memory
by: Gu, Yingjie, et al.
Published: (2026)
by: Gu, Yingjie, et al.
Published: (2026)
Identity Management for Agentic AI: The new frontier of authorization, authentication, and security for an AI agent world
by: South, Tobin, et al.
Published: (2025)
by: South, Tobin, et al.
Published: (2025)
Beyond Static Sandboxing: Learned Capability Governance for Autonomous AI Agents
by: Sidik, Bronislav, et al.
Published: (2026)
by: Sidik, Bronislav, et al.
Published: (2026)
Blind Gods and Broken Screens: Architecting a Secure, Intent-Centric Mobile Agent Operating System
by: Zou, Zhenhua, et al.
Published: (2026)
by: Zou, Zhenhua, et al.
Published: (2026)
Secure Decentralized Learning with Blockchain
by: Zhang, Xiaoxue, et al.
Published: (2023)
by: Zhang, Xiaoxue, et al.
Published: (2023)
Rethinking AI Hardware: A Three-Layer Cognitive Architecture for Autonomous Agents
by: Chen, Li
Published: (2026)
by: Chen, Li
Published: (2026)
ChargingBoul: A Competitive Negotiating Agent with Novel Opponent Modeling
by: Shymanski, Joe
Published: (2025)
by: Shymanski, Joe
Published: (2025)
MFH: A Multi-faceted Heuristic Algorithm Selection Approach for Software Verification
by: Su, Jie, et al.
Published: (2025)
by: Su, Jie, et al.
Published: (2025)
Multi-Paradigm Agent Interaction in Practice:A Systematic Analysis of Generator-Evaluator, ReAct Loop,and Adversarial Evaluation in the buddyMe Framework
by: Wang, Xiaohua, et al.
Published: (2026)
by: Wang, Xiaohua, et al.
Published: (2026)
SafetyDrift: Predicting When AI Agents Cross the Line Before They Actually Do
by: Dhodapkar, Aditya, et al.
Published: (2026)
by: Dhodapkar, Aditya, et al.
Published: (2026)
Large Language Models for Agentic NetOps and AIOps: Architectures, Evaluation, and Safety
by: Bilal, Muhammad, et al.
Published: (2026)
by: Bilal, Muhammad, et al.
Published: (2026)
Context Engineering: From Prompts to Corporate Multi-Agent Architecture
by: Vishnyakova, Vera V.
Published: (2026)
by: Vishnyakova, Vera V.
Published: (2026)
FundaPod: A Multi-Persona Agent Pod Platform with Knowledge Graph Memory for AI-Assisted Fundamental Investment Research
by: Zhu, Di, et al.
Published: (2026)
by: Zhu, Di, et al.
Published: (2026)
AegisShield: Democratizing Cyber Threat Modeling with Generative AI
by: Grofsky, Matthew
Published: (2025)
by: Grofsky, Matthew
Published: (2025)
Enabling Transparent Cyber Threat Intelligence Combining Large Language Models and Domain Ontologies
by: Cotti, Luca, et al.
Published: (2025)
by: Cotti, Luca, et al.
Published: (2025)
Similar Items
-
Hidden in Memory: Sleeper Memory Poisoning in LLM Agents
by: Pulipaka, Sidharth, et al.
Published: (2026) -
HDP: A Lightweight Cryptographic Protocol for Human Delegation Provenance in Agentic AI Systems
by: Dalugoda, Asiri
Published: (2026) -
Towards Optimal Agentic Architectures for Offensive Security Tasks
by: David, Isaac, et al.
Published: (2026) -
Governance Architecture for Autonomous Agent Systems: Threats, Framework, and Engineering Practice
by: Ge, Yuxu
Published: (2026) -
Security Considerations for Multi-agent Systems
by: Nguyen, Tam, et al.
Published: (2026)