Structural Enforcement of Goal Integrity in AI Agents via Separation-of-Powers Architecture
Fuente:
arXiv
Guardado en:
| Autor principal: | Xiang, Rong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Prompts Don't Protect: Architectural Enforcement via MCP Proxy for LLM Tool Access Control
por: Uppala, Rohith
Publicado: (2026)
por: Uppala, Rohith
Publicado: (2026)
MemLineage: Lineage-Guided Enforcement for LLM Agent Memory
por: Ouyang, Ciyan, et al.
Publicado: (2026)
por: Ouyang, Ciyan, et al.
Publicado: (2026)
Behavioral Integrity Verification for AI Agent Skills
por: Wu, Yuhao, et al.
Publicado: (2026)
por: Wu, Yuhao, et al.
Publicado: (2026)
Cryptographic Runtime Governance for Autonomous AI Systems: The Aegis Architecture for Verifiable Policy Enforcement
por: Mazzocchetti, Adam Massimo
Publicado: (2026)
por: Mazzocchetti, Adam Massimo
Publicado: (2026)
Security of AI Agents
por: He, Yifeng, et al.
Publicado: (2024)
por: He, Yifeng, et al.
Publicado: (2024)
AgentWard: A Lifecycle Security Architecture for Autonomous AI Agents
por: Zhang, Yixiang, et al.
Publicado: (2026)
por: Zhang, Yixiang, et al.
Publicado: (2026)
Safeguarding AI Agents: Developing and Analyzing Safety Architectures
por: Domkundwar, Ishaan, et al.
Publicado: (2024)
por: Domkundwar, Ishaan, et al.
Publicado: (2024)
The Granularity Mismatch in Agent Security: Argument-Level Provenance Solves Enforcement and Isolates the LLM Reasoning Bottleneck
por: Fan, Linfeng, et al.
Publicado: (2026)
por: Fan, Linfeng, et al.
Publicado: (2026)
Strengthening Human-Centric Chain-of-Thought Reasoning Integrity in LLMs via a Structured Prompt Framework
por: Zhou, Jiling, et al.
Publicado: (2026)
por: Zhou, Jiling, et al.
Publicado: (2026)
AI-Governed Agent Architecture for Web-Trustworthy Tokenization of Alternative Assets
por: Borjigin, Ailiya, et al.
Publicado: (2025)
por: Borjigin, Ailiya, et al.
Publicado: (2025)
Agent Privilege Separation in OpenClaw: A Structural Defense Against Prompt Injection
por: Cheng, Darren, et al.
Publicado: (2026)
por: Cheng, Darren, et al.
Publicado: (2026)
Towards Secure Agent Skills: Architecture, Threat Taxonomy, and Security Analysis
por: Li, Zhiyuan, et al.
Publicado: (2026)
por: Li, Zhiyuan, et al.
Publicado: (2026)
Caging the Agents: A Zero Trust Security Architecture for Autonomous AI in Healthcare
por: Maiti, Saikat
Publicado: (2026)
por: Maiti, Saikat
Publicado: (2026)
The PBSAI Governance Ecosystem: A Multi-Agent AI Reference Architecture for Securing Enterprise AI Estates
por: Willis, John M.
Publicado: (2026)
por: Willis, John M.
Publicado: (2026)
Securing Generative AI in Healthcare: A Zero-Trust Architecture Powered by Confidential Computing on Google Cloud
por: Amanna, Adaobi, et al.
Publicado: (2025)
por: Amanna, Adaobi, et al.
Publicado: (2025)
Goal-Driven Risk Assessment for LLM-Powered Systems: A Healthcare Case Study
por: Nagaraja, Neha, et al.
Publicado: (2026)
por: Nagaraja, Neha, et al.
Publicado: (2026)
From Prompt Injections to Protocol Exploits: Threats in LLM-Powered AI Agents Workflows
por: Ferrag, Mohamed Amine, et al.
Publicado: (2025)
por: Ferrag, Mohamed Amine, et al.
Publicado: (2025)
Meta-Sealing: A Revolutionizing Integrity Assurance Protocol for Transparent, Tamper-Proof, and Trustworthy AI System
por: Krishnamoorthy, Mahesh Vaijainthymala
Publicado: (2024)
por: Krishnamoorthy, Mahesh Vaijainthymala
Publicado: (2024)
ESAA-Security: An Event-Sourced, Verifiable Architecture for Agent-Assisted Security Audits of AI-Generated Code
por: Filho, Elzo Brito dos Santos
Publicado: (2026)
por: Filho, Elzo Brito dos Santos
Publicado: (2026)
Secret Collusion among AI Agents: Multi-Agent Deception via Steganography
por: Motwani, Sumeet Ramesh, et al.
Publicado: (2024)
por: Motwani, Sumeet Ramesh, et al.
Publicado: (2024)
Open Digital Rights Enforcement Framework (ODRE): from descriptive to enforceable policies
por: Cimmino, Andrea, et al.
Publicado: (2024)
por: Cimmino, Andrea, et al.
Publicado: (2024)
Enforcing Benign Trajectories: A Behavioral Firewall for Structured-Workflow AI Agents
por: Dang, Hung
Publicado: (2026)
por: Dang, Hung
Publicado: (2026)
Bypassing AI Control Protocols via Agent-as-a-Proxy Attacks
por: Isbarov, Jafar, et al.
Publicado: (2026)
por: Isbarov, Jafar, et al.
Publicado: (2026)
Prompt Flow Integrity to Prevent Privilege Escalation in LLM Agents
por: Kim, Juhee, et al.
Publicado: (2025)
por: Kim, Juhee, et al.
Publicado: (2025)
SplitAgent: A Privacy-Preserving Distributed Architecture for Enterprise-Cloud Agent Collaboration
por: She, Jianshu
Publicado: (2026)
por: She, Jianshu
Publicado: (2026)
AgenticCyber: A GenAI-Powered Multi-Agent System for Multimodal Threat Detection and Adaptive Response in Cybersecurity
por: Roy, Shovan
Publicado: (2025)
por: Roy, Shovan
Publicado: (2025)
AI Agents Under Threat: A Survey of Key Security Challenges and Future Pathways
por: Deng, Zehang, et al.
Publicado: (2024)
por: Deng, Zehang, et al.
Publicado: (2024)
Enabling Regulatory Multi-Agent Collaboration: Architecture, Challenges, and Solutions
por: Hu, Qinnan, et al.
Publicado: (2025)
por: Hu, Qinnan, et al.
Publicado: (2025)
On the Impossibility of Separating Intelligence from Judgment: The Computational Intractability of Filtering for AI Alignment
por: Ball, Sarah, et al.
Publicado: (2025)
por: Ball, Sarah, et al.
Publicado: (2025)
Magika: AI-Powered Content-Type Detection
por: Fratantonio, Yanick, et al.
Publicado: (2024)
por: Fratantonio, Yanick, et al.
Publicado: (2024)
LLM Agent Honeypot: Monitoring AI Hacking Agents in the Wild
por: Reworr, et al.
Publicado: (2024)
por: Reworr, et al.
Publicado: (2024)
Do Vision-Language Models Respect Contextual Integrity in Location Disclosure?
por: Yang, Ruixin, et al.
Publicado: (2026)
por: Yang, Ruixin, et al.
Publicado: (2026)
Scalable GPU-Based Integrity Verification for Large Machine Learning Models
por: Spoczynski, Marcin, et al.
Publicado: (2025)
por: Spoczynski, Marcin, et al.
Publicado: (2025)
AI Identity: Standards, Gaps, and Research Directions for AI Agents
por: Otsuka, Takumi, et al.
Publicado: (2026)
por: Otsuka, Takumi, et al.
Publicado: (2026)
aCAPTCHA: Verifying That an Entity Is a Capable Agent via Asymmetric Hardness
por: Xu, Zuyao, et al.
Publicado: (2026)
por: Xu, Zuyao, et al.
Publicado: (2026)
Architecting Secure AI Agents: Perspectives on System-Level Defenses Against Indirect Prompt Injection Attacks
por: Xiang, Chong, et al.
Publicado: (2026)
por: Xiang, Chong, et al.
Publicado: (2026)
The Hidden Dangers of Browsing AI Agents
por: Mudryi, Mykyta, et al.
Publicado: (2025)
por: Mudryi, Mykyta, et al.
Publicado: (2025)
AgentWall: A Runtime Safety Layer for Local AI Agents
por: Aravind, Ashwin
Publicado: (2026)
por: Aravind, Ashwin
Publicado: (2026)
AudAgent: Automated Auditing of Privacy Policy Compliance in AI Agents
por: Zheng, Ye, et al.
Publicado: (2025)
por: Zheng, Ye, et al.
Publicado: (2025)
Blind PRNG Hijacking: An Undetectable Integrity-Preserving Attack Against LLM Watermarking
por: You, Ziyang, et al.
Publicado: (2026)
por: You, Ziyang, et al.
Publicado: (2026)
Ejemplares similares
-
Prompts Don't Protect: Architectural Enforcement via MCP Proxy for LLM Tool Access Control
por: Uppala, Rohith
Publicado: (2026) -
MemLineage: Lineage-Guided Enforcement for LLM Agent Memory
por: Ouyang, Ciyan, et al.
Publicado: (2026) -
Behavioral Integrity Verification for AI Agent Skills
por: Wu, Yuhao, et al.
Publicado: (2026) -
Cryptographic Runtime Governance for Autonomous AI Systems: The Aegis Architecture for Verifiable Policy Enforcement
por: Mazzocchetti, Adam Massimo
Publicado: (2026) -
Security of AI Agents
por: He, Yifeng, et al.
Publicado: (2024)