VET Your Agent: Towards Host-Independent Autonomy via Verifiable Execution Traces
Fuente:
arXiv
Saved in:
| Main Authors: | Grigor, Artem, de Witt, Christian Schroeder, Birnbach, Simon, Martinovic, Ivan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Architecting Resilient LLM Agents: A Guide to Secure Plan-then-Execute Implementations
by: Del Rosario, Ron F., et al.
Published: (2025)
by: Del Rosario, Ron F., et al.
Published: (2025)
Secret Collusion among AI Agents: Multi-Agent Deception via Steganography
by: Motwani, Sumeet Ramesh, et al.
Published: (2024)
by: Motwani, Sumeet Ramesh, et al.
Published: (2024)
SATversary: Adversarial Attacks and Defenses for Satellite Fingerprinting
by: Smailes, Joshua, et al.
Published: (2025)
by: Smailes, Joshua, et al.
Published: (2025)
Sticky Fingers: Resilience of Satellite Fingerprinting against Jamming Attacks
by: Smailes, Joshua, et al.
Published: (2024)
by: Smailes, Joshua, et al.
Published: (2024)
Agent-Sentry: Bounding LLM Agents via Execution Provenance
by: Sequeira, Rohan, et al.
Published: (2026)
by: Sequeira, Rohan, et al.
Published: (2026)
Your LLM Agent Can Leak Your Data: Data Exfiltration via Backdoored Tool Use
by: Zhang, Wuyang, et al.
Published: (2026)
by: Zhang, Wuyang, et al.
Published: (2026)
aCAPTCHA: Verifying That an Entity Is a Capable Agent via Asymmetric Hardness
by: Xu, Zuyao, et al.
Published: (2026)
by: Xu, Zuyao, et al.
Published: (2026)
Committed SAE-Feature Traces for Audited-Session Substitution Detection in Hosted LLMs
by: Liu, Ziyang
Published: (2026)
by: Liu, Ziyang
Published: (2026)
From Thinker to Society: Security in Hierarchical Autonomy Evolution of AI Agents
by: Zhang, Xiaolei, et al.
Published: (2026)
by: Zhang, Xiaolei, et al.
Published: (2026)
KeySpace: Enhancing Public Key Infrastructure for Interplanetary Networks
by: Smailes, Joshua, et al.
Published: (2024)
by: Smailes, Joshua, et al.
Published: (2024)
VIGIL: Defending LLM Agents Against Tool Stream Injection via Verify-Before-Commit
by: Lin, Junda, et al.
Published: (2026)
by: Lin, Junda, et al.
Published: (2026)
APEX: Agent Payment Execution with Policy for Autonomous Agent API Access
by: Uddin, Mohd Safwan, et al.
Published: (2026)
by: Uddin, Mohd Safwan, et al.
Published: (2026)
Your Semantic-Independent Watermark is Fragile: A Semantic Perturbation Attack against EaaS Watermark
by: Fei, Zekun, et al.
Published: (2024)
by: Fei, Zekun, et al.
Published: (2024)
Hacking CTFs with Plain Agents
by: Turtayev, Rustem, et al.
Published: (2024)
by: Turtayev, Rustem, et al.
Published: (2024)
Towards Reinforcement Learning for Exploration of Speculative Execution Vulnerabilities
by: Lai, Evan, et al.
Published: (2025)
by: Lai, Evan, et al.
Published: (2025)
CapSeal: Capability-Sealed Secret Mediation for Secure Agent Execution
by: Jin, Shutong, et al.
Published: (2026)
by: Jin, Shutong, et al.
Published: (2026)
FinVault: Benchmarking Financial Agent Safety in Execution-Grounded Environments
by: Yang, Zhi, et al.
Published: (2026)
by: Yang, Zhi, et al.
Published: (2026)
Post-Training Local LLM Agents for Linux Privilege Escalation with Verifiable Rewards
by: Normann, Philipp, et al.
Published: (2026)
by: Normann, Philipp, et al.
Published: (2026)
Open Challenges in Multi-Agent Security: Towards Secure Systems of Interacting AI Agents
by: de Witt, Christian Schroeder, et al.
Published: (2025)
by: de Witt, Christian Schroeder, et al.
Published: (2025)
Attestable Audits: Verifiable AI Safety Benchmarks Using Trusted Execution Environments
by: Schnabl, Christoph, et al.
Published: (2025)
by: Schnabl, Christoph, et al.
Published: (2025)
NEST: Nascent Encoded Steganographic Thoughts
by: Karpov, Artem
Published: (2026)
by: Karpov, Artem
Published: (2026)
The Autonomy Tax: Defense Training Breaks LLM Agents
by: Li, Shawn, et al.
Published: (2026)
by: Li, Shawn, et al.
Published: (2026)
The Authorization-Execution Gap Is a Major Safety and Security Problem in Open-World Agents
by: Wu, Baoyuan, et al.
Published: (2026)
by: Wu, Baoyuan, et al.
Published: (2026)
Autonomous Intelligent Agents for Natural-Language-Driven Web Execution with Integrated Security Assurance
by: Pasupuleti, Vinil, et al.
Published: (2026)
by: Pasupuleti, Vinil, et al.
Published: (2026)
Red-Teaming Agent Execution Contexts: Open-World Security Evaluation on OpenClaw
by: Yao, Hongwei, et al.
Published: (2026)
by: Yao, Hongwei, et al.
Published: (2026)
HarmfulSkillBench: How Do Harmful Skills Weaponize Your Agents?
by: Jiang, Yukun, et al.
Published: (2026)
by: Jiang, Yukun, et al.
Published: (2026)
When Convenience Becomes Risk: A Semantic View of Under-Specification in Host-Acting Agents
by: Lu, Di, et al.
Published: (2026)
by: Lu, Di, et al.
Published: (2026)
AgentRAE: Remote Action Execution through Notification-based Visual Backdoors against Screenshots-based Mobile GUI Agents
by: Luo, Yutao, et al.
Published: (2026)
by: Luo, Yutao, et al.
Published: (2026)
IMMACULATE: A Practical LLM Auditing Framework via Verifiable Computation
by: Guo, Yanpei, et al.
Published: (2026)
by: Guo, Yanpei, et al.
Published: (2026)
Jolt Atlas: Verifiable Inference via Lookup Arguments in Zero Knowledge
by: Benno, Wyatt, et al.
Published: (2026)
by: Benno, Wyatt, et al.
Published: (2026)
Security Risks in Tool-Enabled AI Agents: A Systematic Analysis of Privileged Execution Environments
by: Goel, Hardik
Published: (2026)
by: Goel, Hardik
Published: (2026)
ESAA-Security: An Event-Sourced, Verifiable Architecture for Agent-Assisted Security Audits of AI-Generated Code
by: Filho, Elzo Brito dos Santos
Published: (2026)
by: Filho, Elzo Brito dos Santos
Published: (2026)
Protect Your Score: Contact Tracing With Differential Privacy Guarantees
by: Romijnders, Rob, et al.
Published: (2023)
by: Romijnders, Rob, et al.
Published: (2023)
The Art of Building Verifiers for Computer Use Agents
by: Rosset, Corby, et al.
Published: (2026)
by: Rosset, Corby, et al.
Published: (2026)
Agentic AI for Cybersecurity: A Meta-Cognitive Architecture for Governable Autonomy
by: Kojukhov, Andrei, et al.
Published: (2026)
by: Kojukhov, Andrei, et al.
Published: (2026)
Evaluating AI cyber capabilities with crowdsourced elicitation
by: Petrov, Artem, et al.
Published: (2025)
by: Petrov, Artem, et al.
Published: (2025)
Towards Effective Complementary Security Analysis using Large Language Models
by: Wagner, Jonas, et al.
Published: (2025)
by: Wagner, Jonas, et al.
Published: (2025)
Statistical Proof of Execution (SPEX)
by: Dallachiesa, Michele, et al.
Published: (2025)
by: Dallachiesa, Michele, et al.
Published: (2025)
VFEFL: Privacy-Preserving Federated Learning against Malicious Clients via Verifiable Functional Encryption
by: Cai, Nina, et al.
Published: (2025)
by: Cai, Nina, et al.
Published: (2025)
Towards Safe and Honest AI Agents with Neural Self-Other Overlap
by: Carauleanu, Marc, et al.
Published: (2024)
by: Carauleanu, Marc, et al.
Published: (2024)
Similar Items
-
Architecting Resilient LLM Agents: A Guide to Secure Plan-then-Execute Implementations
by: Del Rosario, Ron F., et al.
Published: (2025) -
Secret Collusion among AI Agents: Multi-Agent Deception via Steganography
by: Motwani, Sumeet Ramesh, et al.
Published: (2024) -
SATversary: Adversarial Attacks and Defenses for Satellite Fingerprinting
by: Smailes, Joshua, et al.
Published: (2025) -
Sticky Fingers: Resilience of Satellite Fingerprinting against Jamming Attacks
by: Smailes, Joshua, et al.
Published: (2024) -
Agent-Sentry: Bounding LLM Agents via Execution Provenance
by: Sequeira, Rohan, et al.
Published: (2026)