The Art of Building Verifiers for Computer Use Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Rosset, Corby, Sharma, Pratyusha, Zhao, Andrew, Gonzalez-Fernandez, Miguel, Awadallah, Ahmed |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SentinelAgent: Intent-Verified Delegation Chains for Securing Federal Multi-Agent AI Systems
by: Patil, KrishnaSaiReddy
Published: (2026)
by: Patil, KrishnaSaiReddy
Published: (2026)
From Cloud-Native to Trust-Native: A Protocol for Verifiable Multi-Agent Systems
by: Li, Muyang
Published: (2025)
by: Li, Muyang
Published: (2025)
TessPay: Verify-then-Pay Infrastructure for Trusted Agentic Commerce
by: Goenka, Mehul, et al.
Published: (2026)
by: Goenka, Mehul, et al.
Published: (2026)
Web Fraud Attacks Against LLM-Driven Multi-Agent Systems
by: Kong, Dezhang, et al.
Published: (2025)
by: Kong, Dezhang, et al.
Published: (2025)
Who Owns This Agent? Tracing AI Agents Back to Their Owners
by: Chocron, Ruben, et al.
Published: (2026)
by: Chocron, Ruben, et al.
Published: (2026)
Agents for Agents: An Interrogator-Based Secure Framework for Autonomous Internet of Underwater Things
by: Akarma, Ali, et al.
Published: (2026)
by: Akarma, Ali, et al.
Published: (2026)
Open Challenges in Multi-Agent Security: Towards Secure Systems of Interacting AI Agents
by: de Witt, Christian Schroeder, et al.
Published: (2025)
by: de Witt, Christian Schroeder, et al.
Published: (2025)
Trusted AI Agents in the Cloud
by: Bodea, Teofil, et al.
Published: (2025)
by: Bodea, Teofil, et al.
Published: (2025)
Agent Capability Negotiation and Binding Protocol (ACNBP)
by: Huang, Ken, et al.
Published: (2025)
by: Huang, Ken, et al.
Published: (2025)
Agent Name Service (ANS): A Proof-of-Concept Trust Layer for Secure AI Agent Discovery, Identity, and Governance in Kubernetes
by: Mittal, Akshay, et al.
Published: (2026)
by: Mittal, Akshay, et al.
Published: (2026)
Multi-Agent Actor-Critics in Autonomous Cyber Defense
by: Wang, Mingjun, et al.
Published: (2024)
by: Wang, Mingjun, et al.
Published: (2024)
Chronology of Multi-Agent Interactions for Provenance of Evolving Information
by: Chang, Ching-Chun, et al.
Published: (2025)
by: Chang, Ching-Chun, et al.
Published: (2025)
Skills as Verifiable Artifacts: A Trust Schema and a Biconditional Correctness Criterion for Human-in-the-Loop Agent Runtimes
by: Metere, Alfredo
Published: (2026)
by: Metere, Alfredo
Published: (2026)
Prompt Flow Integrity to Prevent Privilege Escalation in LLM Agents
by: Kim, Juhee, et al.
Published: (2025)
by: Kim, Juhee, et al.
Published: (2025)
Towards Log Analysis with AI Agents: Cowrie Case Study
by: Karaarslan, Enis, et al.
Published: (2025)
by: Karaarslan, Enis, et al.
Published: (2025)
A Vision for Access Control in LLM-based Agent Systems
by: Li, Xinfeng, et al.
Published: (2025)
by: Li, Xinfeng, et al.
Published: (2025)
Heartbeat-Bound Hierarchical Credentials: Cryptographic Revocation for AI Agent Swarms
by: Deochake, Saurabh
Published: (2026)
by: Deochake, Saurabh
Published: (2026)
The Aegis Protocol: A Foundational Security Framework for Autonomous AI Agents
by: Adapala, Sai Teja Reddy, et al.
Published: (2025)
by: Adapala, Sai Teja Reddy, et al.
Published: (2025)
LegalSim: Multi-Agent Simulation of Legal Systems for Discovering Procedural Exploits
by: Badhe, Sanket
Published: (2025)
by: Badhe, Sanket
Published: (2025)
Prompt Infection: LLM-to-LLM Prompt Injection within Multi-Agent Systems
by: Lee, Donghyun, et al.
Published: (2024)
by: Lee, Donghyun, et al.
Published: (2024)
Beyond DNS: Unlocking the Internet of AI Agents via the NANDA Index and Verified AgentFacts
by: Raskar, Ramesh, et al.
Published: (2025)
by: Raskar, Ramesh, et al.
Published: (2025)
CTFExplorer: Evaluating LLM Offensive Agents Through Multi-Target Web CTF Benchmarking
by: Rani, Nanda, et al.
Published: (2026)
by: Rani, Nanda, et al.
Published: (2026)
AutoRISE: Agent-Driven Strategy Evolution for Red-Teaming Large Language Models
by: Gautam, Tanmay, et al.
Published: (2026)
by: Gautam, Tanmay, et al.
Published: (2026)
Digital Identity for Agentic Systems: Toward a Portable Authorization Standard for Autonomous Agents
by: Madhira, Partha
Published: (2026)
by: Madhira, Partha
Published: (2026)
OrchJail: Jailbreaking Tool-Calling Text-to-Image Agents by Orchestration-Guided Fuzzing
by: Chen, Jianming, et al.
Published: (2026)
by: Chen, Jianming, et al.
Published: (2026)
Attack the Messages, Not the Agents: A Multi-round Adaptive Stealthy Tampering Framework for LLM-MAS
by: Yan, Bingyu, et al.
Published: (2025)
by: Yan, Bingyu, et al.
Published: (2025)
GAMMAF: A Common Framework for Graph-Based Anomaly Monitoring Benchmarking in LLM Multi-Agent Systems
by: Mateo-Torrejón, Pablo, et al.
Published: (2026)
by: Mateo-Torrejón, Pablo, et al.
Published: (2026)
Ambient Persuasion in a Deployed AI Agent: Unauthorized Escalation Following Routine Non-Adversarial Content Exposure
by: Cuadros, Diego F., et al.
Published: (2026)
by: Cuadros, Diego F., et al.
Published: (2026)
Explainable and Fine-Grained Safeguarding of LLM Multi-Agent Systems via Bi-Level Graph Anomaly Detection
by: Pan, Junjun, et al.
Published: (2025)
by: Pan, Junjun, et al.
Published: (2025)
CuDA2: An approach for Incorporating Traitor Agents into Cooperative Multi-Agent Systems
by: Chen, Zhen, et al.
Published: (2024)
by: Chen, Zhen, et al.
Published: (2024)
AI Agents with Decentralized Identifiers and Verifiable Credentials
by: Garzon, Sandro Rodriguez, et al.
Published: (2025)
by: Garzon, Sandro Rodriguez, et al.
Published: (2025)
Architectural Obsolescence of Unhardened Agentic-AI Runtimes
by: Metere, Alfredo
Published: (2026)
by: Metere, Alfredo
Published: (2026)
enclawed: A Configurable, Sector-Neutral Hardening Framework for Single-User AI Assistant Gateways
by: Metere, Alfredo
Published: (2026)
by: Metere, Alfredo
Published: (2026)
Governance-Constrained Agentic AI: Blockchain-Enforced Human Oversight for Safety-Critical Wildfire Monitoring
by: Akarma, Ali, et al.
Published: (2026)
by: Akarma, Ali, et al.
Published: (2026)
Protecting Context and Prompts: Deterministic Security for Non-Deterministic AI
by: Rajagopalan, Mohan, et al.
Published: (2026)
by: Rajagopalan, Mohan, et al.
Published: (2026)
Public and private blockchain for decentralized digital building twins and building automation system
by: Ly, Reachsak, et al.
Published: (2026)
by: Ly, Reachsak, et al.
Published: (2026)
Formal Policy Enforcement for Real-World Agentic Systems
by: Palumbo, Nils, et al.
Published: (2026)
by: Palumbo, Nils, et al.
Published: (2026)
RAVEN: Retrieval-Augmented Vulnerability Exploration Network for Memory Corruption Analysis in User Code and Binary Programs
by: Jamwal, Parteek, et al.
Published: (2026)
by: Jamwal, Parteek, et al.
Published: (2026)
Practical challenges of control monitoring in frontier AI deployments
by: Lindner, David, et al.
Published: (2025)
by: Lindner, David, et al.
Published: (2025)
Depending on yourself when you should: Mentoring LLM with RL agents to become the master in cybersecurity games
by: Yan, Yikuan, et al.
Published: (2024)
by: Yan, Yikuan, et al.
Published: (2024)
Similar Items
-
SentinelAgent: Intent-Verified Delegation Chains for Securing Federal Multi-Agent AI Systems
by: Patil, KrishnaSaiReddy
Published: (2026) -
From Cloud-Native to Trust-Native: A Protocol for Verifiable Multi-Agent Systems
by: Li, Muyang
Published: (2025) -
TessPay: Verify-then-Pay Infrastructure for Trusted Agentic Commerce
by: Goenka, Mehul, et al.
Published: (2026) -
Web Fraud Attacks Against LLM-Driven Multi-Agent Systems
by: Kong, Dezhang, et al.
Published: (2025) -
Who Owns This Agent? Tracing AI Agents Back to Their Owners
by: Chocron, Ruben, et al.
Published: (2026)