ProofAgent Harness: Open Infrastructure for Adversarial Evaluation of AI Agents
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Bousetouane, Fouad |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Agentic Systems: A Guide to Transforming Industries with Vertical AI Agents
von: Bousetouane, Fouad
Veröffentlicht: (2025)
von: Bousetouane, Fouad
Veröffentlicht: (2025)
Physical AI Agents: Integrating Cognitive Intelligence with Real-World Action
von: Bousetouane, Fouad
Veröffentlicht: (2025)
von: Bousetouane, Fouad
Veröffentlicht: (2025)
AI Agents Need Memory Control Over More Context
von: Bousetouane, Fouad
Veröffentlicht: (2026)
von: Bousetouane, Fouad
Veröffentlicht: (2026)
Coral Protocol: Open Infrastructure Connecting The Internet of Agents
von: Georgio, Roman J., et al.
Veröffentlicht: (2025)
von: Georgio, Roman J., et al.
Veröffentlicht: (2025)
AI-Olympics: Exploring the Generalization of Agents through Open Competitions
von: Wang, Chen, et al.
Veröffentlicht: (2024)
von: Wang, Chen, et al.
Veröffentlicht: (2024)
Camouflage Adversarial Attacks on Multiple Agent Systems
von: Lu, Ziqing, et al.
Veröffentlicht: (2024)
von: Lu, Ziqing, et al.
Veröffentlicht: (2024)
Strategic Infrastructure Design via Multi-Agent Congestion Games with Joint Placement and Pricing
von: Aminikalibar, Niloofar, et al.
Veröffentlicht: (2026)
von: Aminikalibar, Niloofar, et al.
Veröffentlicht: (2026)
Fault Tolerant Multi-Agent Learning with Adversarial Budget Constraints
von: Mguni, David, et al.
Veröffentlicht: (2025)
von: Mguni, David, et al.
Veröffentlicht: (2025)
Adversarial Attack on Black-Box Multi-Agent by Adaptive Perturbation
von: Chen, Jianming, et al.
Veröffentlicht: (2025)
von: Chen, Jianming, et al.
Veröffentlicht: (2025)
EpochX: Building the Infrastructure for an Emergent Agent Civilization
von: Wang, Huacan, et al.
Veröffentlicht: (2026)
von: Wang, Huacan, et al.
Veröffentlicht: (2026)
HadAgent: Harness-Aware Decentralized Agentic AI Serving with Proof-of-Inference Blockchain Consensus
von: Jimenez, Landy, et al.
Veröffentlicht: (2026)
von: Jimenez, Landy, et al.
Veröffentlicht: (2026)
S-Agents: Self-organizing Agents in Open-ended Environments
von: Chen, Jiaqi, et al.
Veröffentlicht: (2024)
von: Chen, Jiaqi, et al.
Veröffentlicht: (2024)
OpenLens AI: Fully Autonomous Research Agent for Health Infomatics
von: Cheng, Yuxiao, et al.
Veröffentlicht: (2025)
von: Cheng, Yuxiao, et al.
Veröffentlicht: (2025)
RAI: Flexible Agent Framework for Embodied AI
von: Rachwał, Kajetan, et al.
Veröffentlicht: (2025)
von: Rachwał, Kajetan, et al.
Veröffentlicht: (2025)
Hierarchical Pedagogical Oversight: A Multi-Agent Adversarial Framework for Reliable AI Tutoring
von: Sadhu, Saisab, et al.
Veröffentlicht: (2025)
von: Sadhu, Saisab, et al.
Veröffentlicht: (2025)
Evaluating Collective Behaviour of Hundreds of LLM Agents
von: Willis, Richard, et al.
Veröffentlicht: (2026)
von: Willis, Richard, et al.
Veröffentlicht: (2026)
TeamFusion: Supporting Open-ended Teamwork with Multi-Agent Systems
von: Liu, Jiale, et al.
Veröffentlicht: (2026)
von: Liu, Jiale, et al.
Veröffentlicht: (2026)
Octopus Protocol: One-Shot Hardware Discovery and Control for AI Agents via Infrastructure-as-Prompts
von: Simeon, Quilee, et al.
Veröffentlicht: (2026)
von: Simeon, Quilee, et al.
Veröffentlicht: (2026)
Generative AI Agents for Controllable and Protected Content Creation
von: Khan, Haris, et al.
Veröffentlicht: (2026)
von: Khan, Haris, et al.
Veröffentlicht: (2026)
Mesh Memory Protocol: Semantic Infrastructure for Multi-Agent LLM Systems
von: Xu, Hongwei
Veröffentlicht: (2026)
von: Xu, Hongwei
Veröffentlicht: (2026)
Chameleon: Adaptive Adversarial Agents for Scaling-Based Visual Prompt Injection in Multimodal AI Systems
von: Zeeshan, M, et al.
Veröffentlicht: (2025)
von: Zeeshan, M, et al.
Veröffentlicht: (2025)
Enhancing Computational Efficiency in NetLogo: Best Practices for Running Large-Scale Agent-Based Models on AWS and Cloud Infrastructures
von: Duprey, Michael A., et al.
Veröffentlicht: (2026)
von: Duprey, Michael A., et al.
Veröffentlicht: (2026)
Open Challenges in Multi-Agent Security: Towards Secure Systems of Interacting AI Agents
von: de Witt, Christian Schroeder, et al.
Veröffentlicht: (2025)
von: de Witt, Christian Schroeder, et al.
Veröffentlicht: (2025)
Agent Exchange: Shaping the Future of AI Agent Economics
von: Yang, Yingxuan, et al.
Veröffentlicht: (2025)
von: Yang, Yingxuan, et al.
Veröffentlicht: (2025)
Symbiotic Cooperation for Web Agents: Harnessing Complementary Strengths of Large and Small LLMs
von: Zhang, Ruichen, et al.
Veröffentlicht: (2025)
von: Zhang, Ruichen, et al.
Veröffentlicht: (2025)
STLGame: Signal Temporal Logic Games in Adversarial Multi-Agent Systems
von: Yang, Shuo, et al.
Veröffentlicht: (2024)
von: Yang, Shuo, et al.
Veröffentlicht: (2024)
Agent Name Service (ANS): A Proof-of-Concept Trust Layer for Secure AI Agent Discovery, Identity, and Governance in Kubernetes
von: Mittal, Akshay, et al.
Veröffentlicht: (2026)
von: Mittal, Akshay, et al.
Veröffentlicht: (2026)
Evaluating Multi-Agent LLM Architectures for Rare Disease Diagnosis
von: Almasoud, Ahmed
Veröffentlicht: (2026)
von: Almasoud, Ahmed
Veröffentlicht: (2026)
Provably Stable Multi-Agent Routing with Bounded-Delay Adversaries in the Decision Loop
von: Francos, Roee M., et al.
Veröffentlicht: (2025)
von: Francos, Roee M., et al.
Veröffentlicht: (2025)
Second Order State Hallucinations for Adversarial Attack Mitigation in Formation Control of Multi-Agent Systems
von: Patel, Laksh, et al.
Veröffentlicht: (2025)
von: Patel, Laksh, et al.
Veröffentlicht: (2025)
Governing What the EU AI Act Excludes: Accountability for Autonomous AI Agents in Smart City Critical Infrastructure
von: Butt, Talal Ashraf, et al.
Veröffentlicht: (2026)
von: Butt, Talal Ashraf, et al.
Veröffentlicht: (2026)
ROMA: Recursive Open Meta-Agent Framework for Long-Horizon Multi-Agent Systems
von: Alzu'bi, Salaheddin, et al.
Veröffentlicht: (2026)
von: Alzu'bi, Salaheddin, et al.
Veröffentlicht: (2026)
Multi-Agent Craftax: Benchmarking Open-Ended Multi-Agent Reinforcement Learning at the Hyperscale
von: Omari, Bassel Al, et al.
Veröffentlicht: (2025)
von: Omari, Bassel Al, et al.
Veröffentlicht: (2025)
Externalization in LLM Agents: A Unified Review of Memory, Skills, Protocols and Harness Engineering
von: Zhou, Chenyu, et al.
Veröffentlicht: (2026)
von: Zhou, Chenyu, et al.
Veröffentlicht: (2026)
Can AI Agents Agree?
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2026)
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2026)
When AI Agents Disagree Like Humans: Reasoning Trace Analysis for Human-AI Collaborative Moderation
von: Wawer, Michał, et al.
Veröffentlicht: (2026)
von: Wawer, Michał, et al.
Veröffentlicht: (2026)
TAMAS: Benchmarking Adversarial Risks in Multi-Agent LLM Systems
von: Kavathekar, Ishan, et al.
Veröffentlicht: (2025)
von: Kavathekar, Ishan, et al.
Veröffentlicht: (2025)
Design and Evaluation of Multi-Agent AI Oracle Systems for Prediction Market Resolution
von: Kota, Tarun
Veröffentlicht: (2026)
von: Kota, Tarun
Veröffentlicht: (2026)
Agent Contracts: A Formal Framework for Resource-Bounded Autonomous AI Systems
von: Ye, Qing, et al.
Veröffentlicht: (2026)
von: Ye, Qing, et al.
Veröffentlicht: (2026)
Free Agent in Agent-Based Mixture-of-Experts Generative AI Framework
von: Liu, Jung-Hua
Veröffentlicht: (2025)
von: Liu, Jung-Hua
Veröffentlicht: (2025)
Ähnliche Einträge
-
Agentic Systems: A Guide to Transforming Industries with Vertical AI Agents
von: Bousetouane, Fouad
Veröffentlicht: (2025) -
Physical AI Agents: Integrating Cognitive Intelligence with Real-World Action
von: Bousetouane, Fouad
Veröffentlicht: (2025) -
AI Agents Need Memory Control Over More Context
von: Bousetouane, Fouad
Veröffentlicht: (2026) -
Coral Protocol: Open Infrastructure Connecting The Internet of Agents
von: Georgio, Roman J., et al.
Veröffentlicht: (2025) -
AI-Olympics: Exploring the Generalization of Agents through Open Competitions
von: Wang, Chen, et al.
Veröffentlicht: (2024)