Dive into the Agent Matrix: A Realistic Evaluation of Self-Replication Risk in LLM Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Boxuan, Yu, Yi, Guo, Jiaxuan, Shao, Jing |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SkillAdaptor: Self-Adapting Skills for LLM Agents from Trajectories
by: Yu, Zhuoyun, et al.
Published: (2026)
by: Yu, Zhuoyun, et al.
Published: (2026)
LLM Agents Making Agent Tools
by: Wölflein, Georg, et al.
Published: (2025)
by: Wölflein, Georg, et al.
Published: (2025)
SPIO: Ensemble and Selective Strategies via LLM-Based Multi-Agent Planning in Automated Data Science
by: Seo, Wonduk, et al.
Published: (2025)
by: Seo, Wonduk, et al.
Published: (2025)
TrustAgent: Towards Safe and Trustworthy LLM-based Agents
by: Hua, Wenyue, et al.
Published: (2024)
by: Hua, Wenyue, et al.
Published: (2024)
When AI Agents Collude Online: Financial Fraud Risks by Collaborative LLM Agents on Social Platforms
by: Ren, Qibing, et al.
Published: (2025)
by: Ren, Qibing, et al.
Published: (2025)
Can Agents Judge Systematic Reviews Like Humans? Evaluating SLRs with LLM-based Multi-Agent System
by: Mushtaq, Abdullah, et al.
Published: (2025)
by: Mushtaq, Abdullah, et al.
Published: (2025)
MUSE-Autoskill: Self-Evolving Agents via Skill Creation, Memory, Management, and Evaluation
by: Lin, Huawei, et al.
Published: (2026)
by: Lin, Huawei, et al.
Published: (2026)
Opponent Shaping in LLM Agents
by: Segura, Marta Emili Garcia, et al.
Published: (2025)
by: Segura, Marta Emili Garcia, et al.
Published: (2025)
AutoML-Agent: A Multi-Agent LLM Framework for Full-Pipeline AutoML
by: Trirat, Patara, et al.
Published: (2024)
by: Trirat, Patara, et al.
Published: (2024)
$\textit{Agents Under Siege}$: Breaking Pragmatic Multi-Agent LLM Systems with Optimized Prompt Attacks
by: Khan, Rana Muhammad Shahroz, et al.
Published: (2025)
by: Khan, Rana Muhammad Shahroz, et al.
Published: (2025)
MASPO: Joint Prompt Optimization for LLM-based Multi-Agent Systems
by: Wang, Zhexuan, et al.
Published: (2026)
by: Wang, Zhexuan, et al.
Published: (2026)
HealthFlow: A Self-Evolving AI Agent with Meta Planning for Autonomous Healthcare Research
by: Zhu, Yinghao, et al.
Published: (2025)
by: Zhu, Yinghao, et al.
Published: (2025)
AgentRec: Agent Recommendation Using Sentence Embeddings Aligned to Human Feedback
by: Park, Joshua, et al.
Published: (2025)
by: Park, Joshua, et al.
Published: (2025)
MedAgentBoard: Benchmarking Multi-Agent Collaboration with Conventional Methods for Diverse Medical Tasks
by: Zhu, Yinghao, et al.
Published: (2025)
by: Zhu, Yinghao, et al.
Published: (2025)
An Adversary-Resistant Multi-Agent LLM System via Credibility Scoring
by: Ebrahimi, Sana, et al.
Published: (2025)
by: Ebrahimi, Sana, et al.
Published: (2025)
MARCO: Multi-Agent Real-time Chat Orchestration
by: Shrimal, Anubhav, et al.
Published: (2024)
by: Shrimal, Anubhav, et al.
Published: (2024)
AgentForesight: Online Auditing for Early Failure Prediction in Multi-Agent Systems
by: Zhang, Boxuan, et al.
Published: (2026)
by: Zhang, Boxuan, et al.
Published: (2026)
GAMBIT: A Three-Mode Benchmark for Adversarial Robustness in Multi-Agent LLM Collectives
by: Mercier, Alexandre Le, et al.
Published: (2026)
by: Mercier, Alexandre Le, et al.
Published: (2026)
Collaborative Memory: Multi-User Memory Sharing in LLM Agents with Dynamic Access Control
by: Rezazadeh, Alireza, et al.
Published: (2025)
by: Rezazadeh, Alireza, et al.
Published: (2025)
Towards Reliable ML Feature Engineering via Planning in Constrained-Topology of LLM Agents
by: Thakur, Himanshu, et al.
Published: (2026)
by: Thakur, Himanshu, et al.
Published: (2026)
Towards Efficient LLM Grounding for Embodied Multi-Agent Collaboration
by: Zhang, Yang, et al.
Published: (2024)
by: Zhang, Yang, et al.
Published: (2024)
Multi-Agent Design: Optimizing Agents with Better Prompts and Topologies
by: Zhou, Han, et al.
Published: (2025)
by: Zhou, Han, et al.
Published: (2025)
Doctorina MedBench: End-to-End Evaluation of Agent-Based Medical AI
by: Kozlova, Anna, et al.
Published: (2026)
by: Kozlova, Anna, et al.
Published: (2026)
Self-Organized Agents: A LLM Multi-Agent Framework toward Ultra Large-Scale Code Generation and Optimization
by: Ishibashi, Yoichi, et al.
Published: (2024)
by: Ishibashi, Yoichi, et al.
Published: (2024)
RealICU: Do LLM Agents Understand Long-Context ICU Data? A Benchmark Beyond Behavior Imitation
by: Shen, Chengzhi, et al.
Published: (2026)
by: Shen, Chengzhi, et al.
Published: (2026)
KnowAgent: Knowledge-Augmented Planning for LLM-Based Agents
by: Zhu, Yuqi, et al.
Published: (2024)
by: Zhu, Yuqi, et al.
Published: (2024)
Verification-Aware Planning for Multi-Agent Systems
by: Xu, Tianyang, et al.
Published: (2025)
by: Xu, Tianyang, et al.
Published: (2025)
Recursive Agent Optimization
by: Gandhi, Apurva, et al.
Published: (2026)
by: Gandhi, Apurva, et al.
Published: (2026)
Exploring Collaboration Mechanisms for LLM Agents: A Social Psychology View
by: Zhang, Jintian, et al.
Published: (2023)
by: Zhang, Jintian, et al.
Published: (2023)
Memp: Exploring Agent Procedural Memory
by: Fang, Runnan, et al.
Published: (2025)
by: Fang, Runnan, et al.
Published: (2025)
MedAgentBench: A Realistic Virtual EHR Environment to Benchmark Medical LLM Agents
by: Jiang, Yixing, et al.
Published: (2025)
by: Jiang, Yixing, et al.
Published: (2025)
DR. WELL: Dynamic Reasoning and Learning with Symbolic World Model for Embodied LLM-Based Multi-Agent Collaboration
by: Nourzad, Narjes, et al.
Published: (2025)
by: Nourzad, Narjes, et al.
Published: (2025)
Language Agents as Optimizable Graphs
by: Zhuge, Mingchen, et al.
Published: (2024)
by: Zhuge, Mingchen, et al.
Published: (2024)
Can We Predict Before Executing Machine Learning Agents?
by: Zheng, Jingsheng, et al.
Published: (2026)
by: Zheng, Jingsheng, et al.
Published: (2026)
LLM-based Multi-Agent Reinforcement Learning: Current and Future Directions
by: Sun, Chuanneng, et al.
Published: (2024)
by: Sun, Chuanneng, et al.
Published: (2024)
FORGE: Self-Evolving Agent Memory With No Weight Updates via Population Broadcast
by: Bogdanov, Igor, et al.
Published: (2026)
by: Bogdanov, Igor, et al.
Published: (2026)
MAC: Multi-Agent Constitution Learning
by: Thareja, Rushil, et al.
Published: (2026)
by: Thareja, Rushil, et al.
Published: (2026)
Agent Trading Arena: A Study on Numerical Understanding in LLM-Based Agents
by: Ma, Tianmi, et al.
Published: (2025)
by: Ma, Tianmi, et al.
Published: (2025)
Symphony: A Decentralized Multi-Agent Framework for Scalable Collective Intelligence
by: Wang, Ji, et al.
Published: (2025)
by: Wang, Ji, et al.
Published: (2025)
MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
by: Zhu, Kunlun, et al.
Published: (2025)
by: Zhu, Kunlun, et al.
Published: (2025)
Similar Items
-
SkillAdaptor: Self-Adapting Skills for LLM Agents from Trajectories
by: Yu, Zhuoyun, et al.
Published: (2026) -
LLM Agents Making Agent Tools
by: Wölflein, Georg, et al.
Published: (2025) -
SPIO: Ensemble and Selective Strategies via LLM-Based Multi-Agent Planning in Automated Data Science
by: Seo, Wonduk, et al.
Published: (2025) -
TrustAgent: Towards Safe and Trustworthy LLM-based Agents
by: Hua, Wenyue, et al.
Published: (2024) -
When AI Agents Collude Online: Financial Fraud Risks by Collaborative LLM Agents on Social Platforms
by: Ren, Qibing, et al.
Published: (2025)