WESE: Weak Exploration to Strong Exploitation for LLM Agents
Fuente:
arXiv
Salvato in:
| Autori principali: | Huang, Xu, Liu, Weiwen, Chen, Xiaolong, Wang, Xingmei, Lian, Defu, Wang, Yasheng, Tang, Ruiming, Chen, Enhong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Understanding the planning of LLM agents: A survey
di: Huang, Xu, et al.
Pubblicazione: (2024)
di: Huang, Xu, et al.
Pubblicazione: (2024)
Exploration of LLM Multi-Agent Application Implementation Based on LangGraph+CrewAI
di: Duan, Zhihua, et al.
Pubblicazione: (2024)
di: Duan, Zhihua, et al.
Pubblicazione: (2024)
Holos: A Web-Scale LLM-Based Multi-Agent System for the Agentic Web
di: Nie, Xiaohang, et al.
Pubblicazione: (2026)
di: Nie, Xiaohang, et al.
Pubblicazione: (2026)
AgentArk: Distilling Multi-Agent Intelligence into a Single LLM Agent
di: Luo, Yinyi, et al.
Pubblicazione: (2026)
di: Luo, Yinyi, et al.
Pubblicazione: (2026)
HALO: Hierarchical Autonomous Logic-Oriented Orchestration for Multi-Agent LLM Systems
di: Hou, Zhipeng, et al.
Pubblicazione: (2025)
di: Hou, Zhipeng, et al.
Pubblicazione: (2025)
AIvilization v0: Toward Large-Scale Artificial Social Simulation with a Unified Agent Architecture and Adaptive Agent Profiles
di: Fan, Wenkai, et al.
Pubblicazione: (2026)
di: Fan, Wenkai, et al.
Pubblicazione: (2026)
AgentSchool: An LLM-Powered Multi-Agent Simulation for Education
di: Ye, Yulei, et al.
Pubblicazione: (2026)
di: Ye, Yulei, et al.
Pubblicazione: (2026)
Shapley-Coop: Credit Assignment for Emergent Cooperation in Self-Interested LLM Agents
di: Hua, Yun, et al.
Pubblicazione: (2025)
di: Hua, Yun, et al.
Pubblicazione: (2025)
Evolve as a Team: Collaborative Self-Evolution for LLM-based Multi-Agent Systems
di: Hao, Zhezheng, et al.
Pubblicazione: (2026)
di: Hao, Zhezheng, et al.
Pubblicazione: (2026)
MedSentry: Understanding and Mitigating Safety Risks in Medical LLM Multi-Agent Systems
di: Chen, Kai, et al.
Pubblicazione: (2025)
di: Chen, Kai, et al.
Pubblicazione: (2025)
Stop Wandering, Find the Keys: LLMs Discriminate Key States for Efficient Multi-Agent Exploration
di: Qu, Yun, et al.
Pubblicazione: (2024)
di: Qu, Yun, et al.
Pubblicazione: (2024)
Towards CausalGPT: A Multi-Agent Approach for Faithful Knowledge Reasoning via Promoting Causal Consistency in LLMs
di: Tang, Ziyi, et al.
Pubblicazione: (2023)
di: Tang, Ziyi, et al.
Pubblicazione: (2023)
Understanding the Information Propagation Effects of Communication Topologies in LLM-based Multi-Agent Systems
di: Shen, Xu, et al.
Pubblicazione: (2025)
di: Shen, Xu, et al.
Pubblicazione: (2025)
Understanding Individual Agent Importance in Multi-Agent System via Counterfactual Reasoning
di: Chen, Jianming, et al.
Pubblicazione: (2024)
di: Chen, Jianming, et al.
Pubblicazione: (2024)
SafeSieve: From Heuristics to Experience in Progressive Pruning for LLM-based Multi-Agent Communication
di: Zhang, Ruijia, et al.
Pubblicazione: (2025)
di: Zhang, Ruijia, et al.
Pubblicazione: (2025)
ToMPO: Training LLM Strategic Decision Making from a Multi-Agent Perspective
di: Zhang, Yiwen, et al.
Pubblicazione: (2025)
di: Zhang, Yiwen, et al.
Pubblicazione: (2025)
MetaCogAgent: A Metacognitive Multi-Agent LLM Framework with Self-Aware Task Delegation
di: Wang, Chenyu, et al.
Pubblicazione: (2026)
di: Wang, Chenyu, et al.
Pubblicazione: (2026)
Who's the Mole? Modeling and Detecting Intention-Hiding Malicious Agents in LLM-Based Multi-Agent Systems
di: Xie, Yizhe, et al.
Pubblicazione: (2025)
di: Xie, Yizhe, et al.
Pubblicazione: (2025)
COOP$^2$: Defining, Observing, and Repairing Cooperation in LLM Multi-Agent Systems
di: Yang, Hanqing, et al.
Pubblicazione: (2026)
di: Yang, Hanqing, et al.
Pubblicazione: (2026)
PartnerMAS: An LLM Hierarchical Multi-Agent Framework for Business Partner Selection on High-Dimensional Features
di: Li, Lingyao, et al.
Pubblicazione: (2025)
di: Li, Lingyao, et al.
Pubblicazione: (2025)
MAS-on-the-Fly: Dynamic Adaptation of LLM-based Multi-Agent Systems at Test Time
di: Liu, Guangyi, et al.
Pubblicazione: (2026)
di: Liu, Guangyi, et al.
Pubblicazione: (2026)
Heterogeneous Interaction Modeling With Reduced Accumulated Error for Multi-Agent Trajectory Prediction
di: Chen, Siyuan, et al.
Pubblicazione: (2024)
di: Chen, Siyuan, et al.
Pubblicazione: (2024)
Mesh Memory Protocol: Semantic Infrastructure for Multi-Agent LLM Systems
di: Xu, Hongwei
Pubblicazione: (2026)
di: Xu, Hongwei
Pubblicazione: (2026)
Where Did It Go Wrong? Capability-Oriented Failure Attribution for Vision-and-Language Navigation Agents
di: Chen, Jianming, et al.
Pubblicazione: (2026)
di: Chen, Jianming, et al.
Pubblicazione: (2026)
AgentLite: A Lightweight Library for Building and Advancing Task-Oriented LLM Agent System
di: Liu, Zhiwei, et al.
Pubblicazione: (2024)
di: Liu, Zhiwei, et al.
Pubblicazione: (2024)
Position: Towards a Responsible LLM-empowered Multi-Agent Systems
di: Hu, Jinwei, et al.
Pubblicazione: (2025)
di: Hu, Jinwei, et al.
Pubblicazione: (2025)
PIMAEX: Multi-Agent Exploration through Peer Incentivization
di: Kölle, Michael, et al.
Pubblicazione: (2025)
di: Kölle, Michael, et al.
Pubblicazione: (2025)
MACC: Multi-Agent Collaborative Competition for Scientific Exploration
di: Oyama, Satoshi, et al.
Pubblicazione: (2026)
di: Oyama, Satoshi, et al.
Pubblicazione: (2026)
Imagine, Initialize, and Explore: An Effective Exploration Method in Multi-Agent Reinforcement Learning
di: Liu, Zeyang, et al.
Pubblicazione: (2024)
di: Liu, Zeyang, et al.
Pubblicazione: (2024)
LLM Multi-Agent Systems: Challenges and Open Problems
di: Han, Shanshan, et al.
Pubblicazione: (2024)
di: Han, Shanshan, et al.
Pubblicazione: (2024)
MGCBS: An Optimal and Efficient Algorithm for Solving Multi-Goal Multi-Agent Path Finding Problem
di: Tang, Mingkai, et al.
Pubblicazione: (2024)
di: Tang, Mingkai, et al.
Pubblicazione: (2024)
Coevolving with the Other You: Fine-Tuning LLM with Sequential Cooperative Multi-Agent Reinforcement Learning
di: Ma, Hao, et al.
Pubblicazione: (2024)
di: Ma, Hao, et al.
Pubblicazione: (2024)
Memory-Augmented Reinforcement Learning Agent for CAD Generation
di: Xiaolong, Yin, et al.
Pubblicazione: (2026)
di: Xiaolong, Yin, et al.
Pubblicazione: (2026)
Hera: Learning Long-Horizon Coordination for Device-Cloud Collaborative LLM Agents
di: Zhang, Yuxin, et al.
Pubblicazione: (2026)
di: Zhang, Yuxin, et al.
Pubblicazione: (2026)
GenSim: A General Social Simulation Platform with Large Language Model based Agents
di: Tang, Jiakai, et al.
Pubblicazione: (2024)
di: Tang, Jiakai, et al.
Pubblicazione: (2024)
SAGE: Multi-Agent Self-Evolution for LLM Reasoning
di: Peng, Yulin, et al.
Pubblicazione: (2026)
di: Peng, Yulin, et al.
Pubblicazione: (2026)
Gradientsys: A Multi-Agent LLM Scheduler with ReAct Orchestration
di: Song, Xinyuan, et al.
Pubblicazione: (2025)
di: Song, Xinyuan, et al.
Pubblicazione: (2025)
Exposing Weak Links in Multi-Agent Systems under Adversarial Prompting
di: Arora, Nirmit, et al.
Pubblicazione: (2025)
di: Arora, Nirmit, et al.
Pubblicazione: (2025)
Interpreting Emergent Extreme Events in Multi-Agent Systems
di: Tang, Ling, et al.
Pubblicazione: (2026)
di: Tang, Ling, et al.
Pubblicazione: (2026)
Insider Attacks in Multi-Agent LLM Consensus Systems
di: Sun, Xiaolin, et al.
Pubblicazione: (2026)
di: Sun, Xiaolin, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Understanding the planning of LLM agents: A survey
di: Huang, Xu, et al.
Pubblicazione: (2024) -
Exploration of LLM Multi-Agent Application Implementation Based on LangGraph+CrewAI
di: Duan, Zhihua, et al.
Pubblicazione: (2024) -
Holos: A Web-Scale LLM-Based Multi-Agent System for the Agentic Web
di: Nie, Xiaohang, et al.
Pubblicazione: (2026) -
AgentArk: Distilling Multi-Agent Intelligence into a Single LLM Agent
di: Luo, Yinyi, et al.
Pubblicazione: (2026) -
HALO: Hierarchical Autonomous Logic-Oriented Orchestration for Multi-Agent LLM Systems
di: Hou, Zhipeng, et al.
Pubblicazione: (2025)