Training Generalizable Collaborative Agents via Strategic Risk Aversion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qu, Chengrui, Zhang, Yizhou, Lanzetti, Nicolas, Mazumdar, Eric |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Distributionally Robust Cooperative Multi-Agent Reinforcement Learning via Robust Value Factorization
von: Qu, Chengrui, et al.
Veröffentlicht: (2026)
von: Qu, Chengrui, et al.
Veröffentlicht: (2026)
Provably Convergent Actor-Critic for MARL through Risk-aversion
von: Zhang, Yizhou, et al.
Veröffentlicht: (2026)
von: Zhang, Yizhou, et al.
Veröffentlicht: (2026)
Tractable Equilibrium Computation in Markov Games through Risk Aversion
von: Mazumdar, Eric, et al.
Veröffentlicht: (2024)
von: Mazumdar, Eric, et al.
Veröffentlicht: (2024)
Multi-Agent Path Finding via Offline RL and LLM Collaboration
von: Atasever, Merve, et al.
Veröffentlicht: (2025)
von: Atasever, Merve, et al.
Veröffentlicht: (2025)
CORD: Generalizable Cooperation via Role Diversity
von: Matsuyama, Kanefumi, et al.
Veröffentlicht: (2025)
von: Matsuyama, Kanefumi, et al.
Veröffentlicht: (2025)
Language Agents with Reinforcement Learning for Strategic Play in the Werewolf Game
von: Xu, Zelai, et al.
Veröffentlicht: (2023)
von: Xu, Zelai, et al.
Veröffentlicht: (2023)
Learning Generalizable Skills from Offline Multi-Task Data for Multi-Agent Cooperation
von: Liu, Sicong, et al.
Veröffentlicht: (2025)
von: Liu, Sicong, et al.
Veröffentlicht: (2025)
Explaining Strategic Decisions in Multi-Agent Reinforcement Learning for Aerial Combat Tactics
von: Selmonaj, Ardian, et al.
Veröffentlicht: (2025)
von: Selmonaj, Ardian, et al.
Veröffentlicht: (2025)
Deep Reinforcement Learning Agents for Strategic Production Policies in Microeconomic Market Simulations
von: Garrido-Merchán, Eduardo C., et al.
Veröffentlicht: (2024)
von: Garrido-Merchán, Eduardo C., et al.
Veröffentlicht: (2024)
Decentralized and Lifelong-Adaptive Multi-Agent Collaborative Learning
von: Tang, Shuo, et al.
Veröffentlicht: (2024)
von: Tang, Shuo, et al.
Veröffentlicht: (2024)
Skill-Aligned Fairness in Multi-Agent Learning for Collaboration in Healthcare
von: Ekpo, Promise Osaine, et al.
Veröffentlicht: (2025)
von: Ekpo, Promise Osaine, et al.
Veröffentlicht: (2025)
PTDE: Personalized Training with Distilled Execution for Multi-Agent Reinforcement Learning
von: Chen, Yiqun, et al.
Veröffentlicht: (2022)
von: Chen, Yiqun, et al.
Veröffentlicht: (2022)
RiskQ: Risk-sensitive Multi-Agent Reinforcement Learning Value Factorization
von: Shen, Siqi, et al.
Veröffentlicht: (2023)
von: Shen, Siqi, et al.
Veröffentlicht: (2023)
MAGIC-MASK: Multi-Agent Guided Inter-Agent Collaboration with Mask-Based Explainability for Reinforcement Learning
von: Maliha, Maisha, et al.
Veröffentlicht: (2025)
von: Maliha, Maisha, et al.
Veröffentlicht: (2025)
Generative Multi-Agent Collaboration in Embodied AI: A Systematic Review
von: Wu, Di, et al.
Veröffentlicht: (2025)
von: Wu, Di, et al.
Veröffentlicht: (2025)
Massively Multiagent Minigames for Training Generalist Agents
von: Choe, Kyoung Whan, et al.
Veröffentlicht: (2024)
von: Choe, Kyoung Whan, et al.
Veröffentlicht: (2024)
Benchmarking Generalizable Bimanual Manipulation: RoboTwin Dual-Arm Collaboration Challenge at CVPR 2025 MEIS Workshop
von: Chen, Tianxing, et al.
Veröffentlicht: (2025)
von: Chen, Tianxing, et al.
Veröffentlicht: (2025)
RiskAgent: Synergizing Language Models with Validated Tools for Evidence-Based Risk Prediction
von: Liu, Fenglin, et al.
Veröffentlicht: (2025)
von: Liu, Fenglin, et al.
Veröffentlicht: (2025)
A Data-Driven Discretized CS:GO Simulation Environment to Facilitate Strategic Multi-Agent Planning Research
von: Wang, Yunzhe, et al.
Veröffentlicht: (2025)
von: Wang, Yunzhe, et al.
Veröffentlicht: (2025)
Evaluating Collaborative and Autonomous Agents in Data-Stream-Supported Coordination of Mobile Crowdsourcing
von: Bruns, Ralf, et al.
Veröffentlicht: (2024)
von: Bruns, Ralf, et al.
Veröffentlicht: (2024)
OMAC: A Holistic Optimization Framework for LLM-Based Multi-Agent Collaboration
von: Li, Shijun, et al.
Veröffentlicht: (2025)
von: Li, Shijun, et al.
Veröffentlicht: (2025)
MAC: Masked Agent Collaboration Boosts Large Language Model Medical Decision-Making
von: Peng, Zhihao, et al.
Veröffentlicht: (2025)
von: Peng, Zhihao, et al.
Veröffentlicht: (2025)
Talk Structurally, Act Hierarchically: A Collaborative Framework for LLM Multi-Agent Systems
von: Wang, Zhao, et al.
Veröffentlicht: (2025)
von: Wang, Zhao, et al.
Veröffentlicht: (2025)
Revisiting Multi-Agent World Modeling from a Diffusion-Inspired Perspective
von: Zhang, Yang, et al.
Veröffentlicht: (2025)
von: Zhang, Yang, et al.
Veröffentlicht: (2025)
Learning to Communicate and Collaborate in a Competitive Multi-Agent Setup to Clean the Ocean from Macroplastics
von: Siedler, Philipp Dominic
Veröffentlicht: (2023)
von: Siedler, Philipp Dominic
Veröffentlicht: (2023)
Thought Communication in Multiagent Collaboration
von: Zheng, Yujia, et al.
Veröffentlicht: (2025)
von: Zheng, Yujia, et al.
Veröffentlicht: (2025)
How to Train a Leader: Hierarchical Reasoning in Multi-Agent LLMs
von: Estornell, Andrew, et al.
Veröffentlicht: (2025)
von: Estornell, Andrew, et al.
Veröffentlicht: (2025)
ThinkTank: A Framework for Generalizing Domain-Specific AI Agent Systems into Universal Collaborative Intelligence Platforms
von: Surabhi, Praneet Sai Madhu, et al.
Veröffentlicht: (2025)
von: Surabhi, Praneet Sai Madhu, et al.
Veröffentlicht: (2025)
PIVOT: Bridging Planning and Execution in LLM Agents via Trajectory Refinement
von: Zhang, Tuo, et al.
Veröffentlicht: (2026)
von: Zhang, Tuo, et al.
Veröffentlicht: (2026)
Modèles de Substitution pour les Modèles à base d'Agents : Enjeux, Méthodes et Applications
von: Saves, Paul, et al.
Veröffentlicht: (2025)
von: Saves, Paul, et al.
Veröffentlicht: (2025)
Agentic Neural Networks: Self-Evolving Multi-Agent Systems via Textual Backpropagation
von: Ma, Xiaowen, et al.
Veröffentlicht: (2025)
von: Ma, Xiaowen, et al.
Veröffentlicht: (2025)
Multi-Agent Diagnostics for Robustness via Illuminated Diversity
von: Samvelyan, Mikayel, et al.
Veröffentlicht: (2024)
von: Samvelyan, Mikayel, et al.
Veröffentlicht: (2024)
ProAgent: Building Proactive Cooperative Agents with Large Language Models
von: Zhang, Ceyao, et al.
Veröffentlicht: (2023)
von: Zhang, Ceyao, et al.
Veröffentlicht: (2023)
Fast Peer Adaptation with Context-aware Exploration
von: Ma, Long, et al.
Veröffentlicht: (2024)
von: Ma, Long, et al.
Veröffentlicht: (2024)
Proactive Agent Research Environment: Simulating Active Users to Evaluate Proactive Assistants
von: Nathani, Deepak, et al.
Veröffentlicht: (2026)
von: Nathani, Deepak, et al.
Veröffentlicht: (2026)
QSIM: Mitigating Overestimation in Multi-Agent Reinforcement Learning via Action Similarity Weighted Q-Learning
von: Li, Yuanjun, et al.
Veröffentlicht: (2026)
von: Li, Yuanjun, et al.
Veröffentlicht: (2026)
A Semi Centralized Training Decentralized Execution Architecture for Multi Agent Deep Reinforcement Learning in Traffic Signal Control
von: Rezaali, Arash, et al.
Veröffentlicht: (2025)
von: Rezaali, Arash, et al.
Veröffentlicht: (2025)
A Comprehensive Review of AI Agents: Transforming Possibilities in Technology and Beyond
von: Qu, Xiaodong, et al.
Veröffentlicht: (2025)
von: Qu, Xiaodong, et al.
Veröffentlicht: (2025)
ChipMATE: Multi-Agent Training via Reinforcement Learning for Enhanced RTL Generation
von: Yu, Zhongkai, et al.
Veröffentlicht: (2026)
von: Yu, Zhongkai, et al.
Veröffentlicht: (2026)
WideSeek-R1: Exploring Width Scaling for Broad Information Seeking via Multi-Agent Reinforcement Learning
von: Xu, Zelai, et al.
Veröffentlicht: (2026)
von: Xu, Zelai, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Distributionally Robust Cooperative Multi-Agent Reinforcement Learning via Robust Value Factorization
von: Qu, Chengrui, et al.
Veröffentlicht: (2026) -
Provably Convergent Actor-Critic for MARL through Risk-aversion
von: Zhang, Yizhou, et al.
Veröffentlicht: (2026) -
Tractable Equilibrium Computation in Markov Games through Risk Aversion
von: Mazumdar, Eric, et al.
Veröffentlicht: (2024) -
Multi-Agent Path Finding via Offline RL and LLM Collaboration
von: Atasever, Merve, et al.
Veröffentlicht: (2025) -
CORD: Generalizable Cooperation via Role Diversity
von: Matsuyama, Kanefumi, et al.
Veröffentlicht: (2025)