Peer Learning: Learning Complex Policies in Groups from Scratch via Action Recommendations
Fuente:
arXiv
Saved in:
| Main Authors: | Derstroff, Cedric, Cerrato, Mattia, Brugger, Jannis, Peters, Jan, Kramer, Stefan |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reinforcement Learning Enabled Peer-to-Peer Energy Trading for Dairy Farms
by: Shah, Mian Ibad Ali, et al.
Published: (2024)
by: Shah, Mian Ibad Ali, et al.
Published: (2024)
QSIM: Mitigating Overestimation in Multi-Agent Reinforcement Learning via Action Similarity Weighted Q-Learning
by: Li, Yuanjun, et al.
Published: (2026)
by: Li, Yuanjun, et al.
Published: (2026)
Learning Partial Action Replacement in Offline MARL
by: Jin, Yue, et al.
Published: (2026)
by: Jin, Yue, et al.
Published: (2026)
ScholarPeer: A Context-Aware Multi-Agent Framework for Automated Peer Review
by: Goyal, Palash, et al.
Published: (2026)
by: Goyal, Palash, et al.
Published: (2026)
GHQ: Grouped Hybrid Q Learning for Heterogeneous Cooperative Multi-agent Reinforcement Learning
by: Yu, Xiaoyang, et al.
Published: (2023)
by: Yu, Xiaoyang, et al.
Published: (2023)
Automated Scientific Discovery: From Equation Discovery to Autonomous Discovery Systems
by: Kramer, Stefan, et al.
Published: (2023)
by: Kramer, Stefan, et al.
Published: (2023)
PMAT: Optimizing Action Generation Order in Multi-Agent Reinforcement Learning
by: Hu, Kun, et al.
Published: (2025)
by: Hu, Kun, et al.
Published: (2025)
Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems
by: Hu, Liyuan
Published: (2025)
by: Hu, Liyuan
Published: (2025)
Group-Aware Coordination Graph for Multi-Agent Reinforcement Learning
by: Duan, Wei, et al.
Published: (2024)
by: Duan, Wei, et al.
Published: (2024)
Imitating Mistakes in a Learning Companion AI Agent for Online Peer Learning
by: Moribe, Sosui, et al.
Published: (2025)
by: Moribe, Sosui, et al.
Published: (2025)
Neural-Guided Equation Discovery
by: Brugger, Jannis, et al.
Published: (2025)
by: Brugger, Jannis, et al.
Published: (2025)
Redistributing Rewards Across Time and Agents for Multi-Agent Reinforcement Learning
by: Kapoor, Aditya, et al.
Published: (2025)
by: Kapoor, Aditya, et al.
Published: (2025)
Fast Peer Adaptation with Context-aware Exploration
by: Ma, Long, et al.
Published: (2024)
by: Ma, Long, et al.
Published: (2024)
Towards Generalizability of Multi-Agent Reinforcement Learning in Graphs with Recurrent Message Passing
by: Weil, Jannis, et al.
Published: (2024)
by: Weil, Jannis, et al.
Published: (2024)
Centralized Permutation Equivariant Policy for Cooperative Multi-Agent Reinforcement Learning
by: Xu, Zhuofan, et al.
Published: (2025)
by: Xu, Zhuofan, et al.
Published: (2025)
Descent-Guided Policy Gradient for Scalable Cooperative Multi-Agent Learning
by: Yang, Shan, et al.
Published: (2026)
by: Yang, Shan, et al.
Published: (2026)
Deep Reinforcement Learning Agents for Strategic Production Policies in Microeconomic Market Simulations
by: Garrido-Merchán, Eduardo C., et al.
Published: (2024)
by: Garrido-Merchán, Eduardo C., et al.
Published: (2024)
Using Deep Q-Learning to Dynamically Toggle between Push/Pull Actions in Computational Trust Mechanisms
by: Lygizou, Zoi, et al.
Published: (2024)
by: Lygizou, Zoi, et al.
Published: (2024)
MESA: Cooperative Meta-Exploration in Multi-Agent Learning through Exploiting State-Action Space Structure
by: Zhang, Zhicheng, et al.
Published: (2024)
by: Zhang, Zhicheng, et al.
Published: (2024)
Learning Complex Teamwork Tasks Using a Given Sub-task Decomposition
by: Fosong, Elliot, et al.
Published: (2023)
by: Fosong, Elliot, et al.
Published: (2023)
MeloTune: On-Device Arousal Learning and Peer-to-Peer Mood Coupling for Proactive Music Curation
by: Xu, Hongwei
Published: (2026)
by: Xu, Hongwei
Published: (2026)
AutoAct: Automatic Agent Learning from Scratch for QA via Self-Planning
by: Qiao, Shuofei, et al.
Published: (2024)
by: Qiao, Shuofei, et al.
Published: (2024)
A Hierarchical Framework with Spatio-Temporal Consistency Learning for Emergence Detection in Complex Adaptive Systems
by: Chen, Siyuan, et al.
Published: (2024)
by: Chen, Siyuan, et al.
Published: (2024)
Online Action-Stacking Improves Reinforcement Learning Performance for Air Traffic Control
by: Carvell, Ben, et al.
Published: (2026)
by: Carvell, Ben, et al.
Published: (2026)
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise
by: Wu, Xuefei, et al.
Published: (2025)
by: Wu, Xuefei, et al.
Published: (2025)
Decentralizing Multi-Agent Reinforcement Learning with Temporal Causal Information
by: Corazza, Jan, et al.
Published: (2025)
by: Corazza, Jan, et al.
Published: (2025)
Scalable Submodular Policy Optimization via Pruned Submodularity Graph
by: Anand, Aditi, et al.
Published: (2025)
by: Anand, Aditi, et al.
Published: (2025)
PeerGuard: Defending Multi-Agent Systems Against Backdoor Attacks Through Mutual Reasoning
by: Fan, Falong, et al.
Published: (2025)
by: Fan, Falong, et al.
Published: (2025)
Adaptively Coordinating with Novel Partners via Learned Latent Strategies
by: Li, Benjamin, et al.
Published: (2025)
by: Li, Benjamin, et al.
Published: (2025)
Amplifying Exploration in Monte-Carlo Tree Search by Focusing on the Unknown
by: Derstroff, Cedric, et al.
Published: (2024)
by: Derstroff, Cedric, et al.
Published: (2024)
Hierarchical Policy-Gradient Reinforcement Learning for Multi-Agent Shepherding Control of Non-Cohesive Targets
by: Covone, Stefano, et al.
Published: (2025)
by: Covone, Stefano, et al.
Published: (2025)
ANNEAL: Adapting LLM Agents via Governed Symbolic Patch Learning
by: Hakim, Safayat Bin, et al.
Published: (2026)
by: Hakim, Safayat Bin, et al.
Published: (2026)
Emergence of Fair Leaders via Mediators in Multi-Agent Reinforcement Learning
by: Dodwadmath, Akshay, et al.
Published: (2025)
by: Dodwadmath, Akshay, et al.
Published: (2025)
Agent-GSPO: Communication-Efficient Multi-Agent Systems via Group Sequence Policy Optimization
by: Fan, Yijia, et al.
Published: (2025)
by: Fan, Yijia, et al.
Published: (2025)
Maestro: Learning to Collaborate via Conditional Listwise Policy Optimization for Multi-Agent LLMs
by: Yang, Wei, et al.
Published: (2025)
by: Yang, Wei, et al.
Published: (2025)
Learning the Preferences of a Learning Agent
by: Sadek, Karim Abdel, et al.
Published: (2026)
by: Sadek, Karim Abdel, et al.
Published: (2026)
Achieving Collective Welfare in Multi-Agent Reinforcement Learning via Suggestion Sharing
by: Jin, Yue, et al.
Published: (2024)
by: Jin, Yue, et al.
Published: (2024)
Resource Governance in Networked Systems via Integrated Variational Autoencoders and Reinforcement Learning
by: Chen, Qiliang, et al.
Published: (2024)
by: Chen, Qiliang, et al.
Published: (2024)
ToMCAT: Theory-of-Mind for Cooperative Agents in Teams via Multiagent Diffusion Policies
by: Sequeira, Pedro, et al.
Published: (2025)
by: Sequeira, Pedro, et al.
Published: (2025)
Learning to Recommend Multi-Agent Subgraphs from Calling Trees
by: Song, Xinyuan, et al.
Published: (2026)
by: Song, Xinyuan, et al.
Published: (2026)
Similar Items
-
Reinforcement Learning Enabled Peer-to-Peer Energy Trading for Dairy Farms
by: Shah, Mian Ibad Ali, et al.
Published: (2024) -
QSIM: Mitigating Overestimation in Multi-Agent Reinforcement Learning via Action Similarity Weighted Q-Learning
by: Li, Yuanjun, et al.
Published: (2026) -
Learning Partial Action Replacement in Offline MARL
by: Jin, Yue, et al.
Published: (2026) -
ScholarPeer: A Context-Aware Multi-Agent Framework for Automated Peer Review
by: Goyal, Palash, et al.
Published: (2026) -
GHQ: Grouped Hybrid Q Learning for Heterogeneous Cooperative Multi-agent Reinforcement Learning
by: Yu, Xiaoyang, et al.
Published: (2023)