Cooperative Multi-agent RL with Communication Constraints
Fuente:
arXiv
Saved in:
| Main Authors: | Xiong, Nuoya, Singh, Aarti |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Projection Optimization: A General Framework for Multi-Objective and Multi-Group RLHF
by: Xiong, Nuoya, et al.
Published: (2025)
by: Xiong, Nuoya, et al.
Published: (2025)
The Implicit Curriculum: Learning Dynamics in RL with Verifiable Rewards
by: Huang, Yu, et al.
Published: (2026)
by: Huang, Yu, et al.
Published: (2026)
Uniformly Safe RL with Objective Suppression for Multi-Constraint Safety-Critical Applications
by: Zhou, Zihan, et al.
Published: (2024)
by: Zhou, Zihan, et al.
Published: (2024)
Diffusion Models for Offline Multi-agent Reinforcement Learning with Safety Constraints
by: Huang, Jianuo
Published: (2024)
by: Huang, Jianuo
Published: (2024)
Shielded Controller Units for RL with Operational Constraints Applied to Remote Microgrids
by: Nekoei, Hadi, et al.
Published: (2025)
by: Nekoei, Hadi, et al.
Published: (2025)
Beyond Parameter Count: Implicit Bias in Soft Mixture of Experts
by: Chung, Youngseog, et al.
Published: (2024)
by: Chung, Youngseog, et al.
Published: (2024)
Cross-environment Cooperation Enables Zero-shot Multi-agent Coordination
by: Jha, Kunal, et al.
Published: (2025)
by: Jha, Kunal, et al.
Published: (2025)
Token-Level LLM Collaboration via FusionRoute
by: Xiong, Nuoya, et al.
Published: (2026)
by: Xiong, Nuoya, et al.
Published: (2026)
SAC-GLAM: Improving Online RL for LLM agents with Soft Actor-Critic and Hindsight Relabeling
by: Gaven, Loris, et al.
Published: (2024)
by: Gaven, Loris, et al.
Published: (2024)
Performance Comparison of Deep RL Algorithms for Mixed Traffic Cooperative Lane-Changing
by: Yao, Xue, et al.
Published: (2024)
by: Yao, Xue, et al.
Published: (2024)
GHQ: Grouped Hybrid Q Learning for Heterogeneous Cooperative Multi-agent Reinforcement Learning
by: Yu, Xiaoyang, et al.
Published: (2023)
by: Yu, Xiaoyang, et al.
Published: (2023)
LLM-Guided Communication for Cooperative Multi-Agent Reinforcement Learning
by: Bae, Sangjun, et al.
Published: (2026)
by: Bae, Sangjun, et al.
Published: (2026)
Offline Multi-task Transfer RL with Representational Penalization
by: Bose, Avinandan, et al.
Published: (2024)
by: Bose, Avinandan, et al.
Published: (2024)
BlendRL: A Framework for Merging Symbolic and Neural Policy Learning
by: Shindo, Hikaru, et al.
Published: (2024)
by: Shindo, Hikaru, et al.
Published: (2024)
STO-RL: Offline RL under Sparse Rewards via LLM-Guided Subgoal Temporal Order
by: Gu, Chengyang, et al.
Published: (2026)
by: Gu, Chengyang, et al.
Published: (2026)
Imbalanced Gradients in RL Post-Training of Multi-Task LLMs
by: Wu, Runzhe, et al.
Published: (2025)
by: Wu, Runzhe, et al.
Published: (2025)
RL$^3$: Boosting Meta Reinforcement Learning via RL inside RL$^2$
by: Bhatia, Abhinav, et al.
Published: (2023)
by: Bhatia, Abhinav, et al.
Published: (2023)
CaRT: Teaching LLM Agents to Know When They Know Enough
by: Liu, Grace, et al.
Published: (2025)
by: Liu, Grace, et al.
Published: (2025)
Transformers Provably Learn Chain-of-Thought Reasoning with Length Generalization
by: Huang, Yu, et al.
Published: (2025)
by: Huang, Yu, et al.
Published: (2025)
Policy Agnostic RL: Offline RL and Online RL Fine-Tuning of Any Class and Backbone
by: Mark, Max Sobol, et al.
Published: (2024)
by: Mark, Max Sobol, et al.
Published: (2024)
Exponential Topology-enabled Scalable Communication in Multi-agent Reinforcement Learning
by: Li, Xinran, et al.
Published: (2025)
by: Li, Xinran, et al.
Published: (2025)
Variational Offline Multi-agent Skill Discovery
by: Chen, Jiayu, et al.
Published: (2024)
by: Chen, Jiayu, et al.
Published: (2024)
Language-Conditioned Offline RL for Multi-Robot Navigation
by: Morad, Steven, et al.
Published: (2024)
by: Morad, Steven, et al.
Published: (2024)
The Importance of Online Data: Understanding Preference Fine-tuning via Coverage
by: Song, Yuda, et al.
Published: (2024)
by: Song, Yuda, et al.
Published: (2024)
COOL-MC: Verifying and Explaining RL Policies for Multi-bridge Network Maintenance
by: Gross, Dennis
Published: (2026)
by: Gross, Dennis
Published: (2026)
Multi-Objective Instruction-Aware Representation Learning in Procedural Content Generation RL
by: Kim, Sung-Hyun, et al.
Published: (2025)
by: Kim, Sung-Hyun, et al.
Published: (2025)
Diffusion-Driven Semantic Communication for Generative Models with Bandwidth Constraints
by: Guo, Lei, et al.
Published: (2024)
by: Guo, Lei, et al.
Published: (2024)
Forager: a lightweight testbed for continual learning with partial observability in RL
by: Tang, Steven, et al.
Published: (2026)
by: Tang, Steven, et al.
Published: (2026)
Partial Inverse Design of High-Performance Concrete Using Cooperative Neural Networks for Constraint-Aware Mix Generation
by: Nugraha, Agung, et al.
Published: (2025)
by: Nugraha, Agung, et al.
Published: (2025)
An Empirical Study on the Effectiveness of Incorporating Offline RL As Online RL Subroutines
by: Su, Jianhai, et al.
Published: (2025)
by: Su, Jianhai, et al.
Published: (2025)
MADiff: Offline Multi-agent Learning with Diffusion Models
by: Zhu, Zhengbang, et al.
Published: (2023)
by: Zhu, Zhengbang, et al.
Published: (2023)
Adaptive and Robust DBSCAN with Multi-agent Reinforcement Learning
by: Peng, Hao, et al.
Published: (2025)
by: Peng, Hao, et al.
Published: (2025)
Omni-Thinker: Scaling Multi-Task RL in LLMs with Hybrid Reward and Task Scheduling
by: Li, Derek, et al.
Published: (2025)
by: Li, Derek, et al.
Published: (2025)
DeepLTL: Learning to Efficiently Satisfy Complex LTL Specifications for Multi-Task RL
by: Jackermeier, Mathias, et al.
Published: (2024)
by: Jackermeier, Mathias, et al.
Published: (2024)
RL-MSA: a Reinforcement Learning-based Multi-line bus Scheduling Approach
by: Liu, Yingzhuo
Published: (2024)
by: Liu, Yingzhuo
Published: (2024)
RL in Name Only? Analyzing the Structural Assumptions in RL post-training for LLMs
by: Samineni, Soumya Rani, et al.
Published: (2025)
by: Samineni, Soumya Rani, et al.
Published: (2025)
Rethinking RL Evaluation: Can Benchmarks Truly Reveal Failures of RL Methods?
by: Chen, Zihan, et al.
Published: (2025)
by: Chen, Zihan, et al.
Published: (2025)
Scaling Offline Model-Based RL via Jointly-Optimized World-Action Model Pretraining
by: Cheng, Jie, et al.
Published: (2024)
by: Cheng, Jie, et al.
Published: (2024)
Consolidation or Adaptation? PRISM: Disentangling SFT and RL Data via Gradient Concentration
by: Zhao, Yang, et al.
Published: (2026)
by: Zhao, Yang, et al.
Published: (2026)
NeoRL-2: Near Real-World Benchmarks for Offline Reinforcement Learning with Extended Realistic Scenarios
by: Gao, Songyi, et al.
Published: (2025)
by: Gao, Songyi, et al.
Published: (2025)
Similar Items
-
Projection Optimization: A General Framework for Multi-Objective and Multi-Group RLHF
by: Xiong, Nuoya, et al.
Published: (2025) -
The Implicit Curriculum: Learning Dynamics in RL with Verifiable Rewards
by: Huang, Yu, et al.
Published: (2026) -
Uniformly Safe RL with Objective Suppression for Multi-Constraint Safety-Critical Applications
by: Zhou, Zihan, et al.
Published: (2024) -
Diffusion Models for Offline Multi-agent Reinforcement Learning with Safety Constraints
by: Huang, Jianuo
Published: (2024) -
Shielded Controller Units for RL with Operational Constraints Applied to Remote Microgrids
by: Nekoei, Hadi, et al.
Published: (2025)