Communication-Efficient Soft Actor-Critic Policy Collaboration via Regulated Segment Mixture
Fuente:
arXiv
Saved in:
| Main Authors: | Yu, Xiaoxue, Li, Rongpeng, Liang, Chengchao, Zhao, Zhifeng |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Communications-Incentivized Collaborative Reasoning in NetGPT through Agentic Reinforcement Learning
by: Yu, Xiaoxue, et al.
Published: (2026)
by: Yu, Xiaoxue, et al.
Published: (2026)
Robust Event-Triggered Integrated Communication and Control with Graph Information Bottleneck Optimization
by: Wang, Ziqiong, et al.
Published: (2025)
by: Wang, Ziqiong, et al.
Published: (2025)
Multi-agent Uncertainty-Aware Pessimistic Model-Based Reinforcement Learning for Connected Autonomous Vehicles
by: Wen, Ruoqi, et al.
Published: (2025)
by: Wen, Ruoqi, et al.
Published: (2025)
Multi-Agent Probabilistic Ensembles with Trajectory Sampling for Connected Autonomous Vehicles
by: Wen, Ruoqi, et al.
Published: (2023)
by: Wen, Ruoqi, et al.
Published: (2023)
Multi-Agent Conditional Diffusion Model with Mean Field Communication as Wireless Resource Allocation Planner
by: Meng, Kechen, et al.
Published: (2025)
by: Meng, Kechen, et al.
Published: (2025)
Multi-Agent Soft Actor-Critic with Coordinated Loss for Autonomous Mobility-on-Demand Fleet Control
by: Woywood, Zeno, et al.
Published: (2024)
by: Woywood, Zeno, et al.
Published: (2024)
Communication-Efficient Cooperative SLAMMOT via Determining the Number of Collaboration Vehicles
by: Fang, Susu, et al.
Published: (2024)
by: Fang, Susu, et al.
Published: (2024)
RALLY: Role-Adaptive LLM-Driven Yoked Navigation for Agentic UAV Swarms
by: Wang, Ziyao, et al.
Published: (2025)
by: Wang, Ziyao, et al.
Published: (2025)
Towards Efficient Collaboration via Graph Modeling in Reinforcement Learning
by: Fan, Wenzhe, et al.
Published: (2024)
by: Fan, Wenzhe, et al.
Published: (2024)
Nash Soft Actor-Critic LEO Satellite Handover Management Algorithm for Flying Vehicles
by: Chen, Jinxuan, et al.
Published: (2024)
by: Chen, Jinxuan, et al.
Published: (2024)
GTDE: Grouped Training with Decentralized Execution for Multi-agent Actor-Critic
by: Li, Mengxian, et al.
Published: (2024)
by: Li, Mengxian, et al.
Published: (2024)
Multi-Agent LLM Actor-Critic Framework for Social Robot Navigation
by: Wang, Weizheng, et al.
Published: (2025)
by: Wang, Weizheng, et al.
Published: (2025)
Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic
by: Liu, Shuo, et al.
Published: (2026)
by: Liu, Shuo, et al.
Published: (2026)
MetaCrit: A Critical Thinking Framework for Self-Regulated LLM Reasoning
by: Hou, Xinmeng, et al.
Published: (2025)
by: Hou, Xinmeng, et al.
Published: (2025)
On-policy Actor-Critic Reinforcement Learning for Multi-UAV Exploration
by: Farid, Ali Moltajaei, et al.
Published: (2024)
by: Farid, Ali Moltajaei, et al.
Published: (2024)
Multi-Agent Actor-Critic Generative AI for Query Resolution and Analysis
by: Rahman, Mohammad Wali Ur, et al.
Published: (2025)
by: Rahman, Mohammad Wali Ur, et al.
Published: (2025)
IBSEN: Director-Actor Agent Collaboration for Controllable and Interactive Drama Script Generation
by: Han, Senyu, et al.
Published: (2024)
by: Han, Senyu, et al.
Published: (2024)
Adaptive AUV Hunting Policy with Covert Communication via Diffusion Model
by: Guo, Xu, et al.
Published: (2025)
by: Guo, Xu, et al.
Published: (2025)
Goal Recognition using Actor-Critic Optimization
by: Nageris, Ben, et al.
Published: (2024)
by: Nageris, Ben, et al.
Published: (2024)
Multi-agent Off-policy Actor-Critic Reinforcement Learning for Partially Observable Environments
by: Zhaikhan, Ainur, et al.
Published: (2024)
by: Zhaikhan, Ainur, et al.
Published: (2024)
Gated Coordination for Efficient Multi-Agent Collaboration in Minecraft Game
by: Jian, HuaDong, et al.
Published: (2026)
by: Jian, HuaDong, et al.
Published: (2026)
Collaborative Belief Reasoning with LLMs for Efficient Multi-Agent Collaboration
by: Wang, Zhimin, et al.
Published: (2025)
by: Wang, Zhimin, et al.
Published: (2025)
Belief-Driven Multi-Agent Collaboration via Approximate Perfect Bayesian Equilibrium for Social Simulation
by: Fang, Weiwei, et al.
Published: (2026)
by: Fang, Weiwei, et al.
Published: (2026)
Multi-Agent Actor-Critics in Autonomous Cyber Defense
by: Wang, Mingjun, et al.
Published: (2024)
by: Wang, Mingjun, et al.
Published: (2024)
SC-MAS: Constructing Cost-Efficient Multi-Agent Systems with Edge-Level Heterogeneous Collaboration
by: Zhao, Di, et al.
Published: (2026)
by: Zhao, Di, et al.
Published: (2026)
Synthesis of Communication Policies for Multi-Agent Systems Robust to Communication Restrictions
by: Soudijani, Saleh, et al.
Published: (2025)
by: Soudijani, Saleh, et al.
Published: (2025)
ACCoRD: Actor-Critic Conflict Resolution with Deep learning for O-RAN xApps
by: Adamczyk, Cezary, et al.
Published: (2026)
by: Adamczyk, Cezary, et al.
Published: (2026)
Learning Efficient Communication Protocols for Multi-Agent Reinforcement Learning
by: Zhang, Xinren, et al.
Published: (2025)
by: Zhang, Xinren, et al.
Published: (2025)
Agent-GSPO: Communication-Efficient Multi-Agent Systems via Group Sequence Policy Optimization
by: Fan, Yijia, et al.
Published: (2025)
by: Fan, Yijia, et al.
Published: (2025)
DTPPO: Dual-Transformer Encoder-based Proximal Policy Optimization for Multi-UAV Navigation in Unseen Complex Environments
by: Wei, Anning, et al.
Published: (2024)
by: Wei, Anning, et al.
Published: (2024)
Collaborative Scheduling of Time-dependent UAVs,Vehicles and Workers for Crowdsensing in Disaster Response
by: Han, Lei, et al.
Published: (2025)
by: Han, Lei, et al.
Published: (2025)
AgentCDM: Enhancing Multi-Agent Collaborative Decision-Making via ACH-Inspired Structured Reasoning
by: Zhao, Xuyang, et al.
Published: (2025)
by: Zhao, Xuyang, et al.
Published: (2025)
Maestro: Learning to Collaborate via Conditional Listwise Policy Optimization for Multi-Agent LLMs
by: Yang, Wei, et al.
Published: (2025)
by: Yang, Wei, et al.
Published: (2025)
Generalization of Heterogeneous Multi-Robot Policies via Awareness and Communication of Capabilities
by: Howell, Pierce, et al.
Published: (2024)
by: Howell, Pierce, et al.
Published: (2024)
Breaking the Communication-Accuracy Trade-off: A Sparsified Information Diffusion Framework for Multi-Agent Collaborative Perception
by: Zha, Jirong, et al.
Published: (2026)
by: Zha, Jirong, et al.
Published: (2026)
CoCoL: A Communication Efficient Decentralized Collaborative Method for Multi-Robot Systems
by: Huang, Jiaxi, et al.
Published: (2025)
by: Huang, Jiaxi, et al.
Published: (2025)
Finite-Time Global Optimality Convergence in Deep Neural Actor-Critic Methods for Decentralized Multi-Agent Reinforcement Learning
by: Zhang, Zhiyao, et al.
Published: (2025)
by: Zhang, Zhiyao, et al.
Published: (2025)
Sample and Communication Efficient Fully Decentralized MARL Policy Evaluation via a New Approach: Local TD update
by: Hairi, Fnu, et al.
Published: (2024)
by: Hairi, Fnu, et al.
Published: (2024)
Thought Communication in Multiagent Collaboration
by: Zheng, Yujia, et al.
Published: (2025)
by: Zheng, Yujia, et al.
Published: (2025)
Beyond the Individual: Virtualizing Multi-Disciplinary Reasoning for Clinical Intake via Collaborative Agents
by: Chen, Huangwei, et al.
Published: (2026)
by: Chen, Huangwei, et al.
Published: (2026)
Similar Items
-
Communications-Incentivized Collaborative Reasoning in NetGPT through Agentic Reinforcement Learning
by: Yu, Xiaoxue, et al.
Published: (2026) -
Robust Event-Triggered Integrated Communication and Control with Graph Information Bottleneck Optimization
by: Wang, Ziqiong, et al.
Published: (2025) -
Multi-agent Uncertainty-Aware Pessimistic Model-Based Reinforcement Learning for Connected Autonomous Vehicles
by: Wen, Ruoqi, et al.
Published: (2025) -
Multi-Agent Probabilistic Ensembles with Trajectory Sampling for Connected Autonomous Vehicles
by: Wen, Ruoqi, et al.
Published: (2023) -
Multi-Agent Conditional Diffusion Model with Mean Field Communication as Wireless Resource Allocation Planner
by: Meng, Kechen, et al.
Published: (2025)