Mixed Q-Functionals: Advancing Value-Based Methods in Cooperative MARL with Continuous Action Domains
Fuente:
arXiv
Guardado en:
| Autores principales: | Findik, Yasin, Ahmadzadeh, S. Reza |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Collaborative Adaptation for Recovery from Unforeseen Malfunctions in Discrete and Continuous MARL Domains
por: Findik, Yasin, et al.
Publicado: (2024)
por: Findik, Yasin, et al.
Publicado: (2024)
Relational Q-Functionals: Multi-Agent Learning to Recover from Unforeseen Robot Malfunctions in Continuous Action Domains
por: Findik, Yasin, et al.
Publicado: (2024)
por: Findik, Yasin, et al.
Publicado: (2024)
A Learning Framework For Cooperative Collision Avoidance of UAV Swarms Leveraging Domain Knowledge
por: Huang, Shuangyao, et al.
Publicado: (2025)
por: Huang, Shuangyao, et al.
Publicado: (2025)
Remembering the Markov Property in Cooperative MARL
por: Tessera, Kale-ab Abebe, et al.
Publicado: (2025)
por: Tessera, Kale-ab Abebe, et al.
Publicado: (2025)
Probing Dec-POMDP Reasoning in Cooperative MARL
por: Tessera, Kale-ab, et al.
Publicado: (2026)
por: Tessera, Kale-ab, et al.
Publicado: (2026)
Generalising Multi-Agent Cooperation through Task-Agnostic Communication
por: Jayalath, Dulhan, et al.
Publicado: (2024)
por: Jayalath, Dulhan, et al.
Publicado: (2024)
Innate-Values-driven Reinforcement Learning based Cooperative Multi-Agent Cognitive Modeling
por: Yang, Qin
Publicado: (2024)
por: Yang, Qin
Publicado: (2024)
Mixed Traffic Control and Coordination from Pixels
por: Villarreal, Michael, et al.
Publicado: (2023)
por: Villarreal, Michael, et al.
Publicado: (2023)
Cooperative Target Detection with AUVs: A Dual-Timescale Hierarchical MARDL Approach
por: Xueyao, Zhang, et al.
Publicado: (2025)
por: Xueyao, Zhang, et al.
Publicado: (2025)
SafeDiver: Cooperative AUV-USV Assisted Diver Communication via Multi-agent Reinforcement Learning Approach
por: Deng, Tinglong, et al.
Publicado: (2025)
por: Deng, Tinglong, et al.
Publicado: (2025)
Coordination Failure in Cooperative Offline MARL
por: Tilbury, Callum Rhys, et al.
Publicado: (2024)
por: Tilbury, Callum Rhys, et al.
Publicado: (2024)
Learning to Control and Coordinate Mixed Traffic Through Robot Vehicles at Complex and Unsignalized Intersections
por: Wang, Dawei, et al.
Publicado: (2023)
por: Wang, Dawei, et al.
Publicado: (2023)
CrazyMARL: Decentralized Direct Motor Control Policies for Cooperative Aerial Transport of Cable-Suspended Payloads
por: Lorentz, Viktor, et al.
Publicado: (2025)
por: Lorentz, Viktor, et al.
Publicado: (2025)
Mixed-Reality Digital Twins: Leveraging the Physical and Virtual Worlds for Hybrid Sim2Real Transition of Multi-Agent Reinforcement Learning Policies
por: Samak, Chinmay Vilas, et al.
Publicado: (2024)
por: Samak, Chinmay Vilas, et al.
Publicado: (2024)
Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL
por: Zhang, Songyuan, et al.
Publicado: (2025)
por: Zhang, Songyuan, et al.
Publicado: (2025)
MARL-LNS: Cooperative Multi-agent Reinforcement Learning via Large Neighborhoods Search
por: Chen, Weizhe, et al.
Publicado: (2024)
por: Chen, Weizhe, et al.
Publicado: (2024)
CoCoL: A Communication Efficient Decentralized Collaborative Method for Multi-Robot Systems
por: Huang, Jiaxi, et al.
Publicado: (2025)
por: Huang, Jiaxi, et al.
Publicado: (2025)
Learning Partial Action Replacement in Offline MARL
por: Jin, Yue, et al.
Publicado: (2026)
por: Jin, Yue, et al.
Publicado: (2026)
Scenario-Based Curriculum Generation for Multi-Agent Autonomous Driving
por: Brunnbauer, Axel, et al.
Publicado: (2024)
por: Brunnbauer, Axel, et al.
Publicado: (2024)
A Reinforcement Learning Inspired Latent Yield Based Adaptive Algorithm Switching Mechanism
por: Nair, Jayprakash S., et al.
Publicado: (2026)
por: Nair, Jayprakash S., et al.
Publicado: (2026)
Towards Global Optimality in Cooperative MARL with the Transformation And Distillation Framework
por: Ye, Jianing, et al.
Publicado: (2022)
por: Ye, Jianing, et al.
Publicado: (2022)
Multi-Agent Inverse Q-Learning from Demonstrations
por: Haynam, Nathaniel, et al.
Publicado: (2025)
por: Haynam, Nathaniel, et al.
Publicado: (2025)
Multi-Agent Coordination in Autonomous Vehicle Routing: A Simulation-Based Study of Communication, Memory, and Routing Loops
por: Saifullah, KM Khalid, et al.
Publicado: (2025)
por: Saifullah, KM Khalid, et al.
Publicado: (2025)
Energy-Efficient Power Control for Multiple-Task Split Inference in UAVs: A Tiny Learning-Based Approach
por: Zhao, Chenxi, et al.
Publicado: (2023)
por: Zhao, Chenxi, et al.
Publicado: (2023)
Partial Action Replacement: Tackling Distribution Shift in Offline MARL
por: Jin, Yue, et al.
Publicado: (2025)
por: Jin, Yue, et al.
Publicado: (2025)
MFC-EQ: Mean-Field Control with Envelope Q-Learning for Moving Decentralized Agents in Formation
por: Lin, Qiushi, et al.
Publicado: (2024)
por: Lin, Qiushi, et al.
Publicado: (2024)
InteRACT: Transformer Models for Human Intent Prediction Conditioned on Robot Actions
por: Kedia, Kushal, et al.
Publicado: (2023)
por: Kedia, Kushal, et al.
Publicado: (2023)
Online Action-Stacking Improves Reinforcement Learning Performance for Air Traffic Control
por: Carvell, Ben, et al.
Publicado: (2026)
por: Carvell, Ben, et al.
Publicado: (2026)
SkyRover: A Modular Simulator for Cross-Domain Pathfinding
por: Ma, Wenhui, et al.
Publicado: (2025)
por: Ma, Wenhui, et al.
Publicado: (2025)
Generalizing Cooperative Eco-driving via Multi-residual Task Learning
por: Jayawardana, Vindula, et al.
Publicado: (2024)
por: Jayawardana, Vindula, et al.
Publicado: (2024)
Multi-Agent Deep Q-Network with Layer-based Communication Channel for Autonomous Internal Logistics Vehicle Scheduling in Smart Manufacturing
por: Feizabadi, Mohammad, et al.
Publicado: (2024)
por: Feizabadi, Mohammad, et al.
Publicado: (2024)
Multi-residual Mixture of Experts Learning for Cooperative Control in Multi-vehicle Systems
por: Jayawardana, Vindula, et al.
Publicado: (2025)
por: Jayawardana, Vindula, et al.
Publicado: (2025)
Co-Optimizing Reconfigurable Environments and Policies for Decentralized Multi-Agent Navigation
por: Gao, Zhan, et al.
Publicado: (2024)
por: Gao, Zhan, et al.
Publicado: (2024)
MAexp: A Generic Platform for RL-based Multi-Agent Exploration
por: Zhu, Shaohao, et al.
Publicado: (2024)
por: Zhu, Shaohao, et al.
Publicado: (2024)
Diffusion-Reinforcement Learning Hierarchical Motion Planning in Multi-agent Adversarial Games
por: Wu, Zixuan, et al.
Publicado: (2024)
por: Wu, Zixuan, et al.
Publicado: (2024)
OPTIMA: Optimized Policy for Intelligent Multi-Agent Systems Enables Coordination-Aware Autonomous Vehicles
por: Du, Rui, et al.
Publicado: (2024)
por: Du, Rui, et al.
Publicado: (2024)
MADRL-based UAVs Trajectory Design with Anti-Collision Mechanism in Vehicular Networks
por: Spampinato, Leonardo, et al.
Publicado: (2024)
por: Spampinato, Leonardo, et al.
Publicado: (2024)
Safe Human Robot Navigation in Warehouse Scenario
por: Farrell, Seth, et al.
Publicado: (2025)
por: Farrell, Seth, et al.
Publicado: (2025)
Architecture for Multi-Unmanned Aerial Vehicles based Autonomous Precision Agriculture Systems
por: Temesgen, Ebasa, et al.
Publicado: (2026)
por: Temesgen, Ebasa, et al.
Publicado: (2026)
Multi-Agent Probabilistic Ensembles with Trajectory Sampling for Connected Autonomous Vehicles
por: Wen, Ruoqi, et al.
Publicado: (2023)
por: Wen, Ruoqi, et al.
Publicado: (2023)
Ejemplares similares
-
Collaborative Adaptation for Recovery from Unforeseen Malfunctions in Discrete and Continuous MARL Domains
por: Findik, Yasin, et al.
Publicado: (2024) -
Relational Q-Functionals: Multi-Agent Learning to Recover from Unforeseen Robot Malfunctions in Continuous Action Domains
por: Findik, Yasin, et al.
Publicado: (2024) -
A Learning Framework For Cooperative Collision Avoidance of UAV Swarms Leveraging Domain Knowledge
por: Huang, Shuangyao, et al.
Publicado: (2025) -
Remembering the Markov Property in Cooperative MARL
por: Tessera, Kale-ab Abebe, et al.
Publicado: (2025) -
Probing Dec-POMDP Reasoning in Cooperative MARL
por: Tessera, Kale-ab, et al.
Publicado: (2026)