Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms
Fuente:
arXiv
Saved in:
| Main Authors: | Ren, Zhenjie, Wei, Xiaoli, Yu, Xiang, Zhou, Xun Yu |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Continuous-time q-learning for mean-field control with common noise, part-I: Theoretical foundations
by: Ren, Zhenjie, et al.
Published: (2026)
by: Ren, Zhenjie, et al.
Published: (2026)
Continuous-time q-learning for mean-field control problems
by: Wei, Xiaoli, et al.
Published: (2023)
by: Wei, Xiaoli, et al.
Published: (2023)
Unified continuous-time q-learning for mean-field game and mean-field control problems
by: Wei, Xiaoli, et al.
Published: (2024)
by: Wei, Xiaoli, et al.
Published: (2024)
Scalable spectral representations for multi-agent reinforcement learning in network MDPs
by: Ren, Zhaolin, et al.
Published: (2024)
by: Ren, Zhaolin, et al.
Published: (2024)
An active learning method for solving competitive multi-agent decision-making and control problems
by: Fabiani, Filippo, et al.
Published: (2022)
by: Fabiani, Filippo, et al.
Published: (2022)
Exploiting inter-agent coupling information for efficient reinforcement learning of cooperative LQR
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
Graphon Mean-Field Control for Cooperative Multi-Agent Reinforcement Learning
by: Hu, Yuanquan, et al.
Published: (2022)
by: Hu, Yuanquan, et al.
Published: (2022)
Hybrid topology control: a dynamic leader-based distributed edge-addition and deletion mechanism
by: Garg, Kunal, et al.
Published: (2026)
by: Garg, Kunal, et al.
Published: (2026)
Regret of exploratory policy improvement and $q$-learning
by: Tang, Wenpin, et al.
Published: (2024)
by: Tang, Wenpin, et al.
Published: (2024)
Data/moment-driven approaches for fast predictive control of collective dynamics
by: Albi, Giacomo, et al.
Published: (2024)
by: Albi, Giacomo, et al.
Published: (2024)
Near-Optimal Online Learning for Multi-Agent Submodular Coordination: Tight Approximation and Communication Efficiency
by: Zhang, Qixin, et al.
Published: (2025)
by: Zhang, Qixin, et al.
Published: (2025)
Distributed Stochastic Zeroth-Order Optimization with Compressed Communication
by: Hua, Youqing, et al.
Published: (2025)
by: Hua, Youqing, et al.
Published: (2025)
Distributed Random Reshuffling Methods with Improved Convergence
by: Huang, Kun, et al.
Published: (2023)
by: Huang, Kun, et al.
Published: (2023)
Distributed fixed-point algorithms for dynamic convex optimization over decentralized and unbalanced wireless networks
by: Agrawal, Navneet, et al.
Published: (2024)
by: Agrawal, Navneet, et al.
Published: (2024)
Performance bound analysis of linear consensus algorithm on strongly connected graphs using effective resistance and reversiblization
by: Yonaiyama, Takumi, et al.
Published: (2025)
by: Yonaiyama, Takumi, et al.
Published: (2025)
Model predictive altitude and velocity control in ergodic potential field directed multi-UAV search
by: Lanča, Luka, et al.
Published: (2024)
by: Lanča, Luka, et al.
Published: (2024)
Continuous-time Discounted Mirror-Descent Dynamics in Monotone Concave Games
by: Gao, Bolin, et al.
Published: (2019)
by: Gao, Bolin, et al.
Published: (2019)
A Generalized Sinkhorn Algorithm for Mean-Field Schrödinger Bridge
by: Eldesoukey, Asmaa, et al.
Published: (2026)
by: Eldesoukey, Asmaa, et al.
Published: (2026)
Distributed Online Submodular Maximization under Communication Delays: A Simultaneous Decision-Making Approach
by: Xu, Zirui, et al.
Published: (2026)
by: Xu, Zirui, et al.
Published: (2026)
Hierarchical Decentralized Stochastic Control for Cyber-Physical Systems
by: Kaza, Kesav, et al.
Published: (2025)
by: Kaza, Kesav, et al.
Published: (2025)
Robust Online Learning over Networks
by: Bastianello, Nicola, et al.
Published: (2023)
by: Bastianello, Nicola, et al.
Published: (2023)
Formation Shape Control using the Gromov-Wasserstein Metric
by: Nakashima, Haruto, et al.
Published: (2025)
by: Nakashima, Haruto, et al.
Published: (2025)
Logarithmically Quantized Distributed Optimization over Dynamic Multi-Agent Networks
by: Doostmohammadian, Mohammadreza, et al.
Published: (2024)
by: Doostmohammadian, Mohammadreza, et al.
Published: (2024)
Principled Learning-to-Communicate with Quasi-Classical Information Structures
by: Liu, Xiangyu, et al.
Published: (2026)
by: Liu, Xiangyu, et al.
Published: (2026)
Optimization and Learning in Open Multi-Agent Systems
by: Deplano, Diego, et al.
Published: (2025)
by: Deplano, Diego, et al.
Published: (2025)
Structured Cooperative Multi-Agent Reinforcement Learning: a Bayesian Network Perspective
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
DeMuon: A Decentralized Muon for Matrix Optimization over Graphs
by: He, Chuan, et al.
Published: (2025)
by: He, Chuan, et al.
Published: (2025)
Fast algorithm for centralized multi-agent maze exploration
by: Crnković, Bojan, et al.
Published: (2023)
by: Crnković, Bojan, et al.
Published: (2023)
Community-based Multi-Agent Reinforcement Learning with Transfer and Active Exploration
by: Shi, Zhaoyang
Published: (2025)
by: Shi, Zhaoyang
Published: (2025)
On the Reliability Limits of LLM-Based Multi-Agent Planning
by: Ao, Ruicheng, et al.
Published: (2026)
by: Ao, Ruicheng, et al.
Published: (2026)
Approximate Linear Programming for Decentralized Policy Iteration in Cooperative Multi-agent Markov Decision Processes
by: Mandal, Lakshmi, et al.
Published: (2023)
by: Mandal, Lakshmi, et al.
Published: (2023)
Analysis of Multiscale Reinforcement Q-Learning Algorithms for Mean Field Control Games
by: Angiuli, Andrea, et al.
Published: (2024)
by: Angiuli, Andrea, et al.
Published: (2024)
Effective Policy Learning for Multi-Agent Online Coordination Beyond Submodular Objectives
by: Zhang, Qixin, et al.
Published: (2025)
by: Zhang, Qixin, et al.
Published: (2025)
High-Probability Convergence in Decentralized Stochastic Optimization with Gradient Tracking
by: Armacki, Aleksandar, et al.
Published: (2026)
by: Armacki, Aleksandar, et al.
Published: (2026)
High-Probability Convergence Guarantees of Decentralized SGD
by: Armacki, Aleksandar, et al.
Published: (2025)
by: Armacki, Aleksandar, et al.
Published: (2025)
Deep Distributed Optimization for Large-Scale Quadratic Programming
by: Saravanos, Augustinos D., et al.
Published: (2024)
by: Saravanos, Augustinos D., et al.
Published: (2024)
Momentum for the Win: Collaborative Federated Reinforcement Learning across Heterogeneous Environments
by: Wang, Han, et al.
Published: (2024)
by: Wang, Han, et al.
Published: (2024)
Diffusion Stochastic Learning Over Adaptive Competing Networks
by: Zhao, Yike, et al.
Published: (2025)
by: Zhao, Yike, et al.
Published: (2025)
Learning Decentralized Partially Observable Mean Field Control for Artificial Collective Behavior
by: Cui, Kai, et al.
Published: (2023)
by: Cui, Kai, et al.
Published: (2023)
Adaptive Decentralized Composite Optimization via Three-Operator Splitting
by: Chen, Xiaokai, et al.
Published: (2026)
by: Chen, Xiaokai, et al.
Published: (2026)
Similar Items
-
Continuous-time q-learning for mean-field control with common noise, part-I: Theoretical foundations
by: Ren, Zhenjie, et al.
Published: (2026) -
Continuous-time q-learning for mean-field control problems
by: Wei, Xiaoli, et al.
Published: (2023) -
Unified continuous-time q-learning for mean-field game and mean-field control problems
by: Wei, Xiaoli, et al.
Published: (2024) -
Scalable spectral representations for multi-agent reinforcement learning in network MDPs
by: Ren, Zhaolin, et al.
Published: (2024) -
An active learning method for solving competitive multi-agent decision-making and control problems
by: Fabiani, Filippo, et al.
Published: (2022)