ARAC: Adaptive Regularized Multi-Agent Soft Actor-Critic in Graph-Structured Adversarial Games
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shi, Ruochuan, Lu, Runyu, Zhu, Yuanheng, Zhao, Dongbin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
R2PS: Worst-Case Robust Real-Time Pursuit Strategies under Partial Observability
von: Lu, Runyu, et al.
Veröffentlicht: (2025)
von: Lu, Runyu, et al.
Veröffentlicht: (2025)
Equilibrium Policy Generalization: A Reinforcement Learning Framework for Cross-Graph Zero-Shot Generalization in Pursuit-Evasion Games
von: Lu, Runyu, et al.
Veröffentlicht: (2025)
von: Lu, Runyu, et al.
Veröffentlicht: (2025)
ISAACS: Iterative Soft Adversarial Actor-Critic for Safety
von: Hsu, Kai-Chieh, et al.
Veröffentlicht: (2022)
von: Hsu, Kai-Chieh, et al.
Veröffentlicht: (2022)
Discretizing Continuous Action Space with Unimodal Probability Distributions for On-Policy Reinforcement Learning
von: Zhu, Yuanyang, et al.
Veröffentlicht: (2024)
von: Zhu, Yuanyang, et al.
Veröffentlicht: (2024)
Revisiting Discrete Soft Actor-Critic
von: Zhou, Haibin, et al.
Veröffentlicht: (2022)
von: Zhou, Haibin, et al.
Veröffentlicht: (2022)
Safe Langevin Soft Actor Critic
von: Keswani, Mahesh, et al.
Veröffentlicht: (2026)
von: Keswani, Mahesh, et al.
Veröffentlicht: (2026)
Meta-DT: Offline Meta-RL as Conditional Sequence Modeling with World Model Disentanglement
von: Wang, Zhi, et al.
Veröffentlicht: (2024)
von: Wang, Zhi, et al.
Veröffentlicht: (2024)
PAC-Bayesian Soft Actor-Critic Learning
von: Tasdighi, Bahareh, et al.
Veröffentlicht: (2023)
von: Tasdighi, Bahareh, et al.
Veröffentlicht: (2023)
Average-Reward Soft Actor-Critic
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
Wasserstein Barycenter Soft Actor-Critic
von: Shahrooei, Zahra, et al.
Veröffentlicht: (2025)
von: Shahrooei, Zahra, et al.
Veröffentlicht: (2025)
Refined Analysis of Entropy-Regularized Actor-Critic
von: Labbi, Safwan, et al.
Veröffentlicht: (2026)
von: Labbi, Safwan, et al.
Veröffentlicht: (2026)
DipLLM: Fine-Tuning LLM for Strategic Decision-making in Diplomacy
von: Xu, Kaixuan, et al.
Veröffentlicht: (2025)
von: Xu, Kaixuan, et al.
Veröffentlicht: (2025)
Multi-Agent Soft Actor-Critic with Coordinated Loss for Autonomous Mobility-on-Demand Fleet Control
von: Woywood, Zeno, et al.
Veröffentlicht: (2024)
von: Woywood, Zeno, et al.
Veröffentlicht: (2024)
$π$-Play: Multi-Agent Self-Play via Privileged Self-Distillation without External Data
von: Zhang, Yaocheng, et al.
Veröffentlicht: (2026)
von: Zhang, Yaocheng, et al.
Veröffentlicht: (2026)
Distributional Soft Actor-Critic with Diffusion Policy
von: Liu, Tong, et al.
Veröffentlicht: (2025)
von: Liu, Tong, et al.
Veröffentlicht: (2025)
Distributional Soft Actor-Critic with Three Refinements
von: Duan, Jingliang, et al.
Veröffentlicht: (2023)
von: Duan, Jingliang, et al.
Veröffentlicht: (2023)
Generative Actor-Critic with Soft Bridge Policies
von: He, Ke, et al.
Veröffentlicht: (2026)
von: He, Ke, et al.
Veröffentlicht: (2026)
Adaptive Ensemble Aggregation for Actor-Critics
von: Werge, Nicklas, et al.
Veröffentlicht: (2025)
von: Werge, Nicklas, et al.
Veröffentlicht: (2025)
Scalable Neighborhood-Based Multi-Agent Actor-Critic
von: Goppelsroeder, Tim, et al.
Veröffentlicht: (2026)
von: Goppelsroeder, Tim, et al.
Veröffentlicht: (2026)
Cooperative Multi-Agent Graph Bandits: UCB Algorithm and Regret Analysis
von: Paschalidis, Phevos, et al.
Veröffentlicht: (2024)
von: Paschalidis, Phevos, et al.
Veröffentlicht: (2024)
DSAC: Distributional Soft Actor-Critic for Risk-Sensitive Reinforcement Learning
von: Ma, Xiaoteng, et al.
Veröffentlicht: (2020)
von: Ma, Xiaoteng, et al.
Veröffentlicht: (2020)
SACn: Soft Actor-Critic with n-step Returns
von: Łyskawa, Jakub, et al.
Veröffentlicht: (2025)
von: Łyskawa, Jakub, et al.
Veröffentlicht: (2025)
Chunking the Critic: A Transformer-based Soft Actor-Critic with N-Step Returns
von: Tian, Dong, et al.
Veröffentlicht: (2025)
von: Tian, Dong, et al.
Veröffentlicht: (2025)
Langevin Soft Actor-Critic: Efficient Exploration through Uncertainty-Driven Critic Learning
von: Ishfaq, Haque, et al.
Veröffentlicht: (2025)
von: Ishfaq, Haque, et al.
Veröffentlicht: (2025)
Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic
von: Vo, Thanh Vinh, et al.
Veröffentlicht: (2025)
von: Vo, Thanh Vinh, et al.
Veröffentlicht: (2025)
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic
von: He, Jiamin, et al.
Veröffentlicht: (2026)
von: He, Jiamin, et al.
Veröffentlicht: (2026)
FM3Q: Factorized Multi-Agent MiniMax Q-Learning for Two-Team Zero-Sum Markov Game
von: Hu, Guangzheng, et al.
Veröffentlicht: (2024)
von: Hu, Guangzheng, et al.
Veröffentlicht: (2024)
Generative Actor Critic
von: Qin, Aoyang, et al.
Veröffentlicht: (2025)
von: Qin, Aoyang, et al.
Veröffentlicht: (2025)
Learning Without Time-Based Embodiment Resets in Soft-Actor Critic
von: Farrahi, Homayoon, et al.
Veröffentlicht: (2025)
von: Farrahi, Homayoon, et al.
Veröffentlicht: (2025)
Actor-Critic without Actor
von: Ki, Donghyeon, et al.
Veröffentlicht: (2025)
von: Ki, Donghyeon, et al.
Veröffentlicht: (2025)
RLAE: Reinforcement Learning-Assisted Ensemble for LLMs
von: Fu, Yuqian, et al.
Veröffentlicht: (2025)
von: Fu, Yuqian, et al.
Veröffentlicht: (2025)
Dissecting Discrete Soft Actor-Critic: Limitations and Principled Alternatives
von: Asad, Reza, et al.
Veröffentlicht: (2025)
von: Asad, Reza, et al.
Veröffentlicht: (2025)
Decorrelated Soft Actor-Critic for Efficient Deep Reinforcement Learning
von: Küçükoğlu, Burcu, et al.
Veröffentlicht: (2025)
von: Küçükoğlu, Burcu, et al.
Veröffentlicht: (2025)
Adversarially Trained Weighted Actor-Critic for Safe Offline Reinforcement Learning
von: Wei, Honghao, et al.
Veröffentlicht: (2024)
von: Wei, Honghao, et al.
Veröffentlicht: (2024)
S$^2$AC: Energy-Based Reinforcement Learning with Stein Soft Actor Critic
von: Messaoud, Safa, et al.
Veröffentlicht: (2024)
von: Messaoud, Safa, et al.
Veröffentlicht: (2024)
DSAC-C: Constrained Maximum Entropy for Robust Discrete Soft-Actor Critic
von: Neo, Dexter, et al.
Veröffentlicht: (2023)
von: Neo, Dexter, et al.
Veröffentlicht: (2023)
ACSAC: Adaptive Chunk Size Actor-Critic with Causal Transformer Q-Network
von: Chen, Qian, et al.
Veröffentlicht: (2026)
von: Chen, Qian, et al.
Veröffentlicht: (2026)
Revisiting On-Policy Distillation: Empirical Failure Modes and Simple Fixes
von: Fu, Yuqian, et al.
Veröffentlicht: (2026)
von: Fu, Yuqian, et al.
Veröffentlicht: (2026)
D2 Actor Critic: Diffusion Actor Meets Distributional Critic
von: Zhang, Lunjun, et al.
Veröffentlicht: (2025)
von: Zhang, Lunjun, et al.
Veröffentlicht: (2025)
Actor-Critic or Critic-Actor? A Tale of Two Time Scales
von: Bhatnagar, Shalabh, et al.
Veröffentlicht: (2022)
von: Bhatnagar, Shalabh, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
R2PS: Worst-Case Robust Real-Time Pursuit Strategies under Partial Observability
von: Lu, Runyu, et al.
Veröffentlicht: (2025) -
Equilibrium Policy Generalization: A Reinforcement Learning Framework for Cross-Graph Zero-Shot Generalization in Pursuit-Evasion Games
von: Lu, Runyu, et al.
Veröffentlicht: (2025) -
ISAACS: Iterative Soft Adversarial Actor-Critic for Safety
von: Hsu, Kai-Chieh, et al.
Veröffentlicht: (2022) -
Discretizing Continuous Action Space with Unimodal Probability Distributions for On-Policy Reinforcement Learning
von: Zhu, Yuanyang, et al.
Veröffentlicht: (2024) -
Revisiting Discrete Soft Actor-Critic
von: Zhou, Haibin, et al.
Veröffentlicht: (2022)