Saved in:
| Main Authors: | Wu, Andy, Lin, Chun-Cheng, Liaw, Rung-Tzuo, Huang, Yuehua, Kuo, Chihjung, Weng, Chia Tong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2510.01083 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Constant in an Ever-Changing World
by: Wu, Andy, et al.
Published: (2025)
by: Wu, Andy, et al.
Published: (2025)
Federated Natural Policy Gradient and Actor Critic Methods for Multi-task Reinforcement Learning
by: Yang, Tong, et al.
Published: (2023)
by: Yang, Tong, et al.
Published: (2023)
Adaptive Ensemble Aggregation for Actor-Critics
by: Werge, Nicklas, et al.
Published: (2025)
by: Werge, Nicklas, et al.
Published: (2025)
Actor-Critic Reinforcement Learning with Phased Actor
by: Wu, Ruofan, et al.
Published: (2024)
by: Wu, Ruofan, et al.
Published: (2024)
Efficient $Q$-Learning and Actor-Critic Methods for Robust Average Reward Reinforcement Learning
by: Xu, Yang, et al.
Published: (2025)
by: Xu, Yang, et al.
Published: (2025)
Cryptocurrency Portfolio Management with Reinforcement Learning: Soft Actor--Critic and Deep Deterministic Policy Gradient Algorithms
by: Paykan, Kamal
Published: (2025)
by: Paykan, Kamal
Published: (2025)
FORLER: Federated Offline Reinforcement Learning with Q-Ensemble and Actor Rectification
by: Qiao, Nan, et al.
Published: (2026)
by: Qiao, Nan, et al.
Published: (2026)
Natural Policy Gradient and Actor Critic Methods for Constrained Multi-Task Reinforcement Learning
by: Zeng, Sihan, et al.
Published: (2024)
by: Zeng, Sihan, et al.
Published: (2024)
Finite-Time Convergence and Sample Complexity of Actor-Critic Multi-Objective Reinforcement Learning
by: Zhou, Tianchen, et al.
Published: (2024)
by: Zhou, Tianchen, et al.
Published: (2024)
Finite-Time Global Optimality Convergence in Deep Neural Actor-Critic Methods for Decentralized Multi-Agent Reinforcement Learning
by: Zhang, Zhiyao, et al.
Published: (2025)
by: Zhang, Zhiyao, et al.
Published: (2025)
On-policy Actor-Critic Reinforcement Learning for Multi-UAV Exploration
by: Farid, Ali Moltajaei, et al.
Published: (2024)
by: Farid, Ali Moltajaei, et al.
Published: (2024)
${\rm E}(3)$-Equivariant Actor-Critic Methods for Cooperative Multi-Agent Reinforcement Learning
by: Chen, Dingyang, et al.
Published: (2023)
by: Chen, Dingyang, et al.
Published: (2023)
Surrogate Ensemble in Expensive Multi-Objective Optimization via Deep Q-Learning
by: Wu, Yuxin, et al.
Published: (2026)
by: Wu, Yuxin, et al.
Published: (2026)
Efficient Exploration in Deep Reinforcement Learning: A Novel Bayesian Actor-Critic Algorithm
by: Rozanov, Nikolai
Published: (2024)
by: Rozanov, Nikolai
Published: (2024)
Quasi-Newton Compatible Actor-Critic for Deterministic Policies
by: Kordabad, Arash Bahari, et al.
Published: (2025)
by: Kordabad, Arash Bahari, et al.
Published: (2025)
Generative Actor Critic
by: Qin, Aoyang, et al.
Published: (2025)
by: Qin, Aoyang, et al.
Published: (2025)
SALE-Based Offline Reinforcement Learning with Ensemble Q-Networks
by: Chun, Zheng
Published: (2025)
by: Chun, Zheng
Published: (2025)
Broad Critic Deep Actor Reinforcement Learning for Continuous Control
by: Thalagala, Shiron, et al.
Published: (2024)
by: Thalagala, Shiron, et al.
Published: (2024)
Decorrelated Soft Actor-Critic for Efficient Deep Reinforcement Learning
by: Küçükoğlu, Burcu, et al.
Published: (2025)
by: Küçükoğlu, Burcu, et al.
Published: (2025)
Relative Importance Sampling for off-Policy Actor-Critic in Deep Reinforcement Learning
by: Humayoo, Mahammad, et al.
Published: (2018)
by: Humayoo, Mahammad, et al.
Published: (2018)
A Single-Loop Deep Actor-Critic Algorithm for Constrained Reinforcement Learning with Provable Convergence
by: Wang, Kexuan, et al.
Published: (2023)
by: Wang, Kexuan, et al.
Published: (2023)
Alleviating Community Fear in Disasters via Multi-Agent Actor-Critic Reinforcement Learning
by: Hakke, Yashodhan D., et al.
Published: (2026)
by: Hakke, Yashodhan D., et al.
Published: (2026)
Multi-agent Off-policy Actor-Critic Reinforcement Learning for Partially Observable Environments
by: Zhaikhan, Ainur, et al.
Published: (2024)
by: Zhaikhan, Ainur, et al.
Published: (2024)
Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control
by: Chen, Donghe, et al.
Published: (2025)
by: Chen, Donghe, et al.
Published: (2025)
Deep Actor-Critics with Tight Risk Certificates
by: Tasdighi, Bahareh, et al.
Published: (2025)
by: Tasdighi, Bahareh, et al.
Published: (2025)
Graphon Mean-Field Subsampling for Cooperative Heterogeneous Multi-Agent Reinforcement Learning
by: Anand, Emile, et al.
Published: (2026)
by: Anand, Emile, et al.
Published: (2026)
A Multi-stage Error Diagnosis for APB Transaction
by: Tsai, Cheng-Yang, et al.
Published: (2025)
by: Tsai, Cheng-Yang, et al.
Published: (2025)
Disentangled Latent Spaces for Reduced Order Models using Deterministic Autoencoders
by: Schwarz, Henning, et al.
Published: (2025)
by: Schwarz, Henning, et al.
Published: (2025)
Multi-Agent Actor-Critics in Autonomous Cyber Defense
by: Wang, Mingjun, et al.
Published: (2024)
by: Wang, Mingjun, et al.
Published: (2024)
Scalable Neighborhood-Based Multi-Agent Actor-Critic
by: Goppelsroeder, Tim, et al.
Published: (2026)
by: Goppelsroeder, Tim, et al.
Published: (2026)
Asymmetric Actor-Critic for Multi-turn LLM Agents
by: Jiang, Shuli, et al.
Published: (2026)
by: Jiang, Shuli, et al.
Published: (2026)
MSSEAC: Multi‐State Soft Elastic Actor‐Critic
by: Yuwan Gu, et al.
Published: (2026)
by: Yuwan Gu, et al.
Published: (2026)
Revisiting Discrete Soft Actor-Critic
by: Zhou, Haibin, et al.
Published: (2022)
by: Zhou, Haibin, et al.
Published: (2022)
Association between ischemic stroke, hemorrhagic stroke, dementia, and rs201118034 among general Taiwanese population
by: Yi‐Chia Liaw, et al.
Published: (2025)
by: Yi‐Chia Liaw, et al.
Published: (2025)
Communicating longevity risk : more things to think about / Liaw Huang, Tom Terry
by: Huang, Liaw
by: Huang, Liaw
Exploration from a Primal-Dual Lens: Value-Incentivized Actor-Critic Methods for Sample-Efficient Online RL
by: Yang, Tong, et al.
Published: (2025)
by: Yang, Tong, et al.
Published: (2025)
Stabilizing the Q-Gradient Field for Policy Smoothness in Actor-Critic
by: Lee, Jeong Woon, et al.
Published: (2026)
by: Lee, Jeong Woon, et al.
Published: (2026)
Letter to the Editor Regarding Determinants of HBeAg Loss During Follow‐Up of a Multiethnic Pediatric Cohort
by: Chia‐Ming Chu, et al.
Published: (2024)
by: Chia‐Ming Chu, et al.
Published: (2024)
Quantum Advantage Actor-Critic for Reinforcement Learning
by: Kölle, Michael, et al.
Published: (2024)
by: Kölle, Michael, et al.
Published: (2024)
Flow Actor-Critic for Offline Reinforcement Learning
by: Chae, Jongseong, et al.
Published: (2026)
by: Chae, Jongseong, et al.
Published: (2026)
Similar Items
-
Constant in an Ever-Changing World
by: Wu, Andy, et al.
Published: (2025) -
Federated Natural Policy Gradient and Actor Critic Methods for Multi-task Reinforcement Learning
by: Yang, Tong, et al.
Published: (2023) -
Adaptive Ensemble Aggregation for Actor-Critics
by: Werge, Nicklas, et al.
Published: (2025) -
Actor-Critic Reinforcement Learning with Phased Actor
by: Wu, Ruofan, et al.
Published: (2024) -
Efficient $Q$-Learning and Actor-Critic Methods for Robust Average Reward Reinforcement Learning
by: Xu, Yang, et al.
Published: (2025)