Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control
Fuente:
arXiv
Guardado en:
| Autores principales: | Chen, Donghe, Peng, Yubin, Zheng, Tengjie, Wang, Han, Qu, Chaoran, Cheng, Lin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Error Distribution Smoothing:Advancing Low-Dimensional Imbalanced Regression
por: Chen, Donghe, et al.
Publicado: (2025)
por: Chen, Donghe, et al.
Publicado: (2025)
Double Actor-Critic with TD Error-Driven Regularization in Reinforcement Learning
por: Chen, Haohui, et al.
Publicado: (2024)
por: Chen, Haohui, et al.
Publicado: (2024)
Flow Actor-Critic for Offline Reinforcement Learning
por: Chae, Jongseong, et al.
Publicado: (2026)
por: Chae, Jongseong, et al.
Publicado: (2026)
Broad Critic Deep Actor Reinforcement Learning for Continuous Control
por: Thalagala, Shiron, et al.
Publicado: (2024)
por: Thalagala, Shiron, et al.
Publicado: (2024)
Relative Importance Sampling for off-Policy Actor-Critic in Deep Reinforcement Learning
por: Humayoo, Mahammad, et al.
Publicado: (2018)
por: Humayoo, Mahammad, et al.
Publicado: (2018)
Monte Carlo Beam Search for Actor-Critic Reinforcement Learning in Continuous Control
por: Alzorgan, Hazim, et al.
Publicado: (2025)
por: Alzorgan, Hazim, et al.
Publicado: (2025)
DSAC: Distributional Soft Actor-Critic for Risk-Sensitive Reinforcement Learning
por: Ma, Xiaoteng, et al.
Publicado: (2020)
por: Ma, Xiaoteng, et al.
Publicado: (2020)
Studying the Interplay Between the Actor and Critic Representations in Reinforcement Learning
por: Garcin, Samuel, et al.
Publicado: (2025)
por: Garcin, Samuel, et al.
Publicado: (2025)
Decorrelated Soft Actor-Critic for Efficient Deep Reinforcement Learning
por: Küçükoğlu, Burcu, et al.
Publicado: (2025)
por: Küçükoğlu, Burcu, et al.
Publicado: (2025)
Quantum Advantage Actor-Critic for Reinforcement Learning
por: Kölle, Michael, et al.
Publicado: (2024)
por: Kölle, Michael, et al.
Publicado: (2024)
Proximal Action Replacement for Behavior Cloning Actor-Critic in Offline Reinforcement Learning
por: Dong, Jinzong, et al.
Publicado: (2026)
por: Dong, Jinzong, et al.
Publicado: (2026)
Reinforcement Learning Position Control of a Quadrotor Using Soft Actor-Critic (SAC)
por: Mahran, Youssef, et al.
Publicado: (2025)
por: Mahran, Youssef, et al.
Publicado: (2025)
Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic
por: Vo, Thanh Vinh, et al.
Publicado: (2025)
por: Vo, Thanh Vinh, et al.
Publicado: (2025)
Federated Natural Policy Gradient and Actor Critic Methods for Multi-task Reinforcement Learning
por: Yang, Tong, et al.
Publicado: (2023)
por: Yang, Tong, et al.
Publicado: (2023)
Offline Actor-Critic Reinforcement Learning Scales to Large Models
por: Springenberg, Jost Tobias, et al.
Publicado: (2024)
por: Springenberg, Jost Tobias, et al.
Publicado: (2024)
Efficient $Q$-Learning and Actor-Critic Methods for Robust Average Reward Reinforcement Learning
por: Xu, Yang, et al.
Publicado: (2025)
por: Xu, Yang, et al.
Publicado: (2025)
Human-Readable Programs as Actors of Reinforcement Learning Agents Using Critic-Moderated Evolution
por: Deproost, Senne, et al.
Publicado: (2024)
por: Deproost, Senne, et al.
Publicado: (2024)
Contraction Actor-Critic: Contraction Metric-Guided Reinforcement Learning for Robust Path Tracking
por: Cho, Minjae, et al.
Publicado: (2025)
por: Cho, Minjae, et al.
Publicado: (2025)
Training Free Guided Flow Matching with Optimal Control
por: Wang, Luran, et al.
Publicado: (2024)
por: Wang, Luran, et al.
Publicado: (2024)
DR-SAC: Distributionally Robust Soft Actor-Critic for Reinforcement Learning under Uncertainty
por: Cui, Mingxuan, et al.
Publicado: (2025)
por: Cui, Mingxuan, et al.
Publicado: (2025)
Bidirectional Soft Actor-Critic: Leveraging Forward and Reverse KL Divergence for Efficient Reinforcement Learning
por: Zhang, Yixian, et al.
Publicado: (2025)
por: Zhang, Yixian, et al.
Publicado: (2025)
Revisiting Discrete Soft Actor-Critic
por: Zhou, Haibin, et al.
Publicado: (2022)
por: Zhou, Haibin, et al.
Publicado: (2022)
${\rm E}(3)$-Equivariant Actor-Critic Methods for Cooperative Multi-Agent Reinforcement Learning
por: Chen, Dingyang, et al.
Publicado: (2023)
por: Chen, Dingyang, et al.
Publicado: (2023)
Your Language Model is Its Own Critic: Reinforcement Learning with Value Estimation from Actor's Internal States
por: Choi, Yunho, et al.
Publicado: (2026)
por: Choi, Yunho, et al.
Publicado: (2026)
Functional Critics Are Essential for Actor-Critic: From Off-Policy Stability to Efficient Exploration
por: Bai, Qinxun, et al.
Publicado: (2025)
por: Bai, Qinxun, et al.
Publicado: (2025)
TOPPO: Rethinking PPO for Multi-Task Reinforcement Learning with Critic Balancing
por: Li, Yuanpeng, et al.
Publicado: (2026)
por: Li, Yuanpeng, et al.
Publicado: (2026)
Comparative Analysis of Parameterized Action Actor-Critic Reinforcement Learning Algorithms for Web Search Match Plan Generation
por: Bapoo, Ubayd, et al.
Publicado: (2025)
por: Bapoo, Ubayd, et al.
Publicado: (2025)
Distributional Soft Actor-Critic with Diffusion Policy
por: Liu, Tong, et al.
Publicado: (2025)
por: Liu, Tong, et al.
Publicado: (2025)
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models
por: Peng, Zengqi, et al.
Publicado: (2025)
por: Peng, Zengqi, et al.
Publicado: (2025)
Value Improved Actor Critic Algorithms
por: Oren, Yaniv, et al.
Publicado: (2024)
por: Oren, Yaniv, et al.
Publicado: (2024)
Diffusion Actor-Critic with Entropy Regulator
por: Wang, Yinuo, et al.
Publicado: (2024)
por: Wang, Yinuo, et al.
Publicado: (2024)
Average-Reward Soft Actor-Critic
por: Adamczyk, Jacob, et al.
Publicado: (2025)
por: Adamczyk, Jacob, et al.
Publicado: (2025)
Stability Enhancement in Reinforcement Learning via Adaptive Control Lyapunov Function
por: Chen, Donghe, et al.
Publicado: (2025)
por: Chen, Donghe, et al.
Publicado: (2025)
Fully Spiking Actor Network with Intra-layer Connections for Reinforcement Learning
por: Chen, Ding, et al.
Publicado: (2024)
por: Chen, Ding, et al.
Publicado: (2024)
Finite-time Convergence Analysis of Actor-Critic with Evolving Reward
por: Hu, Rui, et al.
Publicado: (2025)
por: Hu, Rui, et al.
Publicado: (2025)
Limits of Actor-Critic Algorithms for Decision Tree Policies Learning in IBMDPs
por: Kohler, Hector, et al.
Publicado: (2023)
por: Kohler, Hector, et al.
Publicado: (2023)
Enabling Off-Policy Imitation Learning with Deep Actor Critic Stabilization
por: Sen, Sayambhu, et al.
Publicado: (2025)
por: Sen, Sayambhu, et al.
Publicado: (2025)
Neighboring State-based Exploration for Reinforcement Learning
por: Li, Yu-Teng, et al.
Publicado: (2022)
por: Li, Yu-Teng, et al.
Publicado: (2022)
Offline Reinforcement Learning with OOD State Correction and OOD Action Suppression
por: Mao, Yixiu, et al.
Publicado: (2024)
por: Mao, Yixiu, et al.
Publicado: (2024)
A Safety Modulator Actor-Critic Method in Model-Free Safe Reinforcement Learning and Application in UAV Hovering
por: Qi, Qihan, et al.
Publicado: (2024)
por: Qi, Qihan, et al.
Publicado: (2024)
Ejemplares similares
-
Error Distribution Smoothing:Advancing Low-Dimensional Imbalanced Regression
por: Chen, Donghe, et al.
Publicado: (2025) -
Double Actor-Critic with TD Error-Driven Regularization in Reinforcement Learning
por: Chen, Haohui, et al.
Publicado: (2024) -
Flow Actor-Critic for Offline Reinforcement Learning
por: Chae, Jongseong, et al.
Publicado: (2026) -
Broad Critic Deep Actor Reinforcement Learning for Continuous Control
por: Thalagala, Shiron, et al.
Publicado: (2024) -
Relative Importance Sampling for off-Policy Actor-Critic in Deep Reinforcement Learning
por: Humayoo, Mahammad, et al.
Publicado: (2018)