Heavy-Ball Momentum Accelerated Actor-Critic With Function Approximation
Fuente:
arXiv
Guardado en:
| Autores principales: | Dong, Yanjie, Zhang, Haijun, Wang, Gang, Cui, Shisheng, Hu, Xiping |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Non-Asymptotic Analysis for Single-Loop (Natural) Actor-Critic with Compatible Function Approximation
por: Wang, Yudan, et al.
Publicado: (2024)
por: Wang, Yudan, et al.
Publicado: (2024)
Communication-Efficient Heterogeneous Federated Learning with Generalized Heavy-Ball Momentum
por: Zaccone, Riccardo, et al.
Publicado: (2023)
por: Zaccone, Riccardo, et al.
Publicado: (2023)
Functional Critics Are Essential for Actor-Critic: From Off-Policy Stability to Efficient Exploration
por: Bai, Qinxun, et al.
Publicado: (2025)
por: Bai, Qinxun, et al.
Publicado: (2025)
Diffusion Actor-Critic with Entropy Regulator
por: Wang, Yinuo, et al.
Publicado: (2024)
por: Wang, Yinuo, et al.
Publicado: (2024)
Provable Acceleration of Nesterov's Accelerated Gradient Method over Heavy Ball Method in Training Over-Parameterized Neural Networks
por: Liu, Xin, et al.
Publicado: (2022)
por: Liu, Xin, et al.
Publicado: (2022)
Finite-time Convergence Analysis of Actor-Critic with Evolving Reward
por: Hu, Rui, et al.
Publicado: (2025)
por: Hu, Rui, et al.
Publicado: (2025)
Distributional Soft Actor-Critic with Diffusion Policy
por: Liu, Tong, et al.
Publicado: (2025)
por: Liu, Tong, et al.
Publicado: (2025)
Value Improved Actor Critic Algorithms
por: Oren, Yaniv, et al.
Publicado: (2024)
por: Oren, Yaniv, et al.
Publicado: (2024)
Revisiting Discrete Soft Actor-Critic
por: Zhou, Haibin, et al.
Publicado: (2022)
por: Zhou, Haibin, et al.
Publicado: (2022)
Average-Reward Soft Actor-Critic
por: Adamczyk, Jacob, et al.
Publicado: (2025)
por: Adamczyk, Jacob, et al.
Publicado: (2025)
Proximal Action Replacement for Behavior Cloning Actor-Critic in Offline Reinforcement Learning
por: Dong, Jinzong, et al.
Publicado: (2026)
por: Dong, Jinzong, et al.
Publicado: (2026)
Flow Actor-Critic for Offline Reinforcement Learning
por: Chae, Jongseong, et al.
Publicado: (2026)
por: Chae, Jongseong, et al.
Publicado: (2026)
Relative Importance Sampling for off-Policy Actor-Critic in Deep Reinforcement Learning
por: Humayoo, Mahammad, et al.
Publicado: (2018)
por: Humayoo, Mahammad, et al.
Publicado: (2018)
Broad Critic Deep Actor Reinforcement Learning for Continuous Control
por: Thalagala, Shiron, et al.
Publicado: (2024)
por: Thalagala, Shiron, et al.
Publicado: (2024)
Relational Object-Centric Actor-Critic
por: Ugadiarov, Leonid, et al.
Publicado: (2023)
por: Ugadiarov, Leonid, et al.
Publicado: (2023)
DR-SAC: Distributionally Robust Soft Actor-Critic for Reinforcement Learning under Uncertainty
por: Cui, Mingxuan, et al.
Publicado: (2025)
por: Cui, Mingxuan, et al.
Publicado: (2025)
Scalable Neighborhood-Based Multi-Agent Actor-Critic
por: Goppelsroeder, Tim, et al.
Publicado: (2026)
por: Goppelsroeder, Tim, et al.
Publicado: (2026)
Revisiting Mixture Policies in Entropy-Regularized Actor-Critic
por: He, Jiamin, et al.
Publicado: (2026)
por: He, Jiamin, et al.
Publicado: (2026)
SACn: Soft Actor-Critic with n-step Returns
por: Łyskawa, Jakub, et al.
Publicado: (2025)
por: Łyskawa, Jakub, et al.
Publicado: (2025)
Actor-Critics Can Achieve Optimal Sample Efficiency
por: Tan, Kevin, et al.
Publicado: (2025)
por: Tan, Kevin, et al.
Publicado: (2025)
Rethinking Soft Actor-Critic in High-Dimensional Action Spaces: The Cost of Ignoring Distribution Shift
por: Chen, Yanjun, et al.
Publicado: (2024)
por: Chen, Yanjun, et al.
Publicado: (2024)
Collaborative Yet Personalized Policy Training: Single-Timescale Federated Actor-Critic
por: Wang, Leo Muxing, et al.
Publicado: (2026)
por: Wang, Leo Muxing, et al.
Publicado: (2026)
Dissecting Discrete Soft Actor-Critic: Limitations and Principled Alternatives
por: Asad, Reza, et al.
Publicado: (2025)
por: Asad, Reza, et al.
Publicado: (2025)
Studying the Interplay Between the Actor and Critic Representations in Reinforcement Learning
por: Garcin, Samuel, et al.
Publicado: (2025)
por: Garcin, Samuel, et al.
Publicado: (2025)
Ask-AC: An Initiative Advisor-in-the-Loop Actor-Critic Framework
por: Liu, Shunyu, et al.
Publicado: (2022)
por: Liu, Shunyu, et al.
Publicado: (2022)
Decorrelated Soft Actor-Critic for Efficient Deep Reinforcement Learning
por: Küçükoğlu, Burcu, et al.
Publicado: (2025)
por: Küçükoğlu, Burcu, et al.
Publicado: (2025)
Stabilizing the Q-Gradient Field for Policy Smoothness in Actor-Critic
por: Lee, Jeong Woon, et al.
Publicado: (2026)
por: Lee, Jeong Woon, et al.
Publicado: (2026)
Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control
por: Chen, Donghe, et al.
Publicado: (2025)
por: Chen, Donghe, et al.
Publicado: (2025)
(Accelerated) Noise-adaptive Stochastic Heavy-Ball Momentum
por: Dang, Anh, et al.
Publicado: (2024)
por: Dang, Anh, et al.
Publicado: (2024)
A Safety Modulator Actor-Critic Method in Model-Free Safe Reinforcement Learning and Application in UAV Hovering
por: Qi, Qihan, et al.
Publicado: (2024)
por: Qi, Qihan, et al.
Publicado: (2024)
Tackling Heavy-Tailed Rewards in Reinforcement Learning with Function Approximation: Minimax Optimal and Instance-Dependent Regret Bounds
por: Huang, Jiayi, et al.
Publicado: (2023)
por: Huang, Jiayi, et al.
Publicado: (2023)
Offline Actor-Critic Reinforcement Learning Scales to Large Models
por: Springenberg, Jost Tobias, et al.
Publicado: (2024)
por: Springenberg, Jost Tobias, et al.
Publicado: (2024)
Maximum Entropy On-Policy Actor-Critic via Entropy Advantage Estimation
por: Choe, Jean Seong Bjorn, et al.
Publicado: (2024)
por: Choe, Jean Seong Bjorn, et al.
Publicado: (2024)
Double Actor-Critic with TD Error-Driven Regularization in Reinforcement Learning
por: Chen, Haohui, et al.
Publicado: (2024)
por: Chen, Haohui, et al.
Publicado: (2024)
SMAC: Score-Matched Actor-Critics for Robust Offline-to-Online Transfer
por: de Lara, Nathan Samuel, et al.
Publicado: (2026)
por: de Lara, Nathan Samuel, et al.
Publicado: (2026)
Limits of Actor-Critic Algorithms for Decision Tree Policies Learning in IBMDPs
por: Kohler, Hector, et al.
Publicado: (2023)
por: Kohler, Hector, et al.
Publicado: (2023)
Enabling Off-Policy Imitation Learning with Deep Actor Critic Stabilization
por: Sen, Sayambhu, et al.
Publicado: (2025)
por: Sen, Sayambhu, et al.
Publicado: (2025)
DSAC: Distributional Soft Actor-Critic for Risk-Sensitive Reinforcement Learning
por: Ma, Xiaoteng, et al.
Publicado: (2020)
por: Ma, Xiaoteng, et al.
Publicado: (2020)
Nesterov-Accelerated Robust Federated Learning Over Byzantine Adversaries
por: Xu, Lihan, et al.
Publicado: (2025)
por: Xu, Lihan, et al.
Publicado: (2025)
Bidirectional Soft Actor-Critic: Leveraging Forward and Reverse KL Divergence for Efficient Reinforcement Learning
por: Zhang, Yixian, et al.
Publicado: (2025)
por: Zhang, Yixian, et al.
Publicado: (2025)
Ejemplares similares
-
Non-Asymptotic Analysis for Single-Loop (Natural) Actor-Critic with Compatible Function Approximation
por: Wang, Yudan, et al.
Publicado: (2024) -
Communication-Efficient Heterogeneous Federated Learning with Generalized Heavy-Ball Momentum
por: Zaccone, Riccardo, et al.
Publicado: (2023) -
Functional Critics Are Essential for Actor-Critic: From Off-Policy Stability to Efficient Exploration
por: Bai, Qinxun, et al.
Publicado: (2025) -
Diffusion Actor-Critic with Entropy Regulator
por: Wang, Yinuo, et al.
Publicado: (2024) -
Provable Acceleration of Nesterov's Accelerated Gradient Method over Heavy Ball Method in Training Over-Parameterized Neural Networks
por: Liu, Xin, et al.
Publicado: (2022)