Distributional Soft Actor-Critic with Three Refinements
Fuente:
arXiv
Salvato in:
| Autori principali: | Duan, Jingliang, Wang, Wenxuan, Xiao, Liming, Gao, Jiaxin, Li, Shengbo Eben, Liu, Chang, Zhang, Ya-Qin, Cheng, Bo, Li, Keqiang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
On the Optimization Landscape of Observer-based Dynamic Linear Quadratic Control
di: Duan, Jingliang, et al.
Pubblicazione: (2026)
di: Duan, Jingliang, et al.
Pubblicazione: (2026)
Zeroth-Order Actor-Critic: An Evolutionary Framework for Sequential Decision Problems
di: Lei, Yuheng, et al.
Pubblicazione: (2022)
di: Lei, Yuheng, et al.
Pubblicazione: (2022)
Distributional Soft Actor-Critic with Diffusion Policy
di: Liu, Tong, et al.
Pubblicazione: (2025)
di: Liu, Tong, et al.
Pubblicazione: (2025)
Distributional Soft Actor-Critic with Harmonic Gradient for Safe and Efficient Autonomous Driving in Multi-lane Scenarios
di: Zhang, Feihong, et al.
Pubblicazione: (2025)
di: Zhang, Feihong, et al.
Pubblicazione: (2025)
Predictive Lagrangian Optimization for Constrained Reinforcement Learning
di: Zhang, Tianqi, et al.
Pubblicazione: (2025)
di: Zhang, Tianqi, et al.
Pubblicazione: (2025)
Algorithm Design and Comparative Test of Natural Gradient Gaussian Approximation Filter
di: Cao, Wenhan, et al.
Pubblicazione: (2025)
di: Cao, Wenhan, et al.
Pubblicazione: (2025)
On the Stability of Datatic Control Systems
di: Yang, Yujie, et al.
Pubblicazione: (2024)
di: Yang, Yujie, et al.
Pubblicazione: (2024)
Natural Gradient Gaussian Approximation Filter with Positive Definiteness Guarantee
di: Zhang, Tianyi, et al.
Pubblicazione: (2026)
di: Zhang, Tianyi, et al.
Pubblicazione: (2026)
Canonical Form of Datatic Description in Control Systems
di: Zhan, Guojian, et al.
Pubblicazione: (2024)
di: Zhan, Guojian, et al.
Pubblicazione: (2024)
On the Equilibrium between Feasible Zone and Uncertain Model in Safe Exploration
di: Yang, Yujie, et al.
Pubblicazione: (2026)
di: Yang, Yujie, et al.
Pubblicazione: (2026)
Diffusion Actor-Critic with Entropy Regulator
di: Wang, Yinuo, et al.
Pubblicazione: (2024)
di: Wang, Yinuo, et al.
Pubblicazione: (2024)
Nonlinear Bayesian Filtering with Natural Gradient Gaussian Approximation
di: Cao, Wenhan, et al.
Pubblicazione: (2024)
di: Cao, Wenhan, et al.
Pubblicazione: (2024)
Controllability Test for Nonlinear Datatic Systems
di: Yang, Yujie, et al.
Pubblicazione: (2024)
di: Yang, Yujie, et al.
Pubblicazione: (2024)
Dynamic Tactile Sensing System and Soft Actor Critic Reinforcement Learning for Inclusion Characterization
di: Bannan, John, et al.
Pubblicazione: (2026)
di: Bannan, John, et al.
Pubblicazione: (2026)
Wasserstein Barycenter Soft Actor-Critic
di: Shahrooei, Zahra, et al.
Pubblicazione: (2025)
di: Shahrooei, Zahra, et al.
Pubblicazione: (2025)
Curriculum-Based Soft Actor-Critic for Multi-Section R2R Tension Control
di: Li, Shihao, et al.
Pubblicazione: (2026)
di: Li, Shihao, et al.
Pubblicazione: (2026)
Feasible Policy Iteration for Safe Reinforcement Learning
di: Yang, Yujie, et al.
Pubblicazione: (2023)
di: Yang, Yujie, et al.
Pubblicazione: (2023)
Design and Experimental Test of Datatic Approximate Optimal Filter in Nonlinear Dynamic Systems
di: He, Weixian, et al.
Pubblicazione: (2025)
di: He, Weixian, et al.
Pubblicazione: (2025)
Scalable Synthesis of Formally Verified Neural Value Function for Hamilton-Jacobi Reachability Analysis
di: Yang, Yujie, et al.
Pubblicazione: (2024)
di: Yang, Yujie, et al.
Pubblicazione: (2024)
The Feasibility Theory of Constrained Reinforcement Learning: A Tutorial Study
di: Yang, Yujie, et al.
Pubblicazione: (2024)
di: Yang, Yujie, et al.
Pubblicazione: (2024)
Discretionary Lane-Change Decision and Control via Parameterized Soft Actor-Critic for Hybrid Action Space
di: Lin, Yuan, et al.
Pubblicazione: (2024)
di: Lin, Yuan, et al.
Pubblicazione: (2024)
Convolutional Bayesian Filtering
di: Cao, Wenhan, et al.
Pubblicazione: (2024)
di: Cao, Wenhan, et al.
Pubblicazione: (2024)
An Explicit Discrete-Time Dynamic Vehicle Model with Assured Numerical Stability
di: Zhan, Guojian, et al.
Pubblicazione: (2024)
di: Zhan, Guojian, et al.
Pubblicazione: (2024)
Risk-Sensitive Soft Actor-Critic for Robust Deep Reinforcement Learning under Distribution Shifts
di: Enders, Tobias, et al.
Pubblicazione: (2024)
di: Enders, Tobias, et al.
Pubblicazione: (2024)
ISAACS: Iterative Soft Adversarial Actor-Critic for Safety
di: Hsu, Kai-Chieh, et al.
Pubblicazione: (2022)
di: Hsu, Kai-Chieh, et al.
Pubblicazione: (2022)
Adaptive Soft Actor-Critic Framework for RIS-Assisted and UAV-Aided Communication
di: Adam, Abuzar B. M., et al.
Pubblicazione: (2024)
di: Adam, Abuzar B. M., et al.
Pubblicazione: (2024)
A Three-Dimensional Pursuit-Evasion Game Based on Fuzzy Actor-Critic Learning Algorithm
di: Hu, Penglin
Pubblicazione: (2025)
di: Hu, Penglin
Pubblicazione: (2025)
Multi-Agent Soft Actor-Critic with Coordinated Loss for Autonomous Mobility-on-Demand Fleet Control
di: Woywood, Zeno, et al.
Pubblicazione: (2024)
di: Woywood, Zeno, et al.
Pubblicazione: (2024)
Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm
di: Qiao, Ting, et al.
Pubblicazione: (2024)
di: Qiao, Ting, et al.
Pubblicazione: (2024)
Physics-informed Actor-Critic for Coordination of Virtual Inertia from Power Distribution Systems
di: Stock, Simon, et al.
Pubblicazione: (2024)
di: Stock, Simon, et al.
Pubblicazione: (2024)
Nash Soft Actor-Critic LEO Satellite Handover Management Algorithm for Flying Vehicles
di: Chen, Jinxuan, et al.
Pubblicazione: (2024)
di: Chen, Jinxuan, et al.
Pubblicazione: (2024)
A Distributed Actor-Critic Algorithm for Fixed-Time Consensus in Nonlinear Multi-Agent Systems
di: Delshad, Aria, et al.
Pubblicazione: (2025)
di: Delshad, Aria, et al.
Pubblicazione: (2025)
Exploiting Convexity of Neural Networks in Dynamic Operating Envelope Optimization for Distributed Energy Resources
di: Li, Hongyi, et al.
Pubblicazione: (2025)
di: Li, Hongyi, et al.
Pubblicazione: (2025)
Refined Motion Compensation with Soft Laser Manipulators using Data-Driven Surrogate Models
di: Yan, Yongjun, et al.
Pubblicazione: (2024)
di: Yan, Yongjun, et al.
Pubblicazione: (2024)
Application of Soft Actor-Critic Algorithms in Optimizing Wastewater Treatment with Time Delays Integration
di: Mohammadi, Esmaeel, et al.
Pubblicazione: (2024)
di: Mohammadi, Esmaeel, et al.
Pubblicazione: (2024)
Multi-Robot Multi-Queue Control via Exhaustive Assignment Actor-Critic Learning
di: Merati, Mohammad, et al.
Pubblicazione: (2026)
di: Merati, Mohammad, et al.
Pubblicazione: (2026)
Quasi-Newton Compatible Actor-Critic for Deterministic Policies
di: Kordabad, Arash Bahari, et al.
Pubblicazione: (2025)
di: Kordabad, Arash Bahari, et al.
Pubblicazione: (2025)
Transformer Actor-Critic for Efficient Freshness-Aware Resource Allocation
di: Ansarifard, Maryam, et al.
Pubblicazione: (2026)
di: Ansarifard, Maryam, et al.
Pubblicazione: (2026)
Preparation and Motion Study of Magnetically Driven Micro Soft Robot Mimicking the Cownose Ray
di: Chang, Jiaqing, et al.
Pubblicazione: (2026)
di: Chang, Jiaqing, et al.
Pubblicazione: (2026)
On the Foundation of Distributionally Robust Reinforcement Learning
di: Wang, Shengbo, et al.
Pubblicazione: (2023)
di: Wang, Shengbo, et al.
Pubblicazione: (2023)
Documenti analoghi
-
On the Optimization Landscape of Observer-based Dynamic Linear Quadratic Control
di: Duan, Jingliang, et al.
Pubblicazione: (2026) -
Zeroth-Order Actor-Critic: An Evolutionary Framework for Sequential Decision Problems
di: Lei, Yuheng, et al.
Pubblicazione: (2022) -
Distributional Soft Actor-Critic with Diffusion Policy
di: Liu, Tong, et al.
Pubblicazione: (2025) -
Distributional Soft Actor-Critic with Harmonic Gradient for Safe and Efficient Autonomous Driving in Multi-lane Scenarios
di: Zhang, Feihong, et al.
Pubblicazione: (2025) -
Predictive Lagrangian Optimization for Constrained Reinforcement Learning
di: Zhang, Tianqi, et al.
Pubblicazione: (2025)